Package
extract-pdf
Convert a PDF (URL or ArrayBuffer) into clean HTML with structural tagging — headings, lists, footnotes, code blocks, bold/italic. Slim by default (PDF.js loads at runtime from the pdfjs-serverless CDN build), with optional Granite Docling OCR for pages w
Key numbers
As of 25 Sept 2026
- Weekly downloads
- 3.4k
- - over 12 months
- Gzip bundle size
- -
- No Bundlephobia data
- Dependencies
- 1
- Direct, from package.json
- Last release
- 1 day ago
- v0.1.381
- GitHub stars
- 77
- 11 forks
- Open issues / PRs
- 0 / 0
- Contributors
- -
- Commits, last 52 weeks
- 1,179
Downloads
Last week 2026-09-15 to 2026-09-21
3,388 downloads last week, 24,507 in the last month. No 52-week series has been fetched yet.
- Releases, last 12 months
- -
- Latest version
- 0.1.381, 23 Sept 20261 day ago
- First published
- -
- Last push to GitHub
- 25 Sept 2026same day as the snapshot
- Repository archived
- No
- Open issues
- 0
- Open pull requests
- 0
- Commits, last 52 weeks
- 1,179
- Dependents on npm
- 1
Major versions
-
Recent releases
-
- Module format
- ESM + CommonJSFrom package.json type, main and exports fields
- TypeScript types
- Bundled
- Node.js engine range
- Not declared
- Runtimes
- -
- Unpacked size (npm)
- 263 kB
- Dependencies (1)
grab-url
npm install extract-pdf- jsDelivr
https://cdn.jsdelivr.net/npm/extract-pdf@0.1.381/ - unpkg
https://unpkg.com/extract-pdf@0.1.381/ - esm.sh
https://esm.sh/extract-pdf@0.1.381
Indexability: Not indexed (discovery only): downloads_week 3388 < 10000.