Skip to content
JS
Package

extract-pdf

Convert a PDF (URL or ArrayBuffer) into clean HTML with structural tagging — headings, lists, footnotes, code blocks, bold/italic. Slim by default (PDF.js loads at runtime from the pdfjs-serverless CDN build), with optional Granite Docling OCR for pages w

v0.1.381, 23 Sept 2026rights.institute/PROSPERESM + CommonJSTypes: BundledPDF and documents
$ npm install extract-pdf

Key numbers

As of 25 Sept 2026
Weekly downloads
3.4k
- over 12 months
Gzip bundle size
-
No Bundlephobia data
Dependencies
1
Direct, from package.json
Last release
1 day ago
v0.1.381
GitHub stars
77
11 forks
Open issues / PRs
0 / 0
Contributors
-
Commits, last 52 weeks
1,179

Downloads

Last week 2026-09-15 to 2026-09-21
05k10k15k20kSept 2025Mar 2026Sept 2026High 19.9k (8 Sept 2026)Low 0 (23 Sept 2025)
Weekly npm downloads, one point per week, weeks starting 23 Sept 2025 to 15 Sept 2026. The y axis starts at zero.

3,388 downloads last week, 24,507 in the last month. No 52-week series has been fetched yet.

Releases, last 12 months
-
Latest version
0.1.381, 23 Sept 20261 day ago
First published
-
Last push to GitHub
25 Sept 2026same day as the snapshot
Repository archived
No
Open issues
0
Open pull requests
0
Commits, last 52 weeks
1,179
Dependents on npm
1

Major versions

-

Recent releases

-

Module format
ESM + CommonJSFrom package.json type, main and exports fields
TypeScript types
Bundled
Node.js engine range
Not declared
Runtimes
-
Unpacked size (npm)
263 kB
Dependencies (1)
grab-url
npm install extract-pdf
  • jsDelivrhttps://cdn.jsdelivr.net/npm/extract-pdf@0.1.381/
  • unpkghttps://unpkg.com/extract-pdf@0.1.381/
  • esm.shhttps://esm.sh/extract-pdf@0.1.381
Sources: npm registry, npm downloads API, GitHub API. Data fetched 25 Sept 2026.

Indexability: Not indexed (discovery only): downloads_week 3388 < 10000.