Package category
Scraping and browser automation
Headless browsers, crawlers and HTML extraction.
373 packages7 comparisons
Packages compared
373 packages
| Package | Weekly downloads | 12-month change | 52 weeks | Gzip | Last release | Module | Types | Categories |
|---|---|---|---|---|---|---|---|---|
| domparser-darwin-arm64 A super fast html parser and manipulator written in rust. | 4.1k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| @the-convocation/twitter-scraper A port of n0madic/twitter-scraper to Node.js. | 4.1k | - | - | - | 5 months ago 0.22.3 | ESM + CommonJS | Bundled | Scraping and browser automation |
| w3wallets browser wallets for playwright | 4.1k | +86% | - | 3 days ago 1.0.0-beta.13 | CommonJS | Bundled | Testing, Blockchain and Web3 | |
| pageres Capture website screenshots | 4k | +119% | - | - 9.0.0 | ESM only | Bundled | Scraping and browser automation | |
| @takumi-rs/core-win32-arm64-msvc Render OG images from Takumi node trees with native Node.js bindings. No headless browser. | 4k | - | - | - | 10 days ago 2.14.0 | CommonJS | None | Cloud SDKs, Scraping and browser automation |
| openbrand Extract brand assets (logos, colors, backdrops) from any website URL | 4k | - | - | 4 months ago 0.2.3 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| zencf A client library for accessing the CF Bypass. | 3.9k | - | - | 9 months ago 2.0.3 | CommonJS | None | Cloud SDKs, Scraping and browser automation | |
| browserless The headless Chrome/Chromium driver on top of Puppeteer. Take screenshots, generate PDFs, extract text and HTML with a production-ready API. | 3.9k | +110% | - | 4 days ago 13.12.3 | CommonJS | Bundled | Scraping and browser automation, DOM and browser utilities | |
| @testplane/wdio-utils A WDIO helper utility to provide several utility functions used across the project. | 3.8k | - | - | - | 1 month ago 9.5.5 | ESM + CommonJS | Bundled | Utility libraries, Scraping and browser automation |
| @playwright-opentelemetry/trace-api H3-based API library for storing and serving Playwright OpenTelemetry traces in S3-compatible storage. | 3.8k | - | - | - | 1 day ago 0.13.2 | ESM only | Bundled | Monitoring and error tracking, Scraping and browser automation |
| metascraper-lang Metascraper rule to extract the lang from HTML using Open Graph, JSON-LD, and fallback selectors. | 3.6k | +72% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Scraping and browser automation | |
| @webreel/core Core recording engine for webreel - headless Chrome capture, cursor animation, and video compositing. | 3.5k | - | - | - | 6 months ago 0.1.4 | ESM only | Bundled | Scraping and browser automation, Video and audio |
| playwright-ghost Playwright with plugins to be a ghost. | 3.5k | +813% | - | 3 months ago 0.19.0 | ESM only | Bundled | Scraping and browser automation, Polyfills and shims | |
| domparser-linux-arm64-gnu A super fast html parser and manipulator written in rust. | 3.5k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| @brightdata/cli Command-line interface for Bright Data. Scrape, search, extract structured data, and automate browsers directly from your terminal. | 3.4k | - | - | - | 9 days ago 0.3.7 | CommonJS | Bundled | CLI tools and terminal utilities, Scraping and browser automation |
| happy-dom-without-node Happy DOM is a JavaScript implementation of a web browser without its graphical user interface. It includes many web standards from WHATWG DOM and HTML. | 3.4k | +437% | - | 2 years ago 14.12.3 | ESM only | Bundled | Scraping and browser automation | |
| axe-playwright-report Playwright + axe-core integration to run accessibility scans and build HTML dashboard reports. | 3.4k | +257% | - | 2 months ago 1.2.5 | CommonJS | Bundled | Accessibility, Scraping and browser automation | |
| @wreq-js/binding-linux-x64-musl Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible. | 3.3k | - | - | - | 1 month ago 3.2.0 | CommonJS | None | HTTP clients, Scraping and browser automation |
| domparser-linux-arm64-musl A super fast html parser and manipulator written in rust. | 3.3k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| playwright-s3-reporter A Playwright Reporter for uploading traces to S3 compatible services. | 3.2k | +458% | - | 28 days ago 1.5.0 | ESM + CommonJS | Bundled | Scraping and browser automation, Testing | |
| gulp-dom Gulp plugin for generic DOM manipulation | 3.2k | +9% | - | 7 years ago 1.0.0 | CommonJS | None | Scraping and browser automation | |
| ddg-bulk-image-downloader Lazy way to download images from Duck Duck Go search results in bulk | 3.2k | -85% | - | 4 years ago 0.1.11 | CommonJS | Bundled | Scraping and browser automation | |
| qawolf-socket-npm QA Wolf automation testing package with real-world dependencies | 3.1k | +184% | - | 1 day ago 1.0.648 | CommonJS | None | Testing, WebSockets and realtime | |
| metascraper-x Metascraper rules for X (Twitter) posts — author, text, media, and engagement metadata. | 3.1k | +261% | - | 7 days ago 5.58.1 | CommonJS | Bundled | Scraping and browser automation | |
| garmin-connect Makes it simple to interface with Garmin Connect to get or set any data point | 3k | +587% | - | 2 years ago 1.6.2 | CommonJS | Bundled | Scraping and browser automation | |
| playwright-ajv-schema-validator A Playwright plugin for API schema validation against plain JSON schemas, Swagger schema documents. Built on the robust core-ajv-schema-validator plugin and powered by the Ajv JSON Schema Validator, it delivers results in a clear, user-friendly format, si | 3k | -5% | - | 1 year ago 1.0.2 | CommonJS | Bundled | Scraping and browser automation, Testing | |
| images-scraper Simple scraper for Google images using Puppeteer | 2.9k | +460% | - | 1 year ago 7.0.0 | CommonJS | None | Scraping and browser automation | |
| n8n-nodes-puppeteer n8n node for browser automation using Puppeteer | 2.9k | -68% | - | 8 months ago 1.5.0 | CommonJS | None | Scraping and browser automation, PDF and documents | |
| rebrowser-patches Collection of patches for puppeteer and playwright to avoid automation detection and leaks. Helps to avoid Cloudflare and DataDome CAPTCHA pages. Easy to patch/unpatch, can be enabled/disabled on demand. | 2.9k | +357% | - | 1 year ago 1.0.19 | ESM only | None | Scraping and browser automation, Cloud SDKs | |
| nodejs-web-scraper A web scraper for NodeJs | 2.9k | -54% | - | 3 years ago 6.1.3 | CommonJS | None | Scraping and browser automation | |
| jsonld-extract A damn simple tool to extract json-ld metadata from webpage using jquery like api (jQuery, Cheerio, CashDOM, ...). | 2.9k | +1924% | - | 5 years ago 0.0.8 | CommonJS | None | Scraping and browser automation, Parsers and serialisers | |
| domparser-darwin-x64 A super fast html parser and manipulator written in rust. | 2.9k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| domparser-win32-arm64-msvc A super fast html parser and manipulator written in rust. | 2.8k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| google A module to search and scrape google. This is not sponsored, supported, or affiliated with Google Inc. | 2.8k | -50% | - | 10 years ago 2.1.0 | CommonJS | None | Scraping and browser automation | |
| @electrovir/rebrowser-playwright A drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests. | 2.7k | - | - | - | 2 months ago 1.61.101 | ESM + CommonJS | Bundled | Scraping and browser automation |
| @clipboard-health/playwright-reporter-llm Playwright reporter that outputs structured JSON for LLM agents. Minimal console output, flat schema, easy to filter to failures. | 2.7k | - | - | - | 2 days ago 2.11.7 | CommonJS | Bundled | Testing, Scraping and browser automation |
| @browserless/pdf Convert websites to high-quality PDFs with customizable margins, background printing, and optimized scaling. | 2.6k | - | - | - | 4 days ago 13.12.3 | CommonJS | None | Scraping and browser automation, PDF and documents |
| @trybyte/robotstxt-parser Compile robots.txt rules for Google-compatible or strict RFC 9309 matching | 2.6k | - | - | - | 22 days ago 2.0.0 | ESM only | Bundled | Parsers and serialisers, Scraping and browser automation |
| @centralinc/browseragent Browser automation agent using Computer Use with Playwright | 2.6k | - | - | - | 8 months ago 1.9.6 | ESM + CommonJS | Bundled | Scraping and browser automation |
| puppeteer-html-pdf HTML to PDF converter for Node.js | 2.6k | +1% | - | 2 years ago 4.0.8 | CommonJS | Bundled | Scraping and browser automation, PDF and documents | |
| puppeteer-afp Coherent, persistable anti-fingerprinting for Puppeteer. One seed → one consistent browser identity, with proxy/timezone/geo coherence and a fingerprint vault. | 2.5k | -35% | - | 3 months ago 3.0.1 | CommonJS | Bundled | Scraping and browser automation, 3D, WebGL and game engines | |
| cuimp Node wrapper for curl-impersonate (lexiforest) via CLI - Enhanced with raw buffer support and extra curl args | 2.4k | +12412% | - | 1 month ago 2.1.1 | ESM + CommonJS | Bundled | Scraping and browser automation, TypeScript tooling | |
| qwenproxy-cli High-performance OpenAI & Anthropic compatible API gateway for Qwen with multi-account rotation, interactive TUI, and resilient tool calling. | 2.4k | - | - | today 1.3.6 | ESM only | None | CLI tools and terminal utilities, Testing | |
| reverbnation-scraper Simple package to download audio & fetch basic details of the song from reverbnation. | 2.4k | +73% | - | 6 years ago 2.0.0 | CommonJS | None | Scraping and browser automation | |
| beautiful-dom Beautiful-dom is a lightweight library that mirrors the capabilities of the HTML DOM API needed for parsing crawled HTML/XML pages. It models the methods and properties of HTML nodes that are relevant for extracting data from HTML nodes. It is written in | 2.4k | -10% | - | 5 years ago 1.0.9 | CommonJS | Bundled | Parsers and serialisers, Scraping and browser automation | |
| pageres-cli Capture website screenshots | 2.4k | +85% | - | - 9.0.0 | ESM only | None | CLI tools and terminal utilities, Scraping and browser automation | |
| ivya Fork of Playwright's locator resolution | 2.3k | +234% | - | 4 months ago 1.8.2 | ESM only | Bundled | Scraping and browser automation, Testing | |
| @fetcher-sh/api The developer-friendly client for fetcher.sh — 111 pay-per-call web-data endpoints across Twitter/X, YouTube, TikTok, Instagram, Reddit, Google, App Store, and Yelp. Use as an NPM library or CLI. | 2.3k | - | - | - | 1 month ago 1.0.1 | ESM + CommonJS | Bundled | Maps and geolocation, Scraping and browser automation |
| slimdom-sax-parser Parse an XML string to a light-weight spec-compliant document object model, for browser and Node | 2.3k | +22% | - | 4 years ago 1.5.3 | ESM + CommonJS | Bundled | Parsers and serialisers, Scraping and browser automation | |
| x-ray structure any website | 2.3k | +23% | - | 7 years ago 2.3.4 | CommonJS | None | Scraping and browser automation |
12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.
- domparser-darwin-arm64A super fast html parser and manipulator written in rust.
- @the-convocation/twitter-scraperA port of n0madic/twitter-scraper to Node.js.
- w3walletsbrowser wallets for playwright
- pageresCapture website screenshots
- @takumi-rs/core-win32-arm64-msvcRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
- openbrandExtract brand assets (logos, colors, backdrops) from any website URL
- zencfA client library for accessing the CF Bypass.
- browserlessThe headless Chrome/Chromium driver on top of Puppeteer. Take screenshots, generate PDFs, extract text and HTML with a production-ready API.
- @testplane/wdio-utilsA WDIO helper utility to provide several utility functions used across the project.
- @playwright-opentelemetry/trace-apiH3-based API library for storing and serving Playwright OpenTelemetry traces in S3-compatible storage.
- metascraper-langMetascraper rule to extract the lang from HTML using Open Graph, JSON-LD, and fallback selectors.
- @webreel/coreCore recording engine for webreel - headless Chrome capture, cursor animation, and video compositing.
- playwright-ghostPlaywright with plugins to be a ghost.
- domparser-linux-arm64-gnuA super fast html parser and manipulator written in rust.
- @brightdata/cliCommand-line interface for Bright Data. Scrape, search, extract structured data, and automate browsers directly from your terminal.
- happy-dom-without-nodeHappy DOM is a JavaScript implementation of a web browser without its graphical user interface. It includes many web standards from WHATWG DOM and HTML.
- axe-playwright-reportPlaywright + axe-core integration to run accessibility scans and build HTML dashboard reports.
- @wreq-js/binding-linux-x64-muslNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
- domparser-linux-arm64-muslA super fast html parser and manipulator written in rust.
- playwright-s3-reporterA Playwright Reporter for uploading traces to S3 compatible services.
- gulp-domGulp plugin for generic DOM manipulation
- ddg-bulk-image-downloaderLazy way to download images from Duck Duck Go search results in bulk
- qawolf-socket-npmQA Wolf automation testing package with real-world dependencies
- metascraper-xMetascraper rules for X (Twitter) posts — author, text, media, and engagement metadata.
- garmin-connectMakes it simple to interface with Garmin Connect to get or set any data point
- playwright-ajv-schema-validatorA Playwright plugin for API schema validation against plain JSON schemas, Swagger schema documents. Built on the robust core-ajv-schema-validator plugin and powered by the Ajv JSON Schema Validator, it delivers results in a clear, user-friendly format, si
- images-scraperSimple scraper for Google images using Puppeteer
- n8n-nodes-puppeteern8n node for browser automation using Puppeteer
- rebrowser-patchesCollection of patches for puppeteer and playwright to avoid automation detection and leaks. Helps to avoid Cloudflare and DataDome CAPTCHA pages. Easy to patch/unpatch, can be enabled/disabled on demand.
- nodejs-web-scraperA web scraper for NodeJs
- jsonld-extractA damn simple tool to extract json-ld metadata from webpage using jquery like api (jQuery, Cheerio, CashDOM, ...).
- domparser-darwin-x64A super fast html parser and manipulator written in rust.
- domparser-win32-arm64-msvcA super fast html parser and manipulator written in rust.
- googleA module to search and scrape google. This is not sponsored, supported, or affiliated with Google Inc.
- @electrovir/rebrowser-playwrightA drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests.
- @clipboard-health/playwright-reporter-llmPlaywright reporter that outputs structured JSON for LLM agents. Minimal console output, flat schema, easy to filter to failures.
- @browserless/pdfConvert websites to high-quality PDFs with customizable margins, background printing, and optimized scaling.
- @trybyte/robotstxt-parserCompile robots.txt rules for Google-compatible or strict RFC 9309 matching
- @centralinc/browseragentBrowser automation agent using Computer Use with Playwright
- puppeteer-html-pdfHTML to PDF converter for Node.js
- puppeteer-afpCoherent, persistable anti-fingerprinting for Puppeteer. One seed → one consistent browser identity, with proxy/timezone/geo coherence and a fingerprint vault.
- cuimpNode wrapper for curl-impersonate (lexiforest) via CLI - Enhanced with raw buffer support and extra curl args
- qwenproxy-cliHigh-performance OpenAI & Anthropic compatible API gateway for Qwen with multi-account rotation, interactive TUI, and resilient tool calling.
- reverbnation-scraperSimple package to download audio & fetch basic details of the song from reverbnation.
- beautiful-domBeautiful-dom is a lightweight library that mirrors the capabilities of the HTML DOM API needed for parsing crawled HTML/XML pages. It models the methods and properties of HTML nodes that are relevant for extracting data from HTML nodes. It is written in
- pageres-cliCapture website screenshots
- ivyaFork of Playwright's locator resolution
- @fetcher-sh/apiThe developer-friendly client for fetcher.sh — 111 pay-per-call web-data endpoints across Twitter/X, YouTube, TikTok, Instagram, Reddit, Google, App Store, and Yelp. Use as an NPM library or CLI.
- slimdom-sax-parserParse an XML string to a light-weight spec-compliant document object model, for browser and Node
- x-raystructure any website