Skip to content
JS
Package category

Scraping and browser automation

Headless browsers, crawlers and HTML extraction.

373 packages7 comparisons

Packages compared

373 packages
PackageWeekly downloads12-month change52 weeksGzipLast releaseModuleTypesCategories
domparser-darwin-arm64
A super fast html parser and manipulator written in rust.
4.1k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
@the-convocation/twitter-scraper
A port of n0madic/twitter-scraper to Node.js.
4.1k---5 months ago
0.22.3
ESM + CommonJSBundledScraping and browser automation
w3wallets
browser wallets for playwright
4.1k+86%-3 days ago
1.0.0-beta.13
CommonJSBundledTesting, Blockchain and Web3
pageres
Capture website screenshots
4k+119%--
9.0.0
ESM onlyBundledScraping and browser automation
@takumi-rs/core-win32-arm64-msvc
Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
4k---10 days ago
2.14.0
CommonJSNoneCloud SDKs, Scraping and browser automation
openbrand
Extract brand assets (logos, colors, backdrops) from any website URL
4k--4 months ago
0.2.3
ESM + CommonJSBundledScraping and browser automation
zencf
A client library for accessing the CF Bypass.
3.9k--9 months ago
2.0.3
CommonJSNoneCloud SDKs, Scraping and browser automation
browserless
The headless Chrome/Chromium driver on top of Puppeteer. Take screenshots, generate PDFs, extract text and HTML with a production-ready API.
3.9k+110%-4 days ago
13.12.3
CommonJSBundledScraping and browser automation, DOM and browser utilities
@testplane/wdio-utils
A WDIO helper utility to provide several utility functions used across the project.
3.8k---1 month ago
9.5.5
ESM + CommonJSBundledUtility libraries, Scraping and browser automation
@playwright-opentelemetry/trace-api
H3-based API library for storing and serving Playwright OpenTelemetry traces in S3-compatible storage.
3.8k---1 day ago
0.13.2
ESM onlyBundledMonitoring and error tracking, Scraping and browser automation
metascraper-lang
Metascraper rule to extract the lang from HTML using Open Graph, JSON-LD, and fallback selectors.
3.6k+72%-1 month ago
5.56.2
CommonJSBundledScraping and browser automation
@webreel/core
Core recording engine for webreel - headless Chrome capture, cursor animation, and video compositing.
3.5k---6 months ago
0.1.4
ESM onlyBundledScraping and browser automation, Video and audio
playwright-ghost
Playwright with plugins to be a ghost.
3.5k+813%-3 months ago
0.19.0
ESM onlyBundledScraping and browser automation, Polyfills and shims
domparser-linux-arm64-gnu
A super fast html parser and manipulator written in rust.
3.5k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
@brightdata/cli
Command-line interface for Bright Data. Scrape, search, extract structured data, and automate browsers directly from your terminal.
3.4k---9 days ago
0.3.7
CommonJSBundledCLI tools and terminal utilities, Scraping and browser automation
happy-dom-without-node
Happy DOM is a JavaScript implementation of a web browser without its graphical user interface. It includes many web standards from WHATWG DOM and HTML.
3.4k+437%-2 years ago
14.12.3
ESM onlyBundledScraping and browser automation
axe-playwright-report
Playwright + axe-core integration to run accessibility scans and build HTML dashboard reports.
3.4k+257%-2 months ago
1.2.5
CommonJSBundledAccessibility, Scraping and browser automation
@wreq-js/binding-linux-x64-musl
Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
3.3k---1 month ago
3.2.0
CommonJSNoneHTTP clients, Scraping and browser automation
domparser-linux-arm64-musl
A super fast html parser and manipulator written in rust.
3.3k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
playwright-s3-reporter
A Playwright Reporter for uploading traces to S3 compatible services.
3.2k+458%-28 days ago
1.5.0
ESM + CommonJSBundledScraping and browser automation, Testing
gulp-dom
Gulp plugin for generic DOM manipulation
3.2k+9%-7 years ago
1.0.0
CommonJSNoneScraping and browser automation
ddg-bulk-image-downloader
Lazy way to download images from Duck Duck Go search results in bulk
3.2k-85%-4 years ago
0.1.11
CommonJSBundledScraping and browser automation
qawolf-socket-npm
QA Wolf automation testing package with real-world dependencies
3.1k+184%-1 day ago
1.0.648
CommonJSNoneTesting, WebSockets and realtime
metascraper-x
Metascraper rules for X (Twitter) posts — author, text, media, and engagement metadata.
3.1k+261%-7 days ago
5.58.1
CommonJSBundledScraping and browser automation
garmin-connect
Makes it simple to interface with Garmin Connect to get or set any data point
3k+587%-2 years ago
1.6.2
CommonJSBundledScraping and browser automation
playwright-ajv-schema-validator
A Playwright plugin for API schema validation against plain JSON schemas, Swagger schema documents. Built on the robust core-ajv-schema-validator plugin and powered by the Ajv JSON Schema Validator, it delivers results in a clear, user-friendly format, si
3k-5%-1 year ago
1.0.2
CommonJSBundledScraping and browser automation, Testing
images-scraper
Simple scraper for Google images using Puppeteer
2.9k+460%-1 year ago
7.0.0
CommonJSNoneScraping and browser automation
n8n-nodes-puppeteer
n8n node for browser automation using Puppeteer
2.9k-68%-8 months ago
1.5.0
CommonJSNoneScraping and browser automation, PDF and documents
rebrowser-patches
Collection of patches for puppeteer and playwright to avoid automation detection and leaks. Helps to avoid Cloudflare and DataDome CAPTCHA pages. Easy to patch/unpatch, can be enabled/disabled on demand.
2.9k+357%-1 year ago
1.0.19
ESM onlyNoneScraping and browser automation, Cloud SDKs
nodejs-web-scraper
A web scraper for NodeJs
2.9k-54%-3 years ago
6.1.3
CommonJSNoneScraping and browser automation
jsonld-extract
A damn simple tool to extract json-ld metadata from webpage using jquery like api (jQuery, Cheerio, CashDOM, ...).
2.9k+1924%-5 years ago
0.0.8
CommonJSNoneScraping and browser automation, Parsers and serialisers
domparser-darwin-x64
A super fast html parser and manipulator written in rust.
2.9k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
domparser-win32-arm64-msvc
A super fast html parser and manipulator written in rust.
2.8k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
google
A module to search and scrape google. This is not sponsored, supported, or affiliated with Google Inc.
2.8k-50%-10 years ago
2.1.0
CommonJSNoneScraping and browser automation
@electrovir/rebrowser-playwright
A drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests.
2.7k---2 months ago
1.61.101
ESM + CommonJSBundledScraping and browser automation
@clipboard-health/playwright-reporter-llm
Playwright reporter that outputs structured JSON for LLM agents. Minimal console output, flat schema, easy to filter to failures.
2.7k---2 days ago
2.11.7
CommonJSBundledTesting, Scraping and browser automation
@browserless/pdf
Convert websites to high-quality PDFs with customizable margins, background printing, and optimized scaling.
2.6k---4 days ago
13.12.3
CommonJSNoneScraping and browser automation, PDF and documents
@trybyte/robotstxt-parser
Compile robots.txt rules for Google-compatible or strict RFC 9309 matching
2.6k---22 days ago
2.0.0
ESM onlyBundledParsers and serialisers, Scraping and browser automation
@centralinc/browseragent
Browser automation agent using Computer Use with Playwright
2.6k---8 months ago
1.9.6
ESM + CommonJSBundledScraping and browser automation
puppeteer-html-pdf
HTML to PDF converter for Node.js
2.6k+1%-2 years ago
4.0.8
CommonJSBundledScraping and browser automation, PDF and documents
puppeteer-afp
Coherent, persistable anti-fingerprinting for Puppeteer. One seed → one consistent browser identity, with proxy/timezone/geo coherence and a fingerprint vault.
2.5k-35%-3 months ago
3.0.1
CommonJSBundledScraping and browser automation, 3D, WebGL and game engines
cuimp
Node wrapper for curl-impersonate (lexiforest) via CLI - Enhanced with raw buffer support and extra curl args
2.4k+12412%-1 month ago
2.1.1
ESM + CommonJSBundledScraping and browser automation, TypeScript tooling
qwenproxy-cli
High-performance OpenAI & Anthropic compatible API gateway for Qwen with multi-account rotation, interactive TUI, and resilient tool calling.
2.4k--today
1.3.6
ESM onlyNoneCLI tools and terminal utilities, Testing
reverbnation-scraper
Simple package to download audio & fetch basic details of the song from reverbnation.
2.4k+73%-6 years ago
2.0.0
CommonJSNoneScraping and browser automation
beautiful-dom
Beautiful-dom is a lightweight library that mirrors the capabilities of the HTML DOM API needed for parsing crawled HTML/XML pages. It models the methods and properties of HTML nodes that are relevant for extracting data from HTML nodes. It is written in
2.4k-10%-5 years ago
1.0.9
CommonJSBundledParsers and serialisers, Scraping and browser automation
pageres-cli
Capture website screenshots
2.4k+85%--
9.0.0
ESM onlyNoneCLI tools and terminal utilities, Scraping and browser automation
ivya
Fork of Playwright's locator resolution
2.3k+234%-4 months ago
1.8.2
ESM onlyBundledScraping and browser automation, Testing
@fetcher-sh/api
The developer-friendly client for fetcher.sh — 111 pay-per-call web-data endpoints across Twitter/X, YouTube, TikTok, Instagram, Reddit, Google, App Store, and Yelp. Use as an NPM library or CLI.
2.3k---1 month ago
1.0.1
ESM + CommonJSBundledMaps and geolocation, Scraping and browser automation
slimdom-sax-parser
Parse an XML string to a light-weight spec-compliant document object model, for browser and Node
2.3k+22%-4 years ago
1.5.3
ESM + CommonJSBundledParsers and serialisers, Scraping and browser automation
x-ray
structure any website
2.3k+23%-7 years ago
2.3.4
CommonJSNoneScraping and browser automation

12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.

  • domparser-darwin-arm64A super fast html parser and manipulator written in rust.
  • @the-convocation/twitter-scraperA port of n0madic/twitter-scraper to Node.js.
  • w3walletsbrowser wallets for playwright
  • pageresCapture website screenshots
  • @takumi-rs/core-win32-arm64-msvcRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
  • openbrandExtract brand assets (logos, colors, backdrops) from any website URL
  • zencfA client library for accessing the CF Bypass.
  • browserlessThe headless Chrome/Chromium driver on top of Puppeteer. Take screenshots, generate PDFs, extract text and HTML with a production-ready API.
  • @testplane/wdio-utilsA WDIO helper utility to provide several utility functions used across the project.
  • @playwright-opentelemetry/trace-apiH3-based API library for storing and serving Playwright OpenTelemetry traces in S3-compatible storage.
  • metascraper-langMetascraper rule to extract the lang from HTML using Open Graph, JSON-LD, and fallback selectors.
  • @webreel/coreCore recording engine for webreel - headless Chrome capture, cursor animation, and video compositing.
  • playwright-ghostPlaywright with plugins to be a ghost.
  • domparser-linux-arm64-gnuA super fast html parser and manipulator written in rust.
  • @brightdata/cliCommand-line interface for Bright Data. Scrape, search, extract structured data, and automate browsers directly from your terminal.
  • happy-dom-without-nodeHappy DOM is a JavaScript implementation of a web browser without its graphical user interface. It includes many web standards from WHATWG DOM and HTML.
  • axe-playwright-reportPlaywright + axe-core integration to run accessibility scans and build HTML dashboard reports.
  • @wreq-js/binding-linux-x64-muslNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
  • domparser-linux-arm64-muslA super fast html parser and manipulator written in rust.
  • playwright-s3-reporterA Playwright Reporter for uploading traces to S3 compatible services.
  • gulp-domGulp plugin for generic DOM manipulation
  • ddg-bulk-image-downloaderLazy way to download images from Duck Duck Go search results in bulk
  • qawolf-socket-npmQA Wolf automation testing package with real-world dependencies
  • metascraper-xMetascraper rules for X (Twitter) posts — author, text, media, and engagement metadata.
  • garmin-connectMakes it simple to interface with Garmin Connect to get or set any data point
  • playwright-ajv-schema-validatorA Playwright plugin for API schema validation against plain JSON schemas, Swagger schema documents. Built on the robust core-ajv-schema-validator plugin and powered by the Ajv JSON Schema Validator, it delivers results in a clear, user-friendly format, si
  • images-scraperSimple scraper for Google images using Puppeteer
  • n8n-nodes-puppeteern8n node for browser automation using Puppeteer
  • rebrowser-patchesCollection of patches for puppeteer and playwright to avoid automation detection and leaks. Helps to avoid Cloudflare and DataDome CAPTCHA pages. Easy to patch/unpatch, can be enabled/disabled on demand.
  • nodejs-web-scraperA web scraper for NodeJs
  • jsonld-extractA damn simple tool to extract json-ld metadata from webpage using jquery like api (jQuery, Cheerio, CashDOM, ...).
  • domparser-darwin-x64A super fast html parser and manipulator written in rust.
  • domparser-win32-arm64-msvcA super fast html parser and manipulator written in rust.
  • googleA module to search and scrape google. This is not sponsored, supported, or affiliated with Google Inc.
  • @electrovir/rebrowser-playwrightA drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests.
  • @clipboard-health/playwright-reporter-llmPlaywright reporter that outputs structured JSON for LLM agents. Minimal console output, flat schema, easy to filter to failures.
  • @browserless/pdfConvert websites to high-quality PDFs with customizable margins, background printing, and optimized scaling.
  • @trybyte/robotstxt-parserCompile robots.txt rules for Google-compatible or strict RFC 9309 matching
  • @centralinc/browseragentBrowser automation agent using Computer Use with Playwright
  • puppeteer-html-pdfHTML to PDF converter for Node.js
  • puppeteer-afpCoherent, persistable anti-fingerprinting for Puppeteer. One seed → one consistent browser identity, with proxy/timezone/geo coherence and a fingerprint vault.
  • cuimpNode wrapper for curl-impersonate (lexiforest) via CLI - Enhanced with raw buffer support and extra curl args
  • qwenproxy-cliHigh-performance OpenAI & Anthropic compatible API gateway for Qwen with multi-account rotation, interactive TUI, and resilient tool calling.
  • reverbnation-scraperSimple package to download audio & fetch basic details of the song from reverbnation.
  • beautiful-domBeautiful-dom is a lightweight library that mirrors the capabilities of the HTML DOM API needed for parsing crawled HTML/XML pages. It models the methods and properties of HTML nodes that are relevant for extracting data from HTML nodes. It is written in
  • pageres-cliCapture website screenshots
  • ivyaFork of Playwright's locator resolution
  • @fetcher-sh/apiThe developer-friendly client for fetcher.sh — 111 pay-per-call web-data endpoints across Twitter/X, YouTube, TikTok, Instagram, Reddit, Google, App Store, and Yelp. Use as an NPM library or CLI.
  • slimdom-sax-parserParse an XML string to a light-weight spec-compliant document object model, for browser and Node
  • x-raystructure any website