Skip to content
JS
Package category

Scraping and browser automation

Headless browsers, crawlers and HTML extraction.

373 packages7 comparisons

Packages compared

373 packages
PackageWeekly downloads12-month change52 weeksGzipLast releaseModuleTypesCategories
@appium/execute-driver-plugin
Plugin for batching and executing Appium driver commands
25.7k---today
7.0.1
ESM onlyBundledScraping and browser automation, Testing
metascraper-author
Metascraper rule to extract the author from HTML using Open Graph, JSON-LD, and fallback selectors.
25.3k+258%-1 month ago
5.56.2
CommonJSBundledScraping and browser automation
desired-capabilities
utilities for parsing shorthand Selenium capabilities objects
25.2k+11%-11 years ago
0.1.0
CommonJSNoneScraping and browser automation, Testing
@redhat-developer/page-objects
Page Object API implementation for a VS Code editor used by ExTester framework.
24.6k---11 days ago
1.24.0
CommonJSBundledTesting, Scraping and browser automation
unfurl.js
Scraper for oEmbed, Twitter Cards and Open Graph metadata - fast and Promise-based
24.2k+88%-2 years ago
6.4.0
CommonJSBundledScraping and browser automation
app-store-scraper
scrape data from the itunes app store
24.2k+446%-2 years ago
0.18.0
CommonJSNoneScraping and browser automation
penthouse
Generate critical path CSS for web pages
24.1k-29%-4 years ago
2.3.3
CommonJSNoneCSS-in-JS and styling, Scraping and browser automation
@wreq-js/binding-win32-x64-msvc
Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
23.9k---1 month ago
3.2.0
CommonJSNoneHTTP clients, Scraping and browser automation
happo
Catch unexpected visual and accessibility changes and UI bugs
23.3k+31445%-today
6.20.0
ESM onlyBundledTesting, Scraping and browser automation
@roxi/ssr
JSDOM SSR renderer for Svelte
23.2k---5 years ago
0.2.1
CommonJSNoneSvelte, Static site generators and meta-frameworks
@guidepup/playwright
Screen reader automation library for Playwright testing.
22.5k---1 month ago
0.19.1
CommonJSBundledAccessibility, Scraping and browser automation
playwright-smart-reporter
An intelligent Playwright HTML reporter with AI-powered failure analysis, flakiness detection, and performance regression alerts
22.2k--24 days ago
2.2.0
CommonJSBundledTesting, Scraping and browser automation
notion-md-crawler
A library to recursively retrieve and serialize Notion pages with customization for machine learning applications.
21.8k+26%-1 year ago
1.0.2
ESM + CommonJSBundledScraping and browser automation
playwright-test
Run mocha, zora, uvu, tape and benchmark.js scripts inside real browsers with playwright.
19.6k+23%-3 months ago
14.1.15
ESM onlyBundledTesting, Scraping and browser automation
firecrawl-cli
Command-line interface for Firecrawl. Scrape, crawl, and extract data from any website, and search a ~43M-abstract research paper index (PubMed, bioRxiv, medRxiv, arXiv), directly from your terminal.
19.4k--2 days ago
1.24.4
CommonJSBundledCLI tools and terminal utilities, Scraping and browser automation
@page-agent/page-controller
Page controller for page-agent - DOM operations and element interactions
19.3k---19 days ago
1.12.4
ESM onlyBundledDOM and browser utilities, Scraping and browser automation
@antiwork/shortest
AI-powered natural language end-to-end testing framework
18.7k---1 year ago
0.4.9
ESM + CommonJSBundledTesting, Scraping and browser automation
@playwright-testing-library/test
playwright + dom-testing-library
18.6k---3 years ago
4.5.0
CommonJSBundledScraping and browser automation, Testing
npm-license-crawler
Analyzes license information for multiple node.js modules (package.json files) as part of your software project.
18.5k-23%-7 years ago
0.2.1
CommonJSNoneScraping and browser automation
window
Exports a jsdom window object.
18.1k-31%-6 years ago
4.2.7
CommonJSNoneScraping and browser automation, Testing
url-metadata
Fetch a URL and scrape its metadata using Node.js or the browser. Can parse metadata from HTML strings or Response objects as well.
18.1k+21%-today
6.2.0
CommonJSBundledScraping and browser automation
pwa-asset-generator
Automates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.html files with the generated images according to Web App Manifest specs and Apple Hum
17.8k-35%-4 days ago
8.1.7
ESM onlyBundledScraping and browser automation
grunt-contrib-jasmine
Run jasmine specs headlessly through Headless Chrome
17k-15%--
4.0.0
CommonJSNoneScraping and browser automation
metascraper-date
Metascraper rule to extract the date from HTML using Open Graph, JSON-LD, and fallback selectors.
16.9k+214%-1 month ago
5.56.2
CommonJSBundledDate and time, Scraping and browser automation
playwright-ng-schematics
Playwright Angular schematics
16.6k+79%-21 days ago
22.1.2
-NoneTesting, Angular
@wreq-js/binding-linux-x64-gnu
Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
15.7k---1 month ago
3.2.0
CommonJSNoneHTTP clients, Scraping and browser automation
ts-web-scraper
A powerful web scraper for both static and client-side rendered sites using only Bun native APIs
15.2k+541%-1 month ago
0.1.10
ESM onlyBundledTypeScript tooling, Scraping and browser automation
robots-txt-parser
A lightweight robots.txt parser for Node.js with support for wildcards, caching and promises.
14.7k+348%-4 years ago
2.0.3
CommonJSNoneParsers and serialisers, Scraping and browser automation
@factory-js/factory
🏭 The object generator for testing
14.2k---7 months ago
0.4.2
ESM + CommonJSBundledORMs and query builders, Testing
@takumi-rs/core-linux-arm64-musl
Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
14.2k---10 days ago
2.14.0
CommonJSNoneCloud SDKs, Scraping and browser automation
rebrowser-puppeteer
A drop-in replacement for puppeteer patched with rebrowser-patches. It allows to pass modern automation detection tests.
13.7k-56%-1 year ago
24.8.1
ESM + CommonJSBundledScraping and browser automation
@page-agent/core
GUI agent for web applications - add intelligent automation to any webpage with a single script
13.7k---19 days ago
1.12.4
ESM onlyBundledScraping and browser automation, TypeScript tooling
simplecrawler
Very straightforward, event driven web crawler. Features a flexible queue interface and a basic cache mechanism with extensible backend.
13.1k-68%-6 years ago
1.1.9
CommonJSNoneScraping and browser automation
@wdio/static-server-service
A WebdriverIO service that can serve your static websites
13.1k---4 days ago
9.32.0
ESM onlyBundledScraping and browser automation, Testing
rebrowser-playwright-core
A drop-in replacement for playwright-core patched with rebrowser-patches. It allows to pass modern automation detection tests.
13k+453%-1 year ago
1.52.0
ESM + CommonJSBundledScraping and browser automation
@seontechnologies/playwright-utils
A collection of utilities for Playwright.
13k---3 days ago
4.4.1
ESM + CommonJSBundledScraping and browser automation, Testing
@global-cache/playwright
Key-value cache for sharing data between Playwright workers.
12.9k---3 months ago
0.5.1
CommonJSBundledTesting, Scraping and browser automation
nightmare
A high-level browser automation library.
12.6k+21%--
3.0.2
CommonJSNoneScraping and browser automation
@kpmck/ag-grid-core
Framework-agnostic AG Grid DOM helpers for Node-based test tooling
12k---3 months ago
1.0.1
ESM onlyBundledTesting, Scraping and browser automation
@usenotra/geo
Request capture SDK for AI traffic attribution and agent feedback collection
11.2k---7 days ago
0.5.0
ESM onlyBundledCloud SDKs, Vue
@takumi-rs/core-darwin-arm64
Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
11k---10 days ago
2.14.0
CommonJSNoneCloud SDKs, Scraping and browser automation
page-agent
GUI agent for web applications - add intelligent automation to any webpage with a single script
10k+54376%-19 days ago
1.12.4
ESM onlyBundledScraping and browser automation, TypeScript tooling
@takumi-rs/core-darwin-x64
Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
9.9k---10 days ago
2.14.0
CommonJSNoneCloud SDKs, Scraping and browser automation
rebrowser-playwright
A drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests.
9.9k+339%-1 year ago
1.52.0
ESM + CommonJSBundledScraping and browser automation
tuistory
Playwright-like testing for TUI applications
9.9k--1 month ago
0.11.0
ESM onlyBundledTesting, CLI tools and terminal utilities
@mozilla/firefox-devtools-mcp-moz
Model Context Protocol (MCP) server for Firefox DevTools automation (moz build with privileged context support)
9.1k---3 days ago
0.10.4
ESM onlyBundledScraping and browser automation
zombie
Insanely fast, full-stack, headless browser testing using Node.js
9.1k+7%-7 years ago
6.1.4
CommonJSNoneTesting, Scraping and browser automation
playwright-testing-library
playwright + dom-testing-library
9k+177%-3 years ago
4.5.0
CommonJSBundledScraping and browser automation, Testing
machinepack-http
Send HTTP requests, scrape webpages, and stream data in your JavaScript/Node.js/Sails.js app with a simple, `jQuery.get()`-like interface for sending HTTP requests and processing server responses.
8.8k+3%-3 years ago
9.0.0
-NoneScraping and browser automation
metascraper-amazon
Metascraper rules tailored for Amazon pages — richer metadata than generic HTML parsers.
8.6k+75%-1 month ago
5.56.2
CommonJSBundledPayments and commerce, Scraping and browser automation

12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.

  • @appium/execute-driver-pluginPlugin for batching and executing Appium driver commands
  • metascraper-authorMetascraper rule to extract the author from HTML using Open Graph, JSON-LD, and fallback selectors.
  • desired-capabilitiesutilities for parsing shorthand Selenium capabilities objects
  • @redhat-developer/page-objectsPage Object API implementation for a VS Code editor used by ExTester framework.
  • unfurl.jsScraper for oEmbed, Twitter Cards and Open Graph metadata - fast and Promise-based
  • app-store-scraperscrape data from the itunes app store
  • penthouseGenerate critical path CSS for web pages
  • @wreq-js/binding-win32-x64-msvcNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
  • happoCatch unexpected visual and accessibility changes and UI bugs
  • @roxi/ssrJSDOM SSR renderer for Svelte
  • @guidepup/playwrightScreen reader automation library for Playwright testing.
  • playwright-smart-reporterAn intelligent Playwright HTML reporter with AI-powered failure analysis, flakiness detection, and performance regression alerts
  • notion-md-crawlerA library to recursively retrieve and serialize Notion pages with customization for machine learning applications.
  • playwright-testRun mocha, zora, uvu, tape and benchmark.js scripts inside real browsers with playwright.
  • firecrawl-cliCommand-line interface for Firecrawl. Scrape, crawl, and extract data from any website, and search a ~43M-abstract research paper index (PubMed, bioRxiv, medRxiv, arXiv), directly from your terminal.
  • @page-agent/page-controllerPage controller for page-agent - DOM operations and element interactions
  • @antiwork/shortestAI-powered natural language end-to-end testing framework
  • @playwright-testing-library/testplaywright + dom-testing-library
  • npm-license-crawlerAnalyzes license information for multiple node.js modules (package.json files) as part of your software project.
  • windowExports a jsdom window object.
  • url-metadataFetch a URL and scrape its metadata using Node.js or the browser. Can parse metadata from HTML strings or Response objects as well.
  • pwa-asset-generatorAutomates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.html files with the generated images according to Web App Manifest specs and Apple Hum
  • grunt-contrib-jasmineRun jasmine specs headlessly through Headless Chrome
  • metascraper-dateMetascraper rule to extract the date from HTML using Open Graph, JSON-LD, and fallback selectors.
  • playwright-ng-schematicsPlaywright Angular schematics
  • @wreq-js/binding-linux-x64-gnuNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
  • ts-web-scraperA powerful web scraper for both static and client-side rendered sites using only Bun native APIs
  • robots-txt-parserA lightweight robots.txt parser for Node.js with support for wildcards, caching and promises.
  • @factory-js/factory🏭 The object generator for testing
  • @takumi-rs/core-linux-arm64-muslRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
  • rebrowser-puppeteerA drop-in replacement for puppeteer patched with rebrowser-patches. It allows to pass modern automation detection tests.
  • @page-agent/coreGUI agent for web applications - add intelligent automation to any webpage with a single script
  • simplecrawlerVery straightforward, event driven web crawler. Features a flexible queue interface and a basic cache mechanism with extensible backend.
  • @wdio/static-server-serviceA WebdriverIO service that can serve your static websites
  • rebrowser-playwright-coreA drop-in replacement for playwright-core patched with rebrowser-patches. It allows to pass modern automation detection tests.
  • @seontechnologies/playwright-utilsA collection of utilities for Playwright.
  • @global-cache/playwrightKey-value cache for sharing data between Playwright workers.
  • nightmareA high-level browser automation library.
  • @kpmck/ag-grid-coreFramework-agnostic AG Grid DOM helpers for Node-based test tooling
  • @usenotra/geoRequest capture SDK for AI traffic attribution and agent feedback collection
  • @takumi-rs/core-darwin-arm64Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
  • page-agentGUI agent for web applications - add intelligent automation to any webpage with a single script
  • @takumi-rs/core-darwin-x64Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
  • rebrowser-playwrightA drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests.
  • tuistoryPlaywright-like testing for TUI applications
  • @mozilla/firefox-devtools-mcp-mozModel Context Protocol (MCP) server for Firefox DevTools automation (moz build with privileged context support)
  • zombieInsanely fast, full-stack, headless browser testing using Node.js
  • playwright-testing-libraryplaywright + dom-testing-library
  • machinepack-httpSend HTTP requests, scrape webpages, and stream data in your JavaScript/Node.js/Sails.js app with a simple, `jQuery.get()`-like interface for sending HTTP requests and processing server responses.
  • metascraper-amazonMetascraper rules tailored for Amazon pages — richer metadata than generic HTML parsers.