Package category
Scraping and browser automation
Headless browsers, crawlers and HTML extraction.
373 packages7 comparisons
Packages compared
373 packages
| Package | Weekly downloads | 12-month change | 52 weeks | Gzip | Last release | Module | Types | Categories |
|---|---|---|---|---|---|---|---|---|
| @appium/execute-driver-plugin Plugin for batching and executing Appium driver commands | 25.7k | - | - | - | today 7.0.1 | ESM only | Bundled | Scraping and browser automation, Testing |
| metascraper-author Metascraper rule to extract the author from HTML using Open Graph, JSON-LD, and fallback selectors. | 25.3k | +258% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Scraping and browser automation | |
| desired-capabilities utilities for parsing shorthand Selenium capabilities objects | 25.2k | +11% | - | 11 years ago 0.1.0 | CommonJS | None | Scraping and browser automation, Testing | |
| @redhat-developer/page-objects Page Object API implementation for a VS Code editor used by ExTester framework. | 24.6k | - | - | - | 11 days ago 1.24.0 | CommonJS | Bundled | Testing, Scraping and browser automation |
| unfurl.js Scraper for oEmbed, Twitter Cards and Open Graph metadata - fast and Promise-based | 24.2k | +88% | - | 2 years ago 6.4.0 | CommonJS | Bundled | Scraping and browser automation | |
| app-store-scraper scrape data from the itunes app store | 24.2k | +446% | - | 2 years ago 0.18.0 | CommonJS | None | Scraping and browser automation | |
| penthouse Generate critical path CSS for web pages | 24.1k | -29% | - | 4 years ago 2.3.3 | CommonJS | None | CSS-in-JS and styling, Scraping and browser automation | |
| @wreq-js/binding-win32-x64-msvc Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible. | 23.9k | - | - | - | 1 month ago 3.2.0 | CommonJS | None | HTTP clients, Scraping and browser automation |
| happo Catch unexpected visual and accessibility changes and UI bugs | 23.3k | +31445% | - | today 6.20.0 | ESM only | Bundled | Testing, Scraping and browser automation | |
| @roxi/ssr JSDOM SSR renderer for Svelte | 23.2k | - | - | - | 5 years ago 0.2.1 | CommonJS | None | Svelte, Static site generators and meta-frameworks |
| @guidepup/playwright Screen reader automation library for Playwright testing. | 22.5k | - | - | - | 1 month ago 0.19.1 | CommonJS | Bundled | Accessibility, Scraping and browser automation |
| playwright-smart-reporter An intelligent Playwright HTML reporter with AI-powered failure analysis, flakiness detection, and performance regression alerts | 22.2k | - | - | 24 days ago 2.2.0 | CommonJS | Bundled | Testing, Scraping and browser automation | |
| notion-md-crawler A library to recursively retrieve and serialize Notion pages with customization for machine learning applications. | 21.8k | +26% | - | 1 year ago 1.0.2 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| playwright-test Run mocha, zora, uvu, tape and benchmark.js scripts inside real browsers with playwright. | 19.6k | +23% | - | 3 months ago 14.1.15 | ESM only | Bundled | Testing, Scraping and browser automation | |
| firecrawl-cli Command-line interface for Firecrawl. Scrape, crawl, and extract data from any website, and search a ~43M-abstract research paper index (PubMed, bioRxiv, medRxiv, arXiv), directly from your terminal. | 19.4k | - | - | 2 days ago 1.24.4 | CommonJS | Bundled | CLI tools and terminal utilities, Scraping and browser automation | |
| @page-agent/page-controller Page controller for page-agent - DOM operations and element interactions | 19.3k | - | - | - | 19 days ago 1.12.4 | ESM only | Bundled | DOM and browser utilities, Scraping and browser automation |
| @antiwork/shortest AI-powered natural language end-to-end testing framework | 18.7k | - | - | - | 1 year ago 0.4.9 | ESM + CommonJS | Bundled | Testing, Scraping and browser automation |
| @playwright-testing-library/test playwright + dom-testing-library | 18.6k | - | - | - | 3 years ago 4.5.0 | CommonJS | Bundled | Scraping and browser automation, Testing |
| npm-license-crawler Analyzes license information for multiple node.js modules (package.json files) as part of your software project. | 18.5k | -23% | - | 7 years ago 0.2.1 | CommonJS | None | Scraping and browser automation | |
| window Exports a jsdom window object. | 18.1k | -31% | - | 6 years ago 4.2.7 | CommonJS | None | Scraping and browser automation, Testing | |
| url-metadata Fetch a URL and scrape its metadata using Node.js or the browser. Can parse metadata from HTML strings or Response objects as well. | 18.1k | +21% | - | today 6.2.0 | CommonJS | Bundled | Scraping and browser automation | |
| pwa-asset-generator Automates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.html files with the generated images according to Web App Manifest specs and Apple Hum | 17.8k | -35% | - | 4 days ago 8.1.7 | ESM only | Bundled | Scraping and browser automation | |
| grunt-contrib-jasmine Run jasmine specs headlessly through Headless Chrome | 17k | -15% | - | - 4.0.0 | CommonJS | None | Scraping and browser automation | |
| metascraper-date Metascraper rule to extract the date from HTML using Open Graph, JSON-LD, and fallback selectors. | 16.9k | +214% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Date and time, Scraping and browser automation | |
| playwright-ng-schematics Playwright Angular schematics | 16.6k | +79% | - | 21 days ago 22.1.2 | - | None | Testing, Angular | |
| @wreq-js/binding-linux-x64-gnu Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible. | 15.7k | - | - | - | 1 month ago 3.2.0 | CommonJS | None | HTTP clients, Scraping and browser automation |
| ts-web-scraper A powerful web scraper for both static and client-side rendered sites using only Bun native APIs | 15.2k | +541% | - | 1 month ago 0.1.10 | ESM only | Bundled | TypeScript tooling, Scraping and browser automation | |
| robots-txt-parser A lightweight robots.txt parser for Node.js with support for wildcards, caching and promises. | 14.7k | +348% | - | 4 years ago 2.0.3 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| @factory-js/factory 🏭 The object generator for testing | 14.2k | - | - | - | 7 months ago 0.4.2 | ESM + CommonJS | Bundled | ORMs and query builders, Testing |
| @takumi-rs/core-linux-arm64-musl Render OG images from Takumi node trees with native Node.js bindings. No headless browser. | 14.2k | - | - | - | 10 days ago 2.14.0 | CommonJS | None | Cloud SDKs, Scraping and browser automation |
| rebrowser-puppeteer A drop-in replacement for puppeteer patched with rebrowser-patches. It allows to pass modern automation detection tests. | 13.7k | -56% | - | 1 year ago 24.8.1 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| @page-agent/core GUI agent for web applications - add intelligent automation to any webpage with a single script | 13.7k | - | - | - | 19 days ago 1.12.4 | ESM only | Bundled | Scraping and browser automation, TypeScript tooling |
| simplecrawler Very straightforward, event driven web crawler. Features a flexible queue interface and a basic cache mechanism with extensible backend. | 13.1k | -68% | - | 6 years ago 1.1.9 | CommonJS | None | Scraping and browser automation | |
| @wdio/static-server-service A WebdriverIO service that can serve your static websites | 13.1k | - | - | - | 4 days ago 9.32.0 | ESM only | Bundled | Scraping and browser automation, Testing |
| rebrowser-playwright-core A drop-in replacement for playwright-core patched with rebrowser-patches. It allows to pass modern automation detection tests. | 13k | +453% | - | 1 year ago 1.52.0 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| @seontechnologies/playwright-utils A collection of utilities for Playwright. | 13k | - | - | - | 3 days ago 4.4.1 | ESM + CommonJS | Bundled | Scraping and browser automation, Testing |
| @global-cache/playwright Key-value cache for sharing data between Playwright workers. | 12.9k | - | - | - | 3 months ago 0.5.1 | CommonJS | Bundled | Testing, Scraping and browser automation |
| nightmare A high-level browser automation library. | 12.6k | +21% | - | - 3.0.2 | CommonJS | None | Scraping and browser automation | |
| @kpmck/ag-grid-core Framework-agnostic AG Grid DOM helpers for Node-based test tooling | 12k | - | - | - | 3 months ago 1.0.1 | ESM only | Bundled | Testing, Scraping and browser automation |
| @usenotra/geo Request capture SDK for AI traffic attribution and agent feedback collection | 11.2k | - | - | - | 7 days ago 0.5.0 | ESM only | Bundled | Cloud SDKs, Vue |
| @takumi-rs/core-darwin-arm64 Render OG images from Takumi node trees with native Node.js bindings. No headless browser. | 11k | - | - | - | 10 days ago 2.14.0 | CommonJS | None | Cloud SDKs, Scraping and browser automation |
| page-agent GUI agent for web applications - add intelligent automation to any webpage with a single script | 10k | +54376% | - | 19 days ago 1.12.4 | ESM only | Bundled | Scraping and browser automation, TypeScript tooling | |
| @takumi-rs/core-darwin-x64 Render OG images from Takumi node trees with native Node.js bindings. No headless browser. | 9.9k | - | - | - | 10 days ago 2.14.0 | CommonJS | None | Cloud SDKs, Scraping and browser automation |
| rebrowser-playwright A drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests. | 9.9k | +339% | - | 1 year ago 1.52.0 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| tuistory Playwright-like testing for TUI applications | 9.9k | - | - | 1 month ago 0.11.0 | ESM only | Bundled | Testing, CLI tools and terminal utilities | |
| @mozilla/firefox-devtools-mcp-moz Model Context Protocol (MCP) server for Firefox DevTools automation (moz build with privileged context support) | 9.1k | - | - | - | 3 days ago 0.10.4 | ESM only | Bundled | Scraping and browser automation |
| zombie Insanely fast, full-stack, headless browser testing using Node.js | 9.1k | +7% | - | 7 years ago 6.1.4 | CommonJS | None | Testing, Scraping and browser automation | |
| playwright-testing-library playwright + dom-testing-library | 9k | +177% | - | 3 years ago 4.5.0 | CommonJS | Bundled | Scraping and browser automation, Testing | |
| machinepack-http Send HTTP requests, scrape webpages, and stream data in your JavaScript/Node.js/Sails.js app with a simple, `jQuery.get()`-like interface for sending HTTP requests and processing server responses. | 8.8k | +3% | - | 3 years ago 9.0.0 | - | None | Scraping and browser automation | |
| metascraper-amazon Metascraper rules tailored for Amazon pages — richer metadata than generic HTML parsers. | 8.6k | +75% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Payments and commerce, Scraping and browser automation |
12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.
- @appium/execute-driver-pluginPlugin for batching and executing Appium driver commands
- metascraper-authorMetascraper rule to extract the author from HTML using Open Graph, JSON-LD, and fallback selectors.
- desired-capabilitiesutilities for parsing shorthand Selenium capabilities objects
- @redhat-developer/page-objectsPage Object API implementation for a VS Code editor used by ExTester framework.
- unfurl.jsScraper for oEmbed, Twitter Cards and Open Graph metadata - fast and Promise-based
- app-store-scraperscrape data from the itunes app store
- penthouseGenerate critical path CSS for web pages
- @wreq-js/binding-win32-x64-msvcNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
- happoCatch unexpected visual and accessibility changes and UI bugs
- @roxi/ssrJSDOM SSR renderer for Svelte
- @guidepup/playwrightScreen reader automation library for Playwright testing.
- playwright-smart-reporterAn intelligent Playwright HTML reporter with AI-powered failure analysis, flakiness detection, and performance regression alerts
- notion-md-crawlerA library to recursively retrieve and serialize Notion pages with customization for machine learning applications.
- playwright-testRun mocha, zora, uvu, tape and benchmark.js scripts inside real browsers with playwright.
- firecrawl-cliCommand-line interface for Firecrawl. Scrape, crawl, and extract data from any website, and search a ~43M-abstract research paper index (PubMed, bioRxiv, medRxiv, arXiv), directly from your terminal.
- @page-agent/page-controllerPage controller for page-agent - DOM operations and element interactions
- @antiwork/shortestAI-powered natural language end-to-end testing framework
- @playwright-testing-library/testplaywright + dom-testing-library
- npm-license-crawlerAnalyzes license information for multiple node.js modules (package.json files) as part of your software project.
- windowExports a jsdom window object.
- url-metadataFetch a URL and scrape its metadata using Node.js or the browser. Can parse metadata from HTML strings or Response objects as well.
- pwa-asset-generatorAutomates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.html files with the generated images according to Web App Manifest specs and Apple Hum
- grunt-contrib-jasmineRun jasmine specs headlessly through Headless Chrome
- metascraper-dateMetascraper rule to extract the date from HTML using Open Graph, JSON-LD, and fallback selectors.
- playwright-ng-schematicsPlaywright Angular schematics
- @wreq-js/binding-linux-x64-gnuNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
- ts-web-scraperA powerful web scraper for both static and client-side rendered sites using only Bun native APIs
- robots-txt-parserA lightweight robots.txt parser for Node.js with support for wildcards, caching and promises.
- @factory-js/factory🏭 The object generator for testing
- @takumi-rs/core-linux-arm64-muslRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
- rebrowser-puppeteerA drop-in replacement for puppeteer patched with rebrowser-patches. It allows to pass modern automation detection tests.
- @page-agent/coreGUI agent for web applications - add intelligent automation to any webpage with a single script
- simplecrawlerVery straightforward, event driven web crawler. Features a flexible queue interface and a basic cache mechanism with extensible backend.
- @wdio/static-server-serviceA WebdriverIO service that can serve your static websites
- rebrowser-playwright-coreA drop-in replacement for playwright-core patched with rebrowser-patches. It allows to pass modern automation detection tests.
- @seontechnologies/playwright-utilsA collection of utilities for Playwright.
- @global-cache/playwrightKey-value cache for sharing data between Playwright workers.
- nightmareA high-level browser automation library.
- @kpmck/ag-grid-coreFramework-agnostic AG Grid DOM helpers for Node-based test tooling
- @usenotra/geoRequest capture SDK for AI traffic attribution and agent feedback collection
- @takumi-rs/core-darwin-arm64Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
- page-agentGUI agent for web applications - add intelligent automation to any webpage with a single script
- @takumi-rs/core-darwin-x64Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
- rebrowser-playwrightA drop-in replacement for playwright patched with rebrowser-patches. It allows to pass modern automation detection tests.
- tuistoryPlaywright-like testing for TUI applications
- @mozilla/firefox-devtools-mcp-mozModel Context Protocol (MCP) server for Firefox DevTools automation (moz build with privileged context support)
- zombieInsanely fast, full-stack, headless browser testing using Node.js
- playwright-testing-libraryplaywright + dom-testing-library
- machinepack-httpSend HTTP requests, scrape webpages, and stream data in your JavaScript/Node.js/Sails.js app with a simple, `jQuery.get()`-like interface for sending HTTP requests and processing server responses.
- metascraper-amazonMetascraper rules tailored for Amazon pages — richer metadata than generic HTML parsers.