Package category
Scraping and browser automation
Headless browsers, crawlers and HTML extraction.
373 packages7 comparisons
Packages compared
373 packages
| Package | Weekly downloads | 12-month change | 52 weeks | Gzip | Last release | Module | Types | Categories |
|---|---|---|---|---|---|---|---|---|
| @crawlee/linkedom The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer. | 110.9k | - | - | - | 1 month ago 3.18.1 | ESM + CommonJS | Bundled | Scraping and browser automation |
| crawlee The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer. | 109.8k | +210% | - | 1 month ago 3.18.1 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| pa11y-ci Pa11y CI is a CI-centric accessibility test runner, built using Pa11y | 96.6k | +29% | - | 4 months ago 4.1.1 | CommonJS | None | Accessibility, Testing | |
| google-play-scraper scrapes app data from google play store | 93k | +424% | - | 3 months ago 10.1.3 | ESM only | Bundled | Scraping and browser automation | |
| wreq-js Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible. | 91.9k | - | - | 1 month ago 3.2.0 | ESM + CommonJS | Bundled | HTTP clients, Scraping and browser automation | |
| @takumi-rs/core-linux-x64-musl Render OG images from Takumi node trees with native Node.js bindings. No headless browser. | 85.9k | - | - | - | 10 days ago 2.14.0 | CommonJS | None | Cloud SDKs, Scraping and browser automation |
| pi-web-access Web search, URL fetching, GitHub repo cloning, PDF extraction, YouTube video understanding, and local video analysis for Pi coding agent. Supports OpenAI, Brave, Parallel, TinyFish, Search1API, Searchinfinity, Querit, Tavily, Firecrawl, Crawl4AI, Jina, SE | 85.2k | - | - | 2 days ago 0.31.0 | ESM only | None | Scraping and browser automation, TypeScript tooling | |
| sitemapper Parser for XML Sitemaps to be used with Robots.txt and web crawlers | 83.3k | +161% | - | 4 months ago 4.1.6 | ESM only | Bundled | Scraping and browser automation, Parsers and serialisers | |
| sauce-connect-launcher A library to download and launch Sauce Connect. | 82k | -12% | - | 6 years ago 1.3.2 | CommonJS | None | Scraping and browser automation, Testing | |
| metascraper-logo Metascraper rule to extract the logo from HTML using Open Graph, JSON-LD, and fallback selectors. | 81k | +341% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Scraping and browser automation | |
| dom-miner dom-miner — mine any site into a compact DOM map for QA test plans and AI agents | 76.3k | - | - | 1 month ago 0.1.4 | ESM only | Bundled | DOM and browser utilities, Testing | |
| metascraper-description Metascraper rule to extract the description from HTML using Open Graph, JSON-LD, and fallback selectors. | 75k | +155% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Scraping and browser automation | |
| metascraper-image Metascraper rule to extract the image from HTML using Open Graph, JSON-LD, and fallback selectors. | 72.8k | +133% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Scraping and browser automation | |
| metascraper-title Metascraper rule to extract the title from HTML using Open Graph, JSON-LD, and fallback selectors. | 67.4k | +153% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Scraping and browser automation | |
| @axe-core/webdriverjs Provides a method to inject and analyze web pages using axe | 67.3k | - | - | - | 1 month ago 4.13.0 | ESM + CommonJS | Bundled | Accessibility, Testing |
| sitemapd Runtime-neutral sitemap parsing and bounded traversal | 63.4k | - | - | 1 month ago 0.2.2 | ESM only | Bundled | Parsers and serialisers, Scraping and browser automation | |
| metascraper-logo-favicon Metascraper logo fallback that picks favicons and apple-touch icons from HTML. | 62.5k | +149% | - | 7 days ago 5.58.1 | CommonJS | Bundled | Scraping and browser automation | |
| domparser-rs A super fast html parser and manipulator written in rust. | 60.6k | - | - | 5 months ago 0.1.1 | CommonJS | Bundled | Parsers and serialisers, Scraping and browser automation | |
| @wdio/json-reporter A WebdriverIO plugin to report results in json format. | 56.8k | - | - | - | 4 days ago 9.32.0 | ESM + CommonJS | Bundled | Scraping and browser automation, Testing |
| metascraper-url Metascraper rule to extract the url from HTML using Open Graph, JSON-LD, and fallback selectors. | 55.6k | +215% | - | 1 month ago 5.56.2 | CommonJS | Bundled | URLs and query strings, Scraping and browser automation | |
| @mcp-b/transports Browser transport implementations for Model Context Protocol (MCP) - postMessage, Chrome extension messaging, and iframe communication for AI agents and LLMs | 55k | - | - | - | 25 days ago 5.1.0 | ESM only | Bundled | Scraping and browser automation |
| domparser-linux-x64-gnu A super fast html parser and manipulator written in rust. | 53.9k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| selenium-server Selenium in an npm package | 53.3k | +9% | - | 7 years ago 3.141.59 | CommonJS | None | Testing, Scraping and browser automation | |
| rebrowser-puppeteer-core A drop-in replacement for puppeteer-core patched with rebrowser-patches. It allows to pass modern automation detection tests. | 53k | +33% | - | 1 year ago 24.8.1 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| dom-parser Fast dom parser based on regexps | 51.5k | -3% | - | 2 years ago 1.1.5 | CommonJS | Bundled | Parsers and serialisers, Scraping and browser automation | |
| puppeteer-screen-recorder A puppeteer Plugin that uses the native chrome devtool protocol for capturing video frame by frame. Also supports an option to follow pages that are opened by the current page object | 50.7k | -59% | - | 1 year ago 3.0.6 | ESM + CommonJS | Bundled | Scraping and browser automation, Video and audio | |
| cloakbrowser Stealth Chromium that passes every bot detection test. Drop-in Playwright/Puppeteer replacement with source-level fingerprint patches. | 50k | - | - | today 0.5.11 | ESM only | Bundled | Scraping and browser automation | |
| domparser-linux-x64-musl A super fast html parser and manipulator written in rust. | 45.2k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| passmark The open-source AI framework for regression testing. | 44.1k | - | - | 3 months ago 1.0.16 | CommonJS | Bundled | Testing, Scraping and browser automation | |
| @vreden/youtube_scraper A simple YouTube video downloader for audio and video formats with resolusi and quality. | 42.2k | - | - | - | 4 months ago 1.2.9 | CommonJS | None | Video and audio, Scraping and browser automation |
| eslint-config-sheriff A comprehensive and opinionated TypeScript-first ESLint configuration. | 42k | +313% | - | 3 months ago 31.4.0 | ESM only | Bundled | Linting and formatting, React | |
| filereader HTML5 FileAPI `FileReader` for Node.JS. | 40.7k | +116% | - | 11 years ago 0.10.3 | CommonJS | None | Scraping and browser automation | |
| @unlighthouse/client UI Client for Unlighthouse. | 40.1k | - | - | - | 3 days ago 0.18.1 | ESM only | None | Scraping and browser automation |
| storycap A Storybook addon, Save the screenshot image of your stories! via puppeteer. | 39.9k | -30% | - | 2 years ago 5.0.1 | ESM + CommonJS | Bundled | Documentation tooling, Scraping and browser automation | |
| @applitools/eyes-playwright Applitools Eyes SDK for Playwright | 39.6k | - | - | - | 1 day ago 1.49.2 | CommonJS | Bundled | Testing, Scraping and browser automation |
| apify The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer. | 38.6k | +58% | - | 4 months ago 3.7.2 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| @hyperbrowser/sdk Node SDK for Hyperbrowser API | 36.1k | - | - | - | 2 days ago 0.93.0 | CommonJS | Bundled | Scraping and browser automation |
| isbot-fast JavaScript module detecting bots/crawlers/spiders via user-agent | 36k | +59% | - | 6 years ago 1.2.0 | CommonJS | None | Scraping and browser automation | |
| metascraper-publisher Metascraper rule to extract the publisher from HTML using Open Graph, JSON-LD, and fallback selectors. | 33.5k | +135% | - | 1 month ago 5.56.2 | CommonJS | Bundled | Scraping and browser automation | |
| @takumi-rs/core-linux-arm64-gnu Render OG images from Takumi node trees with native Node.js bindings. No headless browser. | 33.4k | - | - | - | 10 days ago 2.14.0 | CommonJS | None | Cloud SDKs, Scraping and browser automation |
| wdio-intercept-service Capture and assert HTTP ajax calls in webdriver.io 🕸 | 32.4k | -32% | - | 2 years ago 4.4.1 | CommonJS | Bundled | Testing, Scraping and browser automation | |
| @datafast/ai-crawl Server-side AI crawler tracking for DataFast | 31.9k | - | - | - | 2 months ago 1.0.9 | ESM + CommonJS | Bundled | Scraping and browser automation |
| tiktok-live-api-sdk TikTok LIVE API SDK for Node.js & TypeScript. Real-time TikTok LIVE chat, gifts, likes, follows, viewer counts and PK battles via the EulerStream managed TikTok LIVE API, with typed clients for webcast signing, rooms, gifts, rankings, LIVE alerts, moderat | 31.5k | - | - | 3 days ago 0.6.1 | ESM only | Bundled | Scraping and browser automation, WebSockets and realtime | |
| vscode-extension-tester ExTester is a package that is designed to help you run UI tests for your Visual Studio Code extensions using selenium-webdriver. | 28.7k | -26% | - | 11 days ago 8.27.0 | CommonJS | Bundled | Testing, Scraping and browser automation | |
| chrome-aws-lambda Chromium Binary for AWS Lambda and Google Cloud Functions | 27.9k | -25% | - | 5 years ago 10.1.0 | CommonJS | Bundled | Cloud SDKs, Scraping and browser automation | |
| @askjo/camofox-browser Headless browser automation server and OpenClaw plugin for AI agents - anti-detection, element refs, and session isolation | 26.8k | - | - | - | 2 days ago 1.17.0 | ESM only | None | Scraping and browser automation |
| wdio-rerun-service A WebdriverIO service to track and stage for re-running failed or flaky Jasmine/Mocha tests or Cucumber Scenarios. | 26.6k | +154% | - | 8 months ago 3.0.1 | ESM + CommonJS | Bundled | Testing, Scraping and browser automation | |
| @prerenderer/renderer-puppeteer A renderer for @prerenderer/prerenderer that uses puppeteer to prerender pages. | 26.3k | - | - | - | 2 years ago 1.2.4 | ESM + CommonJS | Bundled | Scraping and browser automation |
| @redhat-developer/locators Pluggable Page Objects locators for an ExTester framework. | 25.9k | - | - | - | 11 days ago 1.24.0 | CommonJS | Bundled | Testing, Scraping and browser automation |
| grunt-contrib-qunit Run QUnit unit tests in a headless Chrome instance | 25.8k | -35% | - | - 10.2.0 | CommonJS | None | Scraping and browser automation |
12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.
- @crawlee/linkedomThe scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
- crawleeThe scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
- pa11y-ciPa11y CI is a CI-centric accessibility test runner, built using Pa11y
- google-play-scraperscrapes app data from google play store
- wreq-jsNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
- @takumi-rs/core-linux-x64-muslRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
- pi-web-accessWeb search, URL fetching, GitHub repo cloning, PDF extraction, YouTube video understanding, and local video analysis for Pi coding agent. Supports OpenAI, Brave, Parallel, TinyFish, Search1API, Searchinfinity, Querit, Tavily, Firecrawl, Crawl4AI, Jina, SE
- sitemapperParser for XML Sitemaps to be used with Robots.txt and web crawlers
- sauce-connect-launcherA library to download and launch Sauce Connect.
- metascraper-logoMetascraper rule to extract the logo from HTML using Open Graph, JSON-LD, and fallback selectors.
- dom-minerdom-miner — mine any site into a compact DOM map for QA test plans and AI agents
- metascraper-descriptionMetascraper rule to extract the description from HTML using Open Graph, JSON-LD, and fallback selectors.
- metascraper-imageMetascraper rule to extract the image from HTML using Open Graph, JSON-LD, and fallback selectors.
- metascraper-titleMetascraper rule to extract the title from HTML using Open Graph, JSON-LD, and fallback selectors.
- @axe-core/webdriverjsProvides a method to inject and analyze web pages using axe
- sitemapdRuntime-neutral sitemap parsing and bounded traversal
- metascraper-logo-faviconMetascraper logo fallback that picks favicons and apple-touch icons from HTML.
- domparser-rsA super fast html parser and manipulator written in rust.
- @wdio/json-reporterA WebdriverIO plugin to report results in json format.
- metascraper-urlMetascraper rule to extract the url from HTML using Open Graph, JSON-LD, and fallback selectors.
- @mcp-b/transportsBrowser transport implementations for Model Context Protocol (MCP) - postMessage, Chrome extension messaging, and iframe communication for AI agents and LLMs
- domparser-linux-x64-gnuA super fast html parser and manipulator written in rust.
- selenium-serverSelenium in an npm package
- rebrowser-puppeteer-coreA drop-in replacement for puppeteer-core patched with rebrowser-patches. It allows to pass modern automation detection tests.
- dom-parserFast dom parser based on regexps
- puppeteer-screen-recorderA puppeteer Plugin that uses the native chrome devtool protocol for capturing video frame by frame. Also supports an option to follow pages that are opened by the current page object
- cloakbrowserStealth Chromium that passes every bot detection test. Drop-in Playwright/Puppeteer replacement with source-level fingerprint patches.
- domparser-linux-x64-muslA super fast html parser and manipulator written in rust.
- passmarkThe open-source AI framework for regression testing.
- @vreden/youtube_scraperA simple YouTube video downloader for audio and video formats with resolusi and quality.
- eslint-config-sheriffA comprehensive and opinionated TypeScript-first ESLint configuration.
- filereaderHTML5 FileAPI `FileReader` for Node.JS.
- @unlighthouse/clientUI Client for Unlighthouse.
- storycapA Storybook addon, Save the screenshot image of your stories! via puppeteer.
- @applitools/eyes-playwrightApplitools Eyes SDK for Playwright
- apifyThe scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
- @hyperbrowser/sdkNode SDK for Hyperbrowser API
- isbot-fastJavaScript module detecting bots/crawlers/spiders via user-agent
- metascraper-publisherMetascraper rule to extract the publisher from HTML using Open Graph, JSON-LD, and fallback selectors.
- @takumi-rs/core-linux-arm64-gnuRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
- wdio-intercept-serviceCapture and assert HTTP ajax calls in webdriver.io 🕸
- @datafast/ai-crawlServer-side AI crawler tracking for DataFast
- tiktok-live-api-sdkTikTok LIVE API SDK for Node.js & TypeScript. Real-time TikTok LIVE chat, gifts, likes, follows, viewer counts and PK battles via the EulerStream managed TikTok LIVE API, with typed clients for webcast signing, rooms, gifts, rankings, LIVE alerts, moderat
- vscode-extension-testerExTester is a package that is designed to help you run UI tests for your Visual Studio Code extensions using selenium-webdriver.
- chrome-aws-lambdaChromium Binary for AWS Lambda and Google Cloud Functions
- @askjo/camofox-browserHeadless browser automation server and OpenClaw plugin for AI agents - anti-detection, element refs, and session isolation
- wdio-rerun-serviceA WebdriverIO service to track and stage for re-running failed or flaky Jasmine/Mocha tests or Cucumber Scenarios.
- @prerenderer/renderer-puppeteerA renderer for @prerenderer/prerenderer that uses puppeteer to prerender pages.
- @redhat-developer/locatorsPluggable Page Objects locators for an ExTester framework.
- grunt-contrib-qunitRun QUnit unit tests in a headless Chrome instance