Skip to content
JS
Package category

Scraping and browser automation

Headless browsers, crawlers and HTML extraction.

373 packages7 comparisons

Packages compared

373 packages
PackageWeekly downloads12-month change52 weeksGzipLast releaseModuleTypesCategories
@crawlee/linkedom
The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
110.9k---1 month ago
3.18.1
ESM + CommonJSBundledScraping and browser automation
crawlee
The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
109.8k+210%-1 month ago
3.18.1
ESM + CommonJSBundledScraping and browser automation
pa11y-ci
Pa11y CI is a CI-centric accessibility test runner, built using Pa11y
96.6k+29%-4 months ago
4.1.1
CommonJSNoneAccessibility, Testing
google-play-scraper
scrapes app data from google play store
93k+424%-3 months ago
10.1.3
ESM onlyBundledScraping and browser automation
wreq-js
Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
91.9k--1 month ago
3.2.0
ESM + CommonJSBundledHTTP clients, Scraping and browser automation
@takumi-rs/core-linux-x64-musl
Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
85.9k---10 days ago
2.14.0
CommonJSNoneCloud SDKs, Scraping and browser automation
pi-web-access
Web search, URL fetching, GitHub repo cloning, PDF extraction, YouTube video understanding, and local video analysis for Pi coding agent. Supports OpenAI, Brave, Parallel, TinyFish, Search1API, Searchinfinity, Querit, Tavily, Firecrawl, Crawl4AI, Jina, SE
85.2k--2 days ago
0.31.0
ESM onlyNoneScraping and browser automation, TypeScript tooling
sitemapper
Parser for XML Sitemaps to be used with Robots.txt and web crawlers
83.3k+161%-4 months ago
4.1.6
ESM onlyBundledScraping and browser automation, Parsers and serialisers
sauce-connect-launcher
A library to download and launch Sauce Connect.
82k-12%-6 years ago
1.3.2
CommonJSNoneScraping and browser automation, Testing
metascraper-logo
Metascraper rule to extract the logo from HTML using Open Graph, JSON-LD, and fallback selectors.
81k+341%-1 month ago
5.56.2
CommonJSBundledScraping and browser automation
dom-miner
dom-miner — mine any site into a compact DOM map for QA test plans and AI agents
76.3k--1 month ago
0.1.4
ESM onlyBundledDOM and browser utilities, Testing
metascraper-description
Metascraper rule to extract the description from HTML using Open Graph, JSON-LD, and fallback selectors.
75k+155%-1 month ago
5.56.2
CommonJSBundledScraping and browser automation
metascraper-image
Metascraper rule to extract the image from HTML using Open Graph, JSON-LD, and fallback selectors.
72.8k+133%-1 month ago
5.56.2
CommonJSBundledScraping and browser automation
metascraper-title
Metascraper rule to extract the title from HTML using Open Graph, JSON-LD, and fallback selectors.
67.4k+153%-1 month ago
5.56.2
CommonJSBundledScraping and browser automation
@axe-core/webdriverjs
Provides a method to inject and analyze web pages using axe
67.3k---1 month ago
4.13.0
ESM + CommonJSBundledAccessibility, Testing
sitemapd
Runtime-neutral sitemap parsing and bounded traversal
63.4k--1 month ago
0.2.2
ESM onlyBundledParsers and serialisers, Scraping and browser automation
metascraper-logo-favicon
Metascraper logo fallback that picks favicons and apple-touch icons from HTML.
62.5k+149%-7 days ago
5.58.1
CommonJSBundledScraping and browser automation
domparser-rs
A super fast html parser and manipulator written in rust.
60.6k--5 months ago
0.1.1
CommonJSBundledParsers and serialisers, Scraping and browser automation
@wdio/json-reporter
A WebdriverIO plugin to report results in json format.
56.8k---4 days ago
9.32.0
ESM + CommonJSBundledScraping and browser automation, Testing
metascraper-url
Metascraper rule to extract the url from HTML using Open Graph, JSON-LD, and fallback selectors.
55.6k+215%-1 month ago
5.56.2
CommonJSBundledURLs and query strings, Scraping and browser automation
@mcp-b/transports
Browser transport implementations for Model Context Protocol (MCP) - postMessage, Chrome extension messaging, and iframe communication for AI agents and LLMs
55k---25 days ago
5.1.0
ESM onlyBundledScraping and browser automation
domparser-linux-x64-gnu
A super fast html parser and manipulator written in rust.
53.9k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
selenium-server
Selenium in an npm package
53.3k+9%-7 years ago
3.141.59
CommonJSNoneTesting, Scraping and browser automation
rebrowser-puppeteer-core
A drop-in replacement for puppeteer-core patched with rebrowser-patches. It allows to pass modern automation detection tests.
53k+33%-1 year ago
24.8.1
ESM + CommonJSBundledScraping and browser automation
dom-parser
Fast dom parser based on regexps
51.5k-3%-2 years ago
1.1.5
CommonJSBundledParsers and serialisers, Scraping and browser automation
puppeteer-screen-recorder
A puppeteer Plugin that uses the native chrome devtool protocol for capturing video frame by frame. Also supports an option to follow pages that are opened by the current page object
50.7k-59%-1 year ago
3.0.6
ESM + CommonJSBundledScraping and browser automation, Video and audio
cloakbrowser
Stealth Chromium that passes every bot detection test. Drop-in Playwright/Puppeteer replacement with source-level fingerprint patches.
50k--today
0.5.11
ESM onlyBundledScraping and browser automation
domparser-linux-x64-musl
A super fast html parser and manipulator written in rust.
45.2k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
passmark
The open-source AI framework for regression testing.
44.1k--3 months ago
1.0.16
CommonJSBundledTesting, Scraping and browser automation
@vreden/youtube_scraper
A simple YouTube video downloader for audio and video formats with resolusi and quality.
42.2k---4 months ago
1.2.9
CommonJSNoneVideo and audio, Scraping and browser automation
eslint-config-sheriff
A comprehensive and opinionated TypeScript-first ESLint configuration.
42k+313%-3 months ago
31.4.0
ESM onlyBundledLinting and formatting, React
filereader
HTML5 FileAPI `FileReader` for Node.JS.
40.7k+116%-11 years ago
0.10.3
CommonJSNoneScraping and browser automation
@unlighthouse/client
UI Client for Unlighthouse.
40.1k---3 days ago
0.18.1
ESM onlyNoneScraping and browser automation
storycap
A Storybook addon, Save the screenshot image of your stories! via puppeteer.
39.9k-30%-2 years ago
5.0.1
ESM + CommonJSBundledDocumentation tooling, Scraping and browser automation
@applitools/eyes-playwright
Applitools Eyes SDK for Playwright
39.6k---1 day ago
1.49.2
CommonJSBundledTesting, Scraping and browser automation
apify
The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
38.6k+58%-4 months ago
3.7.2
ESM + CommonJSBundledScraping and browser automation
@hyperbrowser/sdk
Node SDK for Hyperbrowser API
36.1k---2 days ago
0.93.0
CommonJSBundledScraping and browser automation
isbot-fast
JavaScript module detecting bots/crawlers/spiders via user-agent
36k+59%-6 years ago
1.2.0
CommonJSNoneScraping and browser automation
metascraper-publisher
Metascraper rule to extract the publisher from HTML using Open Graph, JSON-LD, and fallback selectors.
33.5k+135%-1 month ago
5.56.2
CommonJSBundledScraping and browser automation
@takumi-rs/core-linux-arm64-gnu
Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
33.4k---10 days ago
2.14.0
CommonJSNoneCloud SDKs, Scraping and browser automation
wdio-intercept-service
Capture and assert HTTP ajax calls in webdriver.io 🕸
32.4k-32%-2 years ago
4.4.1
CommonJSBundledTesting, Scraping and browser automation
@datafast/ai-crawl
Server-side AI crawler tracking for DataFast
31.9k---2 months ago
1.0.9
ESM + CommonJSBundledScraping and browser automation
tiktok-live-api-sdk
TikTok LIVE API SDK for Node.js & TypeScript. Real-time TikTok LIVE chat, gifts, likes, follows, viewer counts and PK battles via the EulerStream managed TikTok LIVE API, with typed clients for webcast signing, rooms, gifts, rankings, LIVE alerts, moderat
31.5k--3 days ago
0.6.1
ESM onlyBundledScraping and browser automation, WebSockets and realtime
vscode-extension-tester
ExTester is a package that is designed to help you run UI tests for your Visual Studio Code extensions using selenium-webdriver.
28.7k-26%-11 days ago
8.27.0
CommonJSBundledTesting, Scraping and browser automation
chrome-aws-lambda
Chromium Binary for AWS Lambda and Google Cloud Functions
27.9k-25%-5 years ago
10.1.0
CommonJSBundledCloud SDKs, Scraping and browser automation
@askjo/camofox-browser
Headless browser automation server and OpenClaw plugin for AI agents - anti-detection, element refs, and session isolation
26.8k---2 days ago
1.17.0
ESM onlyNoneScraping and browser automation
wdio-rerun-service
A WebdriverIO service to track and stage for re-running failed or flaky Jasmine/Mocha tests or Cucumber Scenarios.
26.6k+154%-8 months ago
3.0.1
ESM + CommonJSBundledTesting, Scraping and browser automation
@prerenderer/renderer-puppeteer
A renderer for @prerenderer/prerenderer that uses puppeteer to prerender pages.
26.3k---2 years ago
1.2.4
ESM + CommonJSBundledScraping and browser automation
@redhat-developer/locators
Pluggable Page Objects locators for an ExTester framework.
25.9k---11 days ago
1.24.0
CommonJSBundledTesting, Scraping and browser automation
grunt-contrib-qunit
Run QUnit unit tests in a headless Chrome instance
25.8k-35%--
10.2.0
CommonJSNoneScraping and browser automation

12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.

  • @crawlee/linkedomThe scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
  • crawleeThe scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
  • pa11y-ciPa11y CI is a CI-centric accessibility test runner, built using Pa11y
  • google-play-scraperscrapes app data from google play store
  • wreq-jsNode.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
  • @takumi-rs/core-linux-x64-muslRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
  • pi-web-accessWeb search, URL fetching, GitHub repo cloning, PDF extraction, YouTube video understanding, and local video analysis for Pi coding agent. Supports OpenAI, Brave, Parallel, TinyFish, Search1API, Searchinfinity, Querit, Tavily, Firecrawl, Crawl4AI, Jina, SE
  • sitemapperParser for XML Sitemaps to be used with Robots.txt and web crawlers
  • sauce-connect-launcherA library to download and launch Sauce Connect.
  • metascraper-logoMetascraper rule to extract the logo from HTML using Open Graph, JSON-LD, and fallback selectors.
  • dom-minerdom-miner — mine any site into a compact DOM map for QA test plans and AI agents
  • metascraper-descriptionMetascraper rule to extract the description from HTML using Open Graph, JSON-LD, and fallback selectors.
  • metascraper-imageMetascraper rule to extract the image from HTML using Open Graph, JSON-LD, and fallback selectors.
  • metascraper-titleMetascraper rule to extract the title from HTML using Open Graph, JSON-LD, and fallback selectors.
  • @axe-core/webdriverjsProvides a method to inject and analyze web pages using axe
  • sitemapdRuntime-neutral sitemap parsing and bounded traversal
  • metascraper-logo-faviconMetascraper logo fallback that picks favicons and apple-touch icons from HTML.
  • domparser-rsA super fast html parser and manipulator written in rust.
  • @wdio/json-reporterA WebdriverIO plugin to report results in json format.
  • metascraper-urlMetascraper rule to extract the url from HTML using Open Graph, JSON-LD, and fallback selectors.
  • @mcp-b/transportsBrowser transport implementations for Model Context Protocol (MCP) - postMessage, Chrome extension messaging, and iframe communication for AI agents and LLMs
  • domparser-linux-x64-gnuA super fast html parser and manipulator written in rust.
  • selenium-serverSelenium in an npm package
  • rebrowser-puppeteer-coreA drop-in replacement for puppeteer-core patched with rebrowser-patches. It allows to pass modern automation detection tests.
  • dom-parserFast dom parser based on regexps
  • puppeteer-screen-recorderA puppeteer Plugin that uses the native chrome devtool protocol for capturing video frame by frame. Also supports an option to follow pages that are opened by the current page object
  • cloakbrowserStealth Chromium that passes every bot detection test. Drop-in Playwright/Puppeteer replacement with source-level fingerprint patches.
  • domparser-linux-x64-muslA super fast html parser and manipulator written in rust.
  • passmarkThe open-source AI framework for regression testing.
  • @vreden/youtube_scraperA simple YouTube video downloader for audio and video formats with resolusi and quality.
  • eslint-config-sheriffA comprehensive and opinionated TypeScript-first ESLint configuration.
  • filereaderHTML5 FileAPI `FileReader` for Node.JS.
  • @unlighthouse/clientUI Client for Unlighthouse.
  • storycapA Storybook addon, Save the screenshot image of your stories! via puppeteer.
  • @applitools/eyes-playwrightApplitools Eyes SDK for Playwright
  • apifyThe scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
  • @hyperbrowser/sdkNode SDK for Hyperbrowser API
  • isbot-fastJavaScript module detecting bots/crawlers/spiders via user-agent
  • metascraper-publisherMetascraper rule to extract the publisher from HTML using Open Graph, JSON-LD, and fallback selectors.
  • @takumi-rs/core-linux-arm64-gnuRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
  • wdio-intercept-serviceCapture and assert HTTP ajax calls in webdriver.io 🕸
  • @datafast/ai-crawlServer-side AI crawler tracking for DataFast
  • tiktok-live-api-sdkTikTok LIVE API SDK for Node.js & TypeScript. Real-time TikTok LIVE chat, gifts, likes, follows, viewer counts and PK battles via the EulerStream managed TikTok LIVE API, with typed clients for webcast signing, rooms, gifts, rankings, LIVE alerts, moderat
  • vscode-extension-testerExTester is a package that is designed to help you run UI tests for your Visual Studio Code extensions using selenium-webdriver.
  • chrome-aws-lambdaChromium Binary for AWS Lambda and Google Cloud Functions
  • @askjo/camofox-browserHeadless browser automation server and OpenClaw plugin for AI agents - anti-detection, element refs, and session isolation
  • wdio-rerun-serviceA WebdriverIO service to track and stage for re-running failed or flaky Jasmine/Mocha tests or Cucumber Scenarios.
  • @prerenderer/renderer-puppeteerA renderer for @prerenderer/prerenderer that uses puppeteer to prerender pages.
  • @redhat-developer/locatorsPluggable Page Objects locators for an ExTester framework.
  • grunt-contrib-qunitRun QUnit unit tests in a headless Chrome instance