Skip to content
JS
Package category

Scraping and browser automation

Headless browsers, crawlers and HTML extraction.

373 packages7 comparisons

Packages compared

373 packages
PackageWeekly downloads12-month change52 weeksGzipLast releaseModuleTypesCategories
fpscanner
A lightweight browser fingerprinting and bot detection library with encryption, obfuscation, and cross-context validation
8.4k+1372%-1 month ago
1.0.8
ESM + CommonJSBundledScraping and browser automation
playwright-i18next-fixture
<div align="center"> <br> <header> <img src="https://github.com/cubanducko/playwright-i18next-fixture/blob/main/assets/logo.png?raw=true" height="64" /> </header> <br> <h1>playwright-i18next-fixture</h1> <p> 📝 Use your `i18next` translati
8.2k+9%-3 years ago
1.0.0
CommonJSBundledInternationalisation, Testing
capture-website
Capture screenshots of websites
7.9k+51%-10 months ago
5.1.0
ESM onlyBundledScraping and browser automation
@applitools/jsdom
jsdom without canvas 19.0.0
7.6k---4 years ago
1.0.4
CommonJSNoneScraping and browser automation, Node.js utilities
@limrun/base-driver
Base driver class for Appium drivers
7.5k---9 months ago
10.1.2-lim.1
-BundledScraping and browser automation, Testing
mocha-jsdom
Simple integration of jsdom into mocha tests
7.2k-26%--
2.0.0
CommonJSNoneScraping and browser automation, Testing
playwright-prometheus-remote-write-reporter
Playwright prometheus remote write reporter. Send your metrics to prometheus in realtime.
7k+387%-9 months ago
0.2.6
ESM + CommonJSBundledMonitoring and error tracking, Scraping and browser automation
@hyperbrowser/agent
Hyperbrowsers Web Agent
7k---8 months ago
1.1.2
CommonJSBundledScraping and browser automation
metascraper-instagram
Metascraper rules tailored for Instagram pages — richer metadata than generic HTML parsers.
6.9k+1252%-7 days ago
5.58.2
CommonJSBundledScraping and browser automation
html-dnd
HTML Drag and Drop Simulator for E2E testing
6.9k-23%-7 years ago
1.2.1
CommonJSBundledTesting, DOM and browser utilities
playwright-advanced-har
Advanced HAR routing for Playwright
6.7k+32%-6 months ago
1.4.1
CommonJSBundledScraping and browser automation, Testing
@factory-js/prisma-factory
🏭 The FactoryJS plugin for Prisma
6.5k---7 months ago
0.2.3
ESM + CommonJSBundledTesting, ORMs and query builders
playwright-network-cache
Cache network requests in Playwright tests
6.4k+241%-4 months ago
0.3.0
CommonJSBundledCaching, Scraping and browser automation
scrape-it
A Node.js scraper for humans.
6.4k+20%-5 days ago
6.1.15
CommonJSBundledScraping and browser automation
@zenrows/browser-sdk
ZenRows Scraping Browser JavaScript SDK
6.3k---2 years ago
1.1.0
ESM + CommonJSBundledScraping and browser automation, Testing
puppeteer-report
create pdf report with header, footer and page number with puppeteer
6.2k+129%-1 year ago
3.2.0
CommonJSBundledPDF and documents, Scraping and browser automation
@takumi-rs/core-win32-x64-msvc
Render OG images from Takumi node trees with native Node.js bindings. No headless browser.
6.1k---10 days ago
2.14.0
CommonJSNoneCloud SDKs, Scraping and browser automation
website-scraper
Download website to a local directory (including all css, images, js, etc.)
6.1k-30%-10 months ago
6.0.0
ESM onlyNoneScraping and browser automation
olostep
Node.js SDK for the Olostep web data API: scrape, crawl, batch, map, search, AI answers, and scheduled monitors.
6.1k--3 months ago
1.2.2
ESM onlyBundledScraping and browser automation
html-table-to-json
Extracts all tables within a provided html snippet and converts them to JSON objects.
6.1k+114%--
1.0.0
-NoneParsers and serialisers, Scraping and browser automation
@bochilteam/scraper
Browserless scraper module
6k---2 years ago
5.0.1
ESM + CommonJSBundledScraping and browser automation
unfluff
A web page content extractor
5.9k+52%-8 years ago
3.2.0
CommonJSNoneScraping and browser automation
apify-cli
Apify command-line interface (CLI) helps you manage the Apify cloud platform and develop, build, and deploy Apify Actors.
5.9k+137%-23 days ago
1.10.0
ESM onlyNoneCLI tools and terminal utilities, Scraping and browser automation
playwright-client-certificate-login
A playwright script to login using client certificates
5.8k-63%--
0.0.3
CommonJSNoneAuthentication and authorisation, Testing
scrape-it-core
The core scraping functionality of scrape-it.
5.7k+52%-1 year ago
1.0.2
CommonJSNoneScraping and browser automation
htmlmetaparser
A `htmlparser2` handler for parsing rich metadata from HTML. Includes HTML metadata, JSON-LD, RDFa, microdata, OEmbed, Twitter cards and AppLinks.
5.7k+39%-2 years ago
2.1.3
CommonJSBundledScraping and browser automation
rezo
Lightning-fast, enterprise-grade HTTP client for modern JavaScript. Full HTTP/2 support, intelligent cookie management, multiple adapters (HTTP, Fetch, cURL, XHR), streaming, proxy support (HTTP/HTTPS/SOCKS), and cross-environment compatibility.
5.6k--2 months ago
1.0.139
ESM + CommonJSBundledReact, Scraping and browser automation
playwright-recast
Fluent pipeline library for processing Playwright traces into polished demo videos — TTS voiceover, subtitles, speed control, and zoom.
5.6k--28 days ago
0.21.0
ESM onlyBundledTesting, Video and audio
domparser-win32-x64-msvc
A super fast html parser and manipulator written in rust.
5.6k--5 months ago
0.1.1
CommonJSNoneParsers and serialisers, Scraping and browser automation
@crawlee/impit-client
impit-based HTTP client implementation for Crawlee. Impersonates browser requests to avoid bot detection.
5.5k---1 month ago
3.18.1
ESM + CommonJSBundledScraping and browser automation
fiftyone.devicedetection.onpremise
Device detection on-premise services for the 51Degrees Pipeline API
5.4k+481%-7 days ago
4.5.87
CommonJSBundledScraping and browser automation
fiftyone.devicedetection.shared
Shared utilities and base functionality for implementing device detection engines for the 51Degrees Pipeline API in Node.js.
5.4k+375%-7 days ago
4.5.87
CommonJSBundledScraping and browser automation
fiftyone.devicedetection.cloud
Device detection cloud services for the 51Degrees Pipeline API
5.4k+413%-7 days ago
4.5.87
CommonJSBundledScraping and browser automation
puppeteer-autoscroll-down
Handle infinite scroll on websites with puppeteer
5.3k+34%-1 year ago
2.0.1
ESM onlyBundledScraping and browser automation, Parsers and serialisers
sitemap-generator
Easily create XML sitemaps for your website.
5.3k+21%-6 years ago
8.5.1
CommonJSNoneScraping and browser automation
fiftyone.devicedetection
Parse HTTP headers to detect the device type, model, operating system, browser, and crawler information
5.3k+519%-7 days ago
4.5.87
CommonJSNoneScraping and browser automation
@scenarist/playwright-helpers
Playwright test helpers for Scenarist scenario management
5.3k---1 month ago
0.4.14
ESM onlyBundledTesting, Scraping and browser automation
scraperapi-sdk
Node.js SDK for ScraperAPI.com
5.2k-15%-2 years ago
2.0.1
CommonJSNoneScraping and browser automation
auto-playwright
Automate Playwright tests using ChatGPT.
5.1k+172%-1 year ago
1.16.1
CommonJSNoneScraping and browser automation, Testing
@vivliostyle/jsdom
A JavaScript implementation of many web standards
5.1k---4 months ago
25.0.1-vivliostyle-cli.2
CommonJSNoneScraping and browser automation, Node.js utilities
google-news-url-decoder
A Node.js library to decode Google News URLs to their original source URLs.
4.9k--3 months ago
1.2.2
CommonJSNoneURLs and query strings, Scraping and browser automation
betterwright
A persistent, policy-guarded Playwright browser for AI agents with network controls, trusted credential filling, proof screenshots, and CAPTCHA helpers.
4.8k--4 days ago
2.8.8
ESM onlyBundledScraping and browser automation, Testing
top-user-agents
Always up-to-date list of the top 100 most common browser user-agents for HTTP clients.
4.8k+52%-4 days ago
2.1.137
CommonJSBundledScraping and browser automation, DOM and browser utilities
puppeteer-capture
A Puppeteer plugin for capturing page as a video with ultimate quality.
4.7k+21228%-1 month ago
1.58.0
CommonJSBundledScraping and browser automation, Video and audio
@wreq-js/binding-darwin-arm64
Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
4.7k---1 month ago
3.2.0
CommonJSNoneHTTP clients, Scraping and browser automation
es6-crawler-detect
This is an ES6 adaptation of the original PHP library CrawlerDetect, this library will help you detect bots/crawlers/spiders vie the useragent.
4.4k-44%-1 year ago
4.0.2
CommonJSBundledScraping and browser automation
node-wreq
HTTP client with native TLS, HTTP2, JA3, JA4 browser impersonation backed by wreq's Rust core
4.4k+6572%-18 days ago
3.2.1
ESM + CommonJSBundledScraping and browser automation, HTTP clients
@firecrawl/firecrawl-convex
Firecrawl component for Convex: scrape, map, and search the web, and run durable crawls with reactive progress.
4.3k---1 month ago
0.1.1
ESM onlyBundledScraping and browser automation
@scrapeless-ai/sdk
Node SDK for Scrapeless AI
4.3k---9 days ago
1.12.1
ESM + CommonJSBundledScraping and browser automation
googlethis
A simple yet powerful module to retrieve organic search results and much more from Google.
4.2k-53%-3 years ago
1.8.0
CommonJSBundledMaps and geolocation, Scraping and browser automation

12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.

  • fpscannerA lightweight browser fingerprinting and bot detection library with encryption, obfuscation, and cross-context validation
  • playwright-i18next-fixture<div align="center"> <br> <header> <img src="https://github.com/cubanducko/playwright-i18next-fixture/blob/main/assets/logo.png?raw=true" height="64" /> </header> <br> <h1>playwright-i18next-fixture</h1> <p> 📝 Use your `i18next` translati
  • capture-websiteCapture screenshots of websites
  • @applitools/jsdomjsdom without canvas 19.0.0
  • @limrun/base-driverBase driver class for Appium drivers
  • mocha-jsdomSimple integration of jsdom into mocha tests
  • playwright-prometheus-remote-write-reporterPlaywright prometheus remote write reporter. Send your metrics to prometheus in realtime.
  • @hyperbrowser/agentHyperbrowsers Web Agent
  • metascraper-instagramMetascraper rules tailored for Instagram pages — richer metadata than generic HTML parsers.
  • html-dndHTML Drag and Drop Simulator for E2E testing
  • playwright-advanced-harAdvanced HAR routing for Playwright
  • @factory-js/prisma-factory🏭 The FactoryJS plugin for Prisma
  • playwright-network-cacheCache network requests in Playwright tests
  • scrape-itA Node.js scraper for humans.
  • @zenrows/browser-sdkZenRows Scraping Browser JavaScript SDK
  • puppeteer-reportcreate pdf report with header, footer and page number with puppeteer
  • @takumi-rs/core-win32-x64-msvcRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
  • website-scraperDownload website to a local directory (including all css, images, js, etc.)
  • olostepNode.js SDK for the Olostep web data API: scrape, crawl, batch, map, search, AI answers, and scheduled monitors.
  • html-table-to-jsonExtracts all tables within a provided html snippet and converts them to JSON objects.
  • @bochilteam/scraperBrowserless scraper module
  • unfluffA web page content extractor
  • apify-cliApify command-line interface (CLI) helps you manage the Apify cloud platform and develop, build, and deploy Apify Actors.
  • playwright-client-certificate-loginA playwright script to login using client certificates
  • scrape-it-coreThe core scraping functionality of scrape-it.
  • htmlmetaparserA `htmlparser2` handler for parsing rich metadata from HTML. Includes HTML metadata, JSON-LD, RDFa, microdata, OEmbed, Twitter cards and AppLinks.
  • rezoLightning-fast, enterprise-grade HTTP client for modern JavaScript. Full HTTP/2 support, intelligent cookie management, multiple adapters (HTTP, Fetch, cURL, XHR), streaming, proxy support (HTTP/HTTPS/SOCKS), and cross-environment compatibility.
  • playwright-recastFluent pipeline library for processing Playwright traces into polished demo videos — TTS voiceover, subtitles, speed control, and zoom.
  • domparser-win32-x64-msvcA super fast html parser and manipulator written in rust.
  • @crawlee/impit-clientimpit-based HTTP client implementation for Crawlee. Impersonates browser requests to avoid bot detection.
  • fiftyone.devicedetection.onpremiseDevice detection on-premise services for the 51Degrees Pipeline API
  • fiftyone.devicedetection.sharedShared utilities and base functionality for implementing device detection engines for the 51Degrees Pipeline API in Node.js.
  • fiftyone.devicedetection.cloudDevice detection cloud services for the 51Degrees Pipeline API
  • puppeteer-autoscroll-downHandle infinite scroll on websites with puppeteer
  • sitemap-generatorEasily create XML sitemaps for your website.
  • fiftyone.devicedetectionParse HTTP headers to detect the device type, model, operating system, browser, and crawler information
  • @scenarist/playwright-helpersPlaywright test helpers for Scenarist scenario management
  • scraperapi-sdkNode.js SDK for ScraperAPI.com
  • auto-playwrightAutomate Playwright tests using ChatGPT.
  • @vivliostyle/jsdomA JavaScript implementation of many web standards
  • google-news-url-decoderA Node.js library to decode Google News URLs to their original source URLs.
  • betterwrightA persistent, policy-guarded Playwright browser for AI agents with network controls, trusted credential filling, proof screenshots, and CAPTCHA helpers.
  • top-user-agentsAlways up-to-date list of the top 100 most common browser user-agents for HTTP clients.
  • puppeteer-captureA Puppeteer plugin for capturing page as a video with ultimate quality.
  • @wreq-js/binding-darwin-arm64Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
  • es6-crawler-detectThis is an ES6 adaptation of the original PHP library CrawlerDetect, this library will help you detect bots/crawlers/spiders vie the useragent.
  • node-wreqHTTP client with native TLS, HTTP2, JA3, JA4 browser impersonation backed by wreq's Rust core
  • @firecrawl/firecrawl-convexFirecrawl component for Convex: scrape, map, and search the web, and run durable crawls with reactive progress.
  • @scrapeless-ai/sdkNode SDK for Scrapeless AI
  • googlethisA simple yet powerful module to retrieve organic search results and much more from Google.