Package category
Scraping and browser automation
Headless browsers, crawlers and HTML extraction.
373 packages7 comparisons
Packages compared
373 packages
| Package | Weekly downloads | 12-month change | 52 weeks | Gzip | Last release | Module | Types | Categories |
|---|---|---|---|---|---|---|---|---|
| fpscanner A lightweight browser fingerprinting and bot detection library with encryption, obfuscation, and cross-context validation | 8.4k | +1372% | - | 1 month ago 1.0.8 | ESM + CommonJS | Bundled | Scraping and browser automation | |
| playwright-i18next-fixture <div align="center"> <br> <header> <img src="https://github.com/cubanducko/playwright-i18next-fixture/blob/main/assets/logo.png?raw=true" height="64" /> </header> <br> <h1>playwright-i18next-fixture</h1> <p> 📝 Use your `i18next` translati | 8.2k | +9% | - | 3 years ago 1.0.0 | CommonJS | Bundled | Internationalisation, Testing | |
| capture-website Capture screenshots of websites | 7.9k | +51% | - | 10 months ago 5.1.0 | ESM only | Bundled | Scraping and browser automation | |
| @applitools/jsdom jsdom without canvas 19.0.0 | 7.6k | - | - | - | 4 years ago 1.0.4 | CommonJS | None | Scraping and browser automation, Node.js utilities |
| @limrun/base-driver Base driver class for Appium drivers | 7.5k | - | - | - | 9 months ago 10.1.2-lim.1 | - | Bundled | Scraping and browser automation, Testing |
| mocha-jsdom Simple integration of jsdom into mocha tests | 7.2k | -26% | - | - 2.0.0 | CommonJS | None | Scraping and browser automation, Testing | |
| playwright-prometheus-remote-write-reporter Playwright prometheus remote write reporter. Send your metrics to prometheus in realtime. | 7k | +387% | - | 9 months ago 0.2.6 | ESM + CommonJS | Bundled | Monitoring and error tracking, Scraping and browser automation | |
| @hyperbrowser/agent Hyperbrowsers Web Agent | 7k | - | - | - | 8 months ago 1.1.2 | CommonJS | Bundled | Scraping and browser automation |
| metascraper-instagram Metascraper rules tailored for Instagram pages — richer metadata than generic HTML parsers. | 6.9k | +1252% | - | 7 days ago 5.58.2 | CommonJS | Bundled | Scraping and browser automation | |
| html-dnd HTML Drag and Drop Simulator for E2E testing | 6.9k | -23% | - | 7 years ago 1.2.1 | CommonJS | Bundled | Testing, DOM and browser utilities | |
| playwright-advanced-har Advanced HAR routing for Playwright | 6.7k | +32% | - | 6 months ago 1.4.1 | CommonJS | Bundled | Scraping and browser automation, Testing | |
| @factory-js/prisma-factory 🏭 The FactoryJS plugin for Prisma | 6.5k | - | - | - | 7 months ago 0.2.3 | ESM + CommonJS | Bundled | Testing, ORMs and query builders |
| playwright-network-cache Cache network requests in Playwright tests | 6.4k | +241% | - | 4 months ago 0.3.0 | CommonJS | Bundled | Caching, Scraping and browser automation | |
| scrape-it A Node.js scraper for humans. | 6.4k | +20% | - | 5 days ago 6.1.15 | CommonJS | Bundled | Scraping and browser automation | |
| @zenrows/browser-sdk ZenRows Scraping Browser JavaScript SDK | 6.3k | - | - | - | 2 years ago 1.1.0 | ESM + CommonJS | Bundled | Scraping and browser automation, Testing |
| puppeteer-report create pdf report with header, footer and page number with puppeteer | 6.2k | +129% | - | 1 year ago 3.2.0 | CommonJS | Bundled | PDF and documents, Scraping and browser automation | |
| @takumi-rs/core-win32-x64-msvc Render OG images from Takumi node trees with native Node.js bindings. No headless browser. | 6.1k | - | - | - | 10 days ago 2.14.0 | CommonJS | None | Cloud SDKs, Scraping and browser automation |
| website-scraper Download website to a local directory (including all css, images, js, etc.) | 6.1k | -30% | - | 10 months ago 6.0.0 | ESM only | None | Scraping and browser automation | |
| olostep Node.js SDK for the Olostep web data API: scrape, crawl, batch, map, search, AI answers, and scheduled monitors. | 6.1k | - | - | 3 months ago 1.2.2 | ESM only | Bundled | Scraping and browser automation | |
| html-table-to-json Extracts all tables within a provided html snippet and converts them to JSON objects. | 6.1k | +114% | - | - 1.0.0 | - | None | Parsers and serialisers, Scraping and browser automation | |
| @bochilteam/scraper Browserless scraper module | 6k | - | - | - | 2 years ago 5.0.1 | ESM + CommonJS | Bundled | Scraping and browser automation |
| unfluff A web page content extractor | 5.9k | +52% | - | 8 years ago 3.2.0 | CommonJS | None | Scraping and browser automation | |
| apify-cli Apify command-line interface (CLI) helps you manage the Apify cloud platform and develop, build, and deploy Apify Actors. | 5.9k | +137% | - | 23 days ago 1.10.0 | ESM only | None | CLI tools and terminal utilities, Scraping and browser automation | |
| playwright-client-certificate-login A playwright script to login using client certificates | 5.8k | -63% | - | - 0.0.3 | CommonJS | None | Authentication and authorisation, Testing | |
| scrape-it-core The core scraping functionality of scrape-it. | 5.7k | +52% | - | 1 year ago 1.0.2 | CommonJS | None | Scraping and browser automation | |
| htmlmetaparser A `htmlparser2` handler for parsing rich metadata from HTML. Includes HTML metadata, JSON-LD, RDFa, microdata, OEmbed, Twitter cards and AppLinks. | 5.7k | +39% | - | 2 years ago 2.1.3 | CommonJS | Bundled | Scraping and browser automation | |
| rezo Lightning-fast, enterprise-grade HTTP client for modern JavaScript. Full HTTP/2 support, intelligent cookie management, multiple adapters (HTTP, Fetch, cURL, XHR), streaming, proxy support (HTTP/HTTPS/SOCKS), and cross-environment compatibility. | 5.6k | - | - | 2 months ago 1.0.139 | ESM + CommonJS | Bundled | React, Scraping and browser automation | |
| playwright-recast Fluent pipeline library for processing Playwright traces into polished demo videos — TTS voiceover, subtitles, speed control, and zoom. | 5.6k | - | - | 28 days ago 0.21.0 | ESM only | Bundled | Testing, Video and audio | |
| domparser-win32-x64-msvc A super fast html parser and manipulator written in rust. | 5.6k | - | - | 5 months ago 0.1.1 | CommonJS | None | Parsers and serialisers, Scraping and browser automation | |
| @crawlee/impit-client impit-based HTTP client implementation for Crawlee. Impersonates browser requests to avoid bot detection. | 5.5k | - | - | - | 1 month ago 3.18.1 | ESM + CommonJS | Bundled | Scraping and browser automation |
| fiftyone.devicedetection.onpremise Device detection on-premise services for the 51Degrees Pipeline API | 5.4k | +481% | - | 7 days ago 4.5.87 | CommonJS | Bundled | Scraping and browser automation | |
| fiftyone.devicedetection.shared Shared utilities and base functionality for implementing device detection engines for the 51Degrees Pipeline API in Node.js. | 5.4k | +375% | - | 7 days ago 4.5.87 | CommonJS | Bundled | Scraping and browser automation | |
| fiftyone.devicedetection.cloud Device detection cloud services for the 51Degrees Pipeline API | 5.4k | +413% | - | 7 days ago 4.5.87 | CommonJS | Bundled | Scraping and browser automation | |
| puppeteer-autoscroll-down Handle infinite scroll on websites with puppeteer | 5.3k | +34% | - | 1 year ago 2.0.1 | ESM only | Bundled | Scraping and browser automation, Parsers and serialisers | |
| sitemap-generator Easily create XML sitemaps for your website. | 5.3k | +21% | - | 6 years ago 8.5.1 | CommonJS | None | Scraping and browser automation | |
| fiftyone.devicedetection Parse HTTP headers to detect the device type, model, operating system, browser, and crawler information | 5.3k | +519% | - | 7 days ago 4.5.87 | CommonJS | None | Scraping and browser automation | |
| @scenarist/playwright-helpers Playwright test helpers for Scenarist scenario management | 5.3k | - | - | - | 1 month ago 0.4.14 | ESM only | Bundled | Testing, Scraping and browser automation |
| scraperapi-sdk Node.js SDK for ScraperAPI.com | 5.2k | -15% | - | 2 years ago 2.0.1 | CommonJS | None | Scraping and browser automation | |
| auto-playwright Automate Playwright tests using ChatGPT. | 5.1k | +172% | - | 1 year ago 1.16.1 | CommonJS | None | Scraping and browser automation, Testing | |
| @vivliostyle/jsdom A JavaScript implementation of many web standards | 5.1k | - | - | - | 4 months ago 25.0.1-vivliostyle-cli.2 | CommonJS | None | Scraping and browser automation, Node.js utilities |
| google-news-url-decoder A Node.js library to decode Google News URLs to their original source URLs. | 4.9k | - | - | 3 months ago 1.2.2 | CommonJS | None | URLs and query strings, Scraping and browser automation | |
| betterwright A persistent, policy-guarded Playwright browser for AI agents with network controls, trusted credential filling, proof screenshots, and CAPTCHA helpers. | 4.8k | - | - | 4 days ago 2.8.8 | ESM only | Bundled | Scraping and browser automation, Testing | |
| top-user-agents Always up-to-date list of the top 100 most common browser user-agents for HTTP clients. | 4.8k | +52% | - | 4 days ago 2.1.137 | CommonJS | Bundled | Scraping and browser automation, DOM and browser utilities | |
| puppeteer-capture A Puppeteer plugin for capturing page as a video with ultimate quality. | 4.7k | +21228% | - | 1 month ago 1.58.0 | CommonJS | Bundled | Scraping and browser automation, Video and audio | |
| @wreq-js/binding-darwin-arm64 Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible. | 4.7k | - | - | - | 1 month ago 3.2.0 | CommonJS | None | HTTP clients, Scraping and browser automation |
| es6-crawler-detect This is an ES6 adaptation of the original PHP library CrawlerDetect, this library will help you detect bots/crawlers/spiders vie the useragent. | 4.4k | -44% | - | 1 year ago 4.0.2 | CommonJS | Bundled | Scraping and browser automation | |
| node-wreq HTTP client with native TLS, HTTP2, JA3, JA4 browser impersonation backed by wreq's Rust core | 4.4k | +6572% | - | 18 days ago 3.2.1 | ESM + CommonJS | Bundled | Scraping and browser automation, HTTP clients | |
| @firecrawl/firecrawl-convex Firecrawl component for Convex: scrape, map, and search the web, and run durable crawls with reactive progress. | 4.3k | - | - | - | 1 month ago 0.1.1 | ESM only | Bundled | Scraping and browser automation |
| @scrapeless-ai/sdk Node SDK for Scrapeless AI | 4.3k | - | - | - | 9 days ago 1.12.1 | ESM + CommonJS | Bundled | Scraping and browser automation |
| googlethis A simple yet powerful module to retrieve organic search results and much more from Google. | 4.2k | -53% | - | 3 years ago 1.8.0 | CommonJS | Bundled | Maps and geolocation, Scraping and browser automation |
12-month change compares the average of the last 4 weeks of downloads with the first 4 weeks of the 52-week series. Gzip size is for the whole package, as measured by Bundlephobia. "-" means the value has not been fetched.
- fpscannerA lightweight browser fingerprinting and bot detection library with encryption, obfuscation, and cross-context validation
- playwright-i18next-fixture<div align="center"> <br> <header> <img src="https://github.com/cubanducko/playwright-i18next-fixture/blob/main/assets/logo.png?raw=true" height="64" /> </header> <br> <h1>playwright-i18next-fixture</h1> <p> 📝 Use your `i18next` translati
- capture-websiteCapture screenshots of websites
- @applitools/jsdomjsdom without canvas 19.0.0
- @limrun/base-driverBase driver class for Appium drivers
- mocha-jsdomSimple integration of jsdom into mocha tests
- playwright-prometheus-remote-write-reporterPlaywright prometheus remote write reporter. Send your metrics to prometheus in realtime.
- @hyperbrowser/agentHyperbrowsers Web Agent
- metascraper-instagramMetascraper rules tailored for Instagram pages — richer metadata than generic HTML parsers.
- html-dndHTML Drag and Drop Simulator for E2E testing
- playwright-advanced-harAdvanced HAR routing for Playwright
- @factory-js/prisma-factory🏭 The FactoryJS plugin for Prisma
- playwright-network-cacheCache network requests in Playwright tests
- scrape-itA Node.js scraper for humans.
- @zenrows/browser-sdkZenRows Scraping Browser JavaScript SDK
- puppeteer-reportcreate pdf report with header, footer and page number with puppeteer
- @takumi-rs/core-win32-x64-msvcRender OG images from Takumi node trees with native Node.js bindings. No headless browser.
- website-scraperDownload website to a local directory (including all css, images, js, etc.)
- olostepNode.js SDK for the Olostep web data API: scrape, crawl, batch, map, search, AI answers, and scheduled monitors.
- html-table-to-jsonExtracts all tables within a provided html snippet and converts them to JSON objects.
- @bochilteam/scraperBrowserless scraper module
- unfluffA web page content extractor
- apify-cliApify command-line interface (CLI) helps you manage the Apify cloud platform and develop, build, and deploy Apify Actors.
- playwright-client-certificate-loginA playwright script to login using client certificates
- scrape-it-coreThe core scraping functionality of scrape-it.
- htmlmetaparserA `htmlparser2` handler for parsing rich metadata from HTML. Includes HTML metadata, JSON-LD, RDFa, microdata, OEmbed, Twitter cards and AppLinks.
- rezoLightning-fast, enterprise-grade HTTP client for modern JavaScript. Full HTTP/2 support, intelligent cookie management, multiple adapters (HTTP, Fetch, cURL, XHR), streaming, proxy support (HTTP/HTTPS/SOCKS), and cross-environment compatibility.
- playwright-recastFluent pipeline library for processing Playwright traces into polished demo videos — TTS voiceover, subtitles, speed control, and zoom.
- domparser-win32-x64-msvcA super fast html parser and manipulator written in rust.
- @crawlee/impit-clientimpit-based HTTP client implementation for Crawlee. Impersonates browser requests to avoid bot detection.
- fiftyone.devicedetection.onpremiseDevice detection on-premise services for the 51Degrees Pipeline API
- fiftyone.devicedetection.sharedShared utilities and base functionality for implementing device detection engines for the 51Degrees Pipeline API in Node.js.
- fiftyone.devicedetection.cloudDevice detection cloud services for the 51Degrees Pipeline API
- puppeteer-autoscroll-downHandle infinite scroll on websites with puppeteer
- sitemap-generatorEasily create XML sitemaps for your website.
- fiftyone.devicedetectionParse HTTP headers to detect the device type, model, operating system, browser, and crawler information
- @scenarist/playwright-helpersPlaywright test helpers for Scenarist scenario management
- scraperapi-sdkNode.js SDK for ScraperAPI.com
- auto-playwrightAutomate Playwright tests using ChatGPT.
- @vivliostyle/jsdomA JavaScript implementation of many web standards
- google-news-url-decoderA Node.js library to decode Google News URLs to their original source URLs.
- betterwrightA persistent, policy-guarded Playwright browser for AI agents with network controls, trusted credential filling, proof screenshots, and CAPTCHA helpers.
- top-user-agentsAlways up-to-date list of the top 100 most common browser user-agents for HTTP clients.
- puppeteer-captureA Puppeteer plugin for capturing page as a video with ultimate quality.
- @wreq-js/binding-darwin-arm64Node.js/TypeScript HTTP client with browser TLS fingerprint impersonation (JA3/JA4). Bypass Cloudflare and anti-bot detection. Rust-powered, fetch()-compatible.
- es6-crawler-detectThis is an ES6 adaptation of the original PHP library CrawlerDetect, this library will help you detect bots/crawlers/spiders vie the useragent.
- node-wreqHTTP client with native TLS, HTTP2, JA3, JA4 browser impersonation backed by wreq's Rust core
- @firecrawl/firecrawl-convexFirecrawl component for Convex: scrape, map, and search the web, and run durable crawls with reactive progress.
- @scrapeless-ai/sdkNode SDK for Scrapeless AI
- googlethisA simple yet powerful module to retrieve organic search results and much more from Google.