
Crawlee
apify/crawleeWeb scraping and browser automation for Node.js
PersonalEnterpriseAutomationLibrary / SDKPermissive
About
Build reliable crawlers with anti-blocking and storage built in.
From the repository: “Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.”
Alternatives
All alternatives →Browser UseLet AI agents use a web browserLibrary / SDK89
ArduPilotAutopilot for drones, planes and roversLibrary / SDK86
PuppeteerControl Chrome and Firefox from JavaScriptLibrary / SDK85
ScrapyFast, high-level web crawling framework for PythonLibrary / SDK85