Crawlee

apify/crawlee

Web scraping and browser automation for Node.js

PersonalEnterpriseAutomationLibrary / SDKPermissive
crawlee.dev
Preview of Crawlee
Preview of Crawlee at crawlee.dev

About

Build reliable crawlers with anti-blocking and storage built in.

From the repository: “Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.”

Browser UseLet AI agents use a web browserLibrary / SDK89
ArduPilotAutopilot for drones, planes and roversLibrary / SDK86
PuppeteerControl Chrome and Firefox from JavaScriptLibrary / SDK85
ScrapyFast, high-level web crawling framework for PythonLibrary / SDK85

Topics