# Crawl4AI

> Web crawler that turns sites into LLM-ready Markdown

An open-source crawler and scraper built for AI: clean Markdown output, structured extraction and browser control.

- Source: https://github.com/unclecode/crawl4ai
- Homepage: https://crawl4ai.com
- License: Apache-2.0
- Language: Python
- Stars: 84658
- Forks: 8759
- Contributors: 89
- Last commit: 2026-09-25
- Latest release: v0.9.4 (2026-09-23)
- Purpose: AI & LLM tooling, RAG & vector DBs
- Runs on: Docker / self-host, Library / SDK
- For: Personal, Enterprise

## Worth score: 88/100

How much DigGitHub recommends it, from activity, adoption, docs, license and security signals.

- Popularity: 25/25
- Momentum: 10/20
- Maintenance: 25/25
- Community: 13/15
- Readiness: 15/15

## Alternatives

- [LightRAG](https://diggithub.com/HKUDS/LightRAG.md): Simple and fast graph-based RAG
- [gpt-researcher](https://diggithub.com/assafelovic/gpt-researcher.md): Autonomous deep research agent for web and local documents
- [Chroma](https://diggithub.com/chroma-core/chroma.md): Open-source embedding database for AI apps
- [Firecrawl](https://diggithub.com/firecrawl/firecrawl.md): Web data API that turns sites into LLM-ready data
- [RAGFlow](https://diggithub.com/infiniflow/ragflow.md): RAG engine built on deep document understanding
- [MarkItDown](https://diggithub.com/microsoft/markitdown.md): Convert Office files and PDFs to Markdown for LLMs
- [Docling](https://diggithub.com/docling-project/docling.md): Get your documents ready for generative AI
- [OpenViking](https://diggithub.com/volcengine/OpenViking.md): Context database for AI agents

---

Source page: https://diggithub.com/unclecode/crawl4ai
Updated: 2026-10-03
