Skip to content
ai-supply.store
探索分类排行榜社区Agent APIFAQ
登录免费注册
catalog / Data & ETL / Crawl4AI
⇄ConnectorData & ETLFree

Crawl4AI

LLM-friendly open-source web crawler that turns pages into clean Markdown/JSON ready for RAG and agent pipelines.

@ai-supply
安装量42k
↗ 源代码仓库

Crawl4AI

Crawl4AI is an open-source, LLM-friendly web crawler and scraper that converts web pages into clean, structured Markdown or JSON ready to feed into RAG and agent pipelines. It is one of the most popular data-ingestion connectors for AI applications, designed for speed and for output that models can consume directly.

Key features

  • Fast async crawling built on Playwright with browser session reuse
  • LLM-ready Markdown generation with content filtering and pruning
  • CSS/XPath selectors plus LLM-based structured extraction strategies
  • Handles JavaScript-rendered pages, lazy loading, and stealth/anti-bot options
  • Python API and a Docker deployment for scale

Usage note: pip install crawl4ai, run the post-install browser setup, then use the async crawler to fetch a URL and receive cleaned Markdown plus extracted structured data.

Curated mirror of the open-source Crawl4AI (Apache-2.0). Get it from the source.

More from @ai-supply

View profile →
◇MCP server
GitHub MCP Server
Official GitHub MCP server — give your AI agent full read/write access to repos, issues, PRs, and actions.
↓ 771k
⠿Embedding
Sentence Transformers
State-of-the-art sentence and text embeddings — compute semantic similarity, clustering, and dense retrieval.
↓ 751k
◆Skill
NLTK
The Natural Language Toolkit — Python's foundational NLP library for tokenization, POS tagging, parsing, and corpora.
↓ 641k
◇MCP server
MCP TypeScript SDK
Official TypeScript/JavaScript SDK for building MCP servers and clients — the Node.js foundation for the Model Context Protocol.
↓ 629k
ai-supply.store

免费、经过安全审核的 AI 能力——技能、MCP、插件、agent、数据集等一应俱全,每一项都经过安全评级与时效追踪,为人类与 agent 共同打造。

api · v3.1status · all green
联系
support@ai-supply.storesecurity@ai-supply.store
目录
  • 探索
  • 分类
  • 排行榜
  • 基准测试
  • 安全
社区
  • 社区
  • FAQ
面向智能体
  • 快速入门 (60s)
  • 授权智能体
  • Agent API
  • OpenAPI 规范
面向开发者
  • 发布
  • 控制台
账户
  • 创建账户
  • 登录
  • 设置
法律条款
  • 条款
  • 发布者协议
  • 可接受使用政策
  • 隐私政策