Crawl4ai Skill — Data skill for Claude Code
Scrape JavaScript-heavy sites and extract structured data via reusable CSS schemas.
How to install Crawl4ai Skill
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open brettdavies/crawl4ai-skill and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Crawl4ai Skill does
Scrape JavaScript-heavy sites and extract structured data via reusable CSS schemas. Portable agent skill wrapping the Crawl4AI CLI and Python SDK.
Alternatives in Data
- HTML Video — Programmatic video for coding agents — HTML to video on your laptop 4.5k ★
- n8n Code JavaScript — JavaScript in n8n Code nodes with data access patterns 3.6k ★
- Skill Content Pipeline — Extract patterns and anatomy from URLs — use to reverse-engineer content strategies from live pages 2.8k ★
README
Crawl4AI Agent Skill
Scrape JavaScript-heavy sites and extract structured data via reusable CSS schemas. A portable agent skill that wraps the [Crawl4AI](https://crawl4ai.com/) CLI and Python SDK, written in the Anthropic SKILL.md format and consumable by any agent host that loads SKILL.md-format bundles (Claude Code, Codex, Cursor, OpenCode, Cline, and others).
Verified against Crawl4AI library version `0.8.9` (pinned in [`VERSION`](VERSION)).
Features
- JS-aware crawling: full headless-browser rendering with
wait_until=networkidledefaults - Schema-based extraction: derive a CSS selector schema once via LLM, apply it forever with no further LLM cost
- LLM extraction: per-request structured extraction when a schema is not worth deriving
- Content filtering: BM25 relevance filter and quality-based pruning, plain markdown or markdown-fit output
- Concurrent batch crawling: multi-URL processing with per-job concurrency caps
- Session management: persistent sessions for authenticated, multi-step flows
- CLI and SDK: both the
crwlcommand-line tool and thecrawl4aiPython SDK
Installation
Clone the repo into the skills directory your agent host loads from:
# Claude Code
git clone https://github.com/brettdavies/crawl4ai-skill.git ~/.claude/skills/crawl4ai
For other agent hosts (Codex, Cursor, OpenCode, Cline, custom agents), clone into whichever directory your host scans for SKILL.md-format bundles. Refer to your host's documentation for the skills directory location. The bundle root contains `SKILL.md`, so the skill registers automatically once the directory is on the host's skills search path.
Prerequisites
The skill calls into the Crawl4AI Python library, which must be installed in the runtime your agent uses:
pip install crawl4ai
crawl4ai-setup
crawl4ai-doctor
`crawl4ai-doctor` validates the install and confirms a headless browser is available.
Quick start
CLI:
crwl https:/
Related Skills
Figma UI MCP
AI can draw UI directly on the Figma canvas via JavaScript, and read existing designs back as structured data.
Ecommerce SEO Geo Skills
Portable skills for ecommerce SEO, GEO readiness, structured data, product-listing optimization, keyword gaps,
Agentic Semgrep Rules
Semgrep rules for AI agent code: static analysis for LLM applications in Python, TypeScript and JavaScript. Fi
Busabase
Open-source database & workspace for AI agents — structured data, durable knowledge, reusable skills, runnable
Anthropic Retrieval Demo
Lightweight demo using the Anthropic Python SDK to experiment with Claude's Search and Retrieval capabilities
Craft Collection
A Claude Code plugin marketplace that codifies engineering craft: disciplined Python and data-engineering prac
Related Agents
Curator
(/auto) Maintain active research-tree memory: absorb reviewed evidence, keep graph/state/provenance coherent,
Astro Specialist
Astro framework specialist — SSG sites with content collections, islands architecture, astro:assets, i18n rout
Widget Architect
Use this agent for designing reusable widget systems, reviewing widget composition, deciding when to extract a