Crawl4AI

Crawl4AI project · Tools and frameworks · json

User-agent: Crawl4AI
Disallow: /
robots.txt tokenCrawl4AI
User-agent containsCrawl4AI
OperatorCrawl4AI project
CategoryTools and frameworks
robots.txtoperator publishes no robots.txt statement
Verify byno published verification method

What it is

An open-source LLM-oriented crawler and scraper library, run by whoever installs it. Like Scrapy, the default user-agent identifies the software and says nothing about who is behind the request.

What blocking it costs you

You block a library, not an operator: the rule catches a researcher and a bulk scraper equally, and anyone who edits one config line is not caught at all.

Full user-agent string

Crawl4AI

Allow it instead

User-agent: Crawl4AI
Allow: /

Operator documentation: https://github.com/unclecode/crawl4ai
Machine copies: json · markdown
Policies that name this crawler: allow-all · maximum-ai-visibility