AgentBlocking it costs you visibility
FirecrawlAgent
Extraction layer used by thousands of AI apps to read your pages on demand.
- robots.txt token
FirecrawlAgent- Operator
- Firecrawl ↗
- Powers
- custom AI apps and agents built on Firecrawl
- Honours robots.txt
- Yes, documented
Allow it
User-agent: FirecrawlAgent
Allow: /Block it
User-agent: FirecrawlAgent
Disallow: /Need the whole file, with every crawler and your private paths — and the llms.txt to go with it? The generator writes them. AI config generator →
User-Agent
Mozilla/5.0 (compatible; FirecrawlAgent/1.0; +https://firecrawl.dev) Crawlable/1.0; +https://crawlable.fr/probeThis is the string we send when probing. The Crawlable suffix identifies us in your logs.
Check your own site
robots.txt is only half the answer: your CDN can refuse FirecrawlAgent before it ever reads the file. A scan checks both.
Free, no account, about 5 seconds.