Agent · Firecrawl
FirecrawlAgent
A hosted scraping API that turns pages into model-ready markdown for developers building AI agents. Each hit is one developer's job, so volume varies widely by site.
- Operator
- Firecrawl
- Purpose
- Agentic browsers
- Feeds
- Firecrawl (agent scraping)
- robots.txt token
FirecrawlAgentAlso:Firecrawl- Honours robots.txt
- Undocumented
- Documentation
- None published
What it looks like in server logs
A representative user-agent string. Version numbers change; the FirecrawlAgent token does not. Paste your own log line into the user-agent checker to confirm a match.
Mozilla/5.0 (compatible; FirecrawlAgent/1.0)How to block FirecrawlAgent
Add this to the robots.txt at the root of your domain. Robots.txt is advisory: it works only on crawlers that choose to read it. This one is undocumented. To enforce a block, match the user agent at your CDN or web server and return 403.
User-agent: FirecrawlAgent
User-agent: Firecrawl
Disallow: /
To block every agentic browser at once, or to block training while keeping AI search citations, use the robots.txt generator.
How to allow FirecrawlAgent while blocking others
A more specific User-agent group wins over a wildcard. Place this above any User-agent: * block that disallows paths, and FirecrawlAgent keeps full access.
User-agent: FirecrawlAgent
Allow: /
How to verify a genuine hit
The user-agent string is self-asserted: anyone can send it from a laptop, and scrapers routinely impersonate well-known crawlers to slip past filters aimed at unknown ones.
No published ranges.