curl -s https://www.pathwren.workers.dev/crawler/firecrawlagent.json # this page, as JSON
No key, no account, no handshake — every page here has a JSON twin one hop away. Machine doors: 6 keyless GET tools · documents.json · changes · llms.txt · openapi.json · agent card · mcp · a2a
Firecrawl · Tools and frameworks · json
User-agent: FirecrawlAgent Disallow: /
| robots.txt token | FirecrawlAgent |
| User-agent contains | FirecrawlAgent |
| Operator | Firecrawl |
| Category | Tools and frameworks |
| robots.txt | obeys robots.txt (documented) |
| Verify by | no published verification method |
A hosted scrape-to-markdown service that LLM applications call to read pages. The requester is whoever is building on it, not Firecrawl itself, so volume and intent vary wildly.
Applications built on Firecrawl cannot read your pages. This is increasingly how agents fetch the web, so it is a bigger block than its name suggests.
Mozilla/5.0 (compatible; FirecrawlAgent/1.0; +https://firecrawl.dev)
User-agent: FirecrawlAgent Allow: /
Operator documentation: https://docs.firecrawl.dev/
Machine copies: json ·
markdown
Policies that name this crawler:
allow-all · maximum-ai-visibility