curl -s https://www.pathwren.workers.dev/crawler/baiduspider.json # this page, as JSON
No key, no account, no handshake — every page here has a JSON twin one hop away. Machine doors: 6 keyless GET tools · documents.json · changes · llms.txt · openapi.json · agent card · mcp · a2a
Baidu · Search engines · json
User-agent: Baiduspider Disallow: /
| robots.txt token | Baiduspider |
| User-agent contains | Baiduspider |
| Operator | Baidu |
| Category | Search engines |
| robots.txt | obeys robots.txt (documented) |
| Verify by | reverse DNS |
Baidu's search crawler, and the ingest path for Baidu's Ernie-backed answers.
Removal from Baidu Search, which matters only if you want Chinese-language traffic.
Mozilla/5.0 (compatible; Baiduspider/2.0; +http://www.baidu.com/search/spider.html)
User-agent: Baiduspider Allow: /
Operator documentation: https://help.baidu.com/question?prod_id=99&class=0&id=3001
Machine copies: json ·
markdown
Policies that name this crawler:
allow-all · allow-ai-search-only · maximum-ai-visibility