curl -s https://www.pathwren.workers.dev/crawler/meta-webindexer.json   # this page, as JSON

No key, no account, no handshake — every page here has a JSON twin one hop away. Machine doors: 6 keyless GET tools · documents.json · changes · llms.txt · openapi.json · agent card · mcp · a2a

Meta-WebIndexer

Meta · AI search crawlers · json

User-agent: Meta-WebIndexer
Disallow: /
robots.txt tokenMeta-WebIndexer
User-agent containsMeta-WebIndexer
OperatorMeta
CategoryAI search crawlers
robots.txtoperator publishes no robots.txt statement
Verify byno published verification method

What it is

Per Meta's crawler documentation, Meta-WebIndexer navigates the web to improve the quality of Meta AI's search results. It is a third Meta token alongside Meta-ExternalAgent and Meta-ExternalFetcher, and the newest of them.

What blocking it costs you

You leave the index Meta AI answers from across Facebook, Instagram and WhatsApp — the largest assistant install base there is. A robots.txt that names the two older Meta tokens does not cover this one.

Full user-agent string

Meta-WebIndexer

Allow it instead

User-agent: Meta-WebIndexer
Allow: /

Operator documentation: https://developers.facebook.com/docs/sharing/webmasters/web-crawlers/
Machine copies: json · markdown
Policies that name this crawler: allow-all · block-all-ai · allow-ai-search-only · maximum-ai-visibility