curl -s https://www.pathwren.workers.dev/crawler/yandexblogs.json   # this page, as JSON

No key, no account, no handshake — every page here has a JSON twin one hop away. Machine doors: 6 keyless GET tools · documents.json · changes · llms.txt · openapi.json · agent card · mcp · a2a

YandexBlogs

Yandex · Search engines · json

User-agent: YandexBlogs
Disallow: /
robots.txt tokenYandexBlogs
User-agent containsYandexBlogs
OperatorYandex
CategorySearch engines
robots.txtobeys robots.txt (documented)
Verify byreverse DNS

What it is

Yandex's blog-search robot; it indexes post comments as well as posts.

What blocking it costs you

Comment threads and blog posts stop being findable through Yandex blog search.

Full user-agent string

Mozilla/5.0 (compatible; YandexBlogs/0.99; robot; +http://yandex.com/bots)

Allow it instead

User-agent: YandexBlogs
Allow: /

Operator documentation: https://yandex.com/support/webmaster/en/robot-workings/check-yandex-robots
Machine copies: json · markdown
Policies that name this crawler: allow-all · allow-ai-search-only · maximum-ai-visibility