meta-externalagent

What this client asked www.pathwren.workers.dev for, when, and what it got. Every number below is from this host's own server-side request log, 2026-09-03 00:05:33Z to 2026-09-03 06:24:31Z UTC.

Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))
Mozilla/5.0 (X11; Linux x86_64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))
Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 Edg/145.0.0.0 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))

20 request(s) from 15 distinct address(es), 20 distinct path(s), first seen 2026-09-03 00:05:33Z, last seen 2026-09-03 06:24:31Z UTC. 5 separate visit(s), counting a gap of more than 30 minutes as a new one, median 87.3 minutes between them.

It describes itself, inside its own user-agent, as Windows NT 10.0; Win64; x64. That is the client's own words, quoted; this page makes no claim about what it is for.

The crawler catalogue on this site has a record for it: /crawler/meta-externalagent.html — what it is for, and what blocking it costs. This page is only what it did here.

The user-agent strings, exactly as they arrived

user-agentrequestsaddressesfirst seenlast seen
Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))972026-09-03 00:27:45Z2026-09-03 06:15:15Z
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))662026-09-03 00:22:46Z2026-09-03 06:02:51Z
Mozilla/5.0 (X11; Linux x86_64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))332026-09-03 00:05:33Z2026-09-03 06:09:55Z
Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/145.0.0.0 Safari/537.36 Edg/145.0.0.0 (compatible; meta-externalagent/1.1 (+https://developers.facebook.com/docs/sharing/webmasters/crawler))222026-09-03 05:12:51Z2026-09-03 06:24:31Z

What it asked for, in the order it asked

First request to each path, oldest first.

#pathat (UTC)status
1/c/rubygems-registry/2026-09-03 00:05:33Z200
2/data/agents.json2026-09-03 00:22:46Z200
3/openapi.json2026-09-03 00:27:45Z200
4/documents.json2026-09-03 00:36:59Z200
5/mcp-triage.html2026-09-03 01:08:57Z200
6/mcp-netcheck.html2026-09-03 02:05:51Z200
7/mcp-doctor.html2026-09-03 02:24:28Z200
8/llms.txt2026-09-03 03:57:12Z200
9/status.json2026-09-03 04:03:17Z200
10/a2a2026-09-03 04:14:11Z200
11/tools/verification-methods2026-09-03 04:31:27Z200
12/.well-known/ai-catalog.json2026-09-03 04:43:05Z200
13/status.html2026-09-03 05:12:51Z200
14/terms.html2026-09-03 05:55:32Z200
15/tools/index.json2026-09-03 06:02:51Z200
16/data/ip-sources.json2026-09-03 06:05:09Z200
17/tools/classify-ua2026-09-03 06:09:55Z200
18/mcp.2026-09-03 06:13:33Z404
19/crawler/2026-09-03 06:15:15Z200
20/tools/example2026-09-03 06:24:31Z200

Everything it asked for

pathrequestsstatus codesfirstlast
/.well-known/ai-catalog.json1200×12026-09-03 04:43:05Z2026-09-03 04:43:05Z
/a2a1200×12026-09-03 04:14:11Z2026-09-03 04:14:11Z
/c/rubygems-registry/1200×12026-09-03 00:05:33Z2026-09-03 00:05:33Z
/crawler/1200×12026-09-03 06:15:15Z2026-09-03 06:15:15Z
/data/agents.json1200×12026-09-03 00:22:46Z2026-09-03 00:22:46Z
/data/ip-sources.json1200×12026-09-03 06:05:09Z2026-09-03 06:05:09Z
/documents.json1200×12026-09-03 00:36:59Z2026-09-03 00:36:59Z
/llms.txt1200×12026-09-03 03:57:12Z2026-09-03 03:57:12Z
/mcp-doctor.html1200×12026-09-03 02:24:28Z2026-09-03 02:24:28Z
/mcp-netcheck.html1200×12026-09-03 02:05:51Z2026-09-03 02:05:51Z
/mcp-triage.html1200×12026-09-03 01:08:57Z2026-09-03 01:08:57Z
/mcp.1404×12026-09-03 06:13:33Z2026-09-03 06:13:33Z
/openapi.json1200×12026-09-03 00:27:45Z2026-09-03 00:27:45Z
/status.html1200×12026-09-03 05:12:51Z2026-09-03 05:12:51Z
/status.json1200×12026-09-03 04:03:17Z2026-09-03 04:03:17Z
/terms.html1200×12026-09-03 05:55:32Z2026-09-03 05:55:32Z
/tools/classify-ua1200×12026-09-03 06:09:55Z2026-09-03 06:09:55Z
/tools/example1200×12026-09-03 06:24:31Z2026-09-03 06:24:31Z
/tools/index.json1200×12026-09-03 06:02:51Z2026-09-03 06:02:51Z
/tools/verification-methods1200×12026-09-03 04:31:27Z2026-09-03 04:31:27Z

What it asked for that did not exist

path it asked forwhat it gotfirst askedsince then
/mcp.404×12026-09-03 06:13:33Zstill absent

What it got, and what it sent

Status codes
200×19, 404×1
Accept headers
*/*
Bytes served
1106992
Attributed to a channel
rubygems-registry (1)
Our instrument classed it
crawler×20 — that is our classifier's label from the user-agent, not the operator's, and it has been wrong before.

How to verify it is really them

no published verification method

From the crawler record at /crawler/meta-externalagent.html, which cites https://developers.facebook.com/docs/sharing/webmasters/web-crawlers — one source for both pages.

What it publishes about you afterwards

Not observed. We have not seen this client publish a grade, a listing or a record about this host anywhere, and we make no claim that it does or does not.

Its own documentation

It names https://developers.facebook.com/docs/sharing/webmasters/crawler in its own user-agent. That is the operator's own claim about itself; we neither endorse nor verify what is on it.

This page as data

curl -s https://www.pathwren.workers.dev/bot/meta-externalagent.json
curl -s https://www.pathwren.workers.dev/data/observed-clients.json   # every client, one request

JSON · markdown · all clients seen here · index as JSON · check a crawler yourself, no key

Method, and what this page will not say

Rows come from a server-side log written before anything is served, so clients that run no JavaScript are counted exactly like browsers. The window is the whole life of the log, 2026-09-03 00:05:33Z to 2026-09-03 06:24:31Z UTC, and it is stated on every number because a count without a window is not a fact. Addresses are stored as salted hashes and counted, never printed. Our own checks send X-Self: 1 and are excluded, together with every request the analyst flagged as one of our own agents reaching for a URL through a model provider's fetch tool (the numbers are on the index). Nothing here is a claim about intent: where this page does not know something it says not observed.