GPTBot

What this client asked www.pathwren.workers.dev for, when, and what it got. Every number below is from this host's own server-side request log, 2026-09-01 05:23:47Z to 2026-09-03 06:00:46Z UTC.

Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.4; +https://openai.com/gptbot)
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; robots.txt; +https://openai.com/searchbot
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; ChatGPT-User/1.0; +https://openai.com/bot
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.0; +https://openai.com/searchbot
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; GPTBot/1.1; +https://openai.com/gptbot
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; GPTBot/1.2; +https://openai.com/gptbot

1835 request(s) from 11 distinct address(es), 1354 distinct path(s), first seen 2026-09-01 05:23:47Z, last seen 2026-09-03 06:00:46Z UTC. 9 separate visit(s), counting a gap of more than 30 minutes as a new one, median 233.7 minutes between them.

It describes itself, inside its own user-agent, as GPTBot/1.4. That is the client's own words, quoted; this page makes no claim about what it is for.

The crawler catalogue on this site has a record for it: /crawler/gptbot.html — what it is for, and what blocking it costs. This page is only what it did here.

The user-agent strings, exactly as they arrived

user-agentrequestsaddressesfirst seenlast seen
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.4; +https://openai.com/gptbot)182322026-09-01 05:23:47Z2026-09-03 05:34:21Z
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; robots.txt; +https://openai.com/searchbot412026-09-01 05:23:47Z2026-09-03 04:15:23Z
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; ChatGPT-User/1.0; +https://openai.com/bot442026-09-03 00:33:36Z2026-09-03 00:36:00Z
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.0; +https://openai.com/searchbot222026-09-01 11:19:50Z2026-09-02 09:26:36Z
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; GPTBot/1.1; +https://openai.com/gptbot112026-09-03 00:22:52Z2026-09-03 00:22:52Z
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko); compatible; GPTBot/1.2; +https://openai.com/gptbot112026-09-03 06:00:46Z2026-09-03 06:00:46Z

What it asked for, in the order it asked

First request to each path, oldest first, first 40 shown.

#pathat (UTC)status
1/robots.txt2026-09-01 05:23:47Z200
2/icon.png2026-09-01 05:23:47Z200
3/sitemap.xml2026-09-01 05:28:35Z200
4/c/mcp-registry-official/mcp.html2026-09-01 06:38:39Z200
5/c/mcp-registry-official/mcp2026-09-01 06:38:42Z200
6/status.html2026-09-01 06:38:45Z200
7/apple-touch-icon.png2026-09-01 06:38:46Z200
8/llms.txt2026-09-01 06:38:48Z200
9/api.html2026-09-01 06:38:49Z200
10/crawler/2026-09-01 06:38:50Z200
11/mcp.html2026-09-01 06:38:51Z200
12/feed.json2026-09-01 06:38:51Z200
13/data/2026-09-01 06:38:52Z200
14/ip-ranges/2026-09-01 06:38:53Z200
15/feed.xml2026-09-01 06:38:53Z200
16/data/agents.json2026-09-01 06:38:53Z200
17/px.gif2026-09-01 06:38:54Z200
18/mcp-triage.html2026-09-01 06:38:55Z200
19/policy/2026-09-01 06:38:55Z200
20/openapi.json2026-09-01 06:38:55Z200
21/operator/2026-09-01 06:38:56Z200
22/favicon.ico2026-09-01 06:38:56Z200
23/mcp-doctor.html2026-09-01 06:38:57Z200
24/about.html2026-09-01 06:38:58Z200
25/2026-09-01 06:38:58Z200
26/status.json2026-09-01 06:38:58Z200
27/apis.json2026-09-01 06:38:59Z200
28/.well-known/api-catalog2026-09-01 06:39:00Z200
29/llms-full.txt2026-09-01 06:39:01Z200
30/openapi.yaml2026-09-01 06:39:01Z200
31/.well-known/api-onboarding2026-09-01 06:39:02Z200
32/crawler/ai2bot.html2026-09-01 06:39:03Z200
33/crawler/claude-searchbot.html2026-09-01 06:39:03Z200
34/crawler/amazonbot.html2026-09-01 06:39:04Z200
35/crawler/webzio-extended.html2026-09-01 06:39:05Z200
36/crawler/meta-externalagent.html2026-09-01 06:39:05Z200
37/crawler/mistralai-user.html2026-09-01 06:39:06Z200
38/category/ai-search.html2026-09-01 06:39:06Z200
39/crawler/chatgpt-user.html2026-09-01 06:39:07Z200
40/crawler/anthropic-ai.html2026-09-01 06:39:07Z200

Everything it asked for

pathrequestsstatus codesfirstlast
/px.gif450200×4502026-09-01 06:38:54Z2026-09-03 04:24:02Z
/robots.txt6200×62026-09-01 05:23:47Z2026-09-03 04:15:23Z
/3200×32026-09-01 06:38:58Z2026-09-03 06:00:46Z
/sitemap.xml3200×32026-09-01 05:28:35Z2026-09-03 05:34:21Z
/bot/index.html2200×22026-09-01 06:40:58Z2026-09-03 04:15:44Z
/data/agents.csv2200×22026-09-01 06:39:08Z2026-09-03 04:15:50Z
/data/agents.json2200×22026-09-01 06:38:53Z2026-09-03 04:15:40Z
/data/ip-sources.json2200×22026-09-01 06:42:29Z2026-09-03 04:15:52Z
/data/observed-clients.csv2200×22026-09-01 06:41:57Z2026-09-03 04:15:51Z
/data/observed-clients.json2200×22026-09-01 06:42:50Z2026-09-03 04:15:53Z
/data/robots-tokens.txt2200×22026-09-01 06:41:48Z2026-09-03 04:15:53Z
/data/ua-regex.json2200×22026-09-01 06:42:33Z2026-09-03 04:15:40Z
/data/ua-regex.txt2200×22026-09-01 06:41:52Z2026-09-03 04:15:41Z
/data/user-agents.txt2200×22026-09-01 06:39:22Z2026-09-03 04:15:53Z
/llms.txt2200×22026-09-01 06:38:48Z2026-09-03 00:36:00Z
/mcp.html2200×22026-09-01 06:38:51Z2026-09-03 00:33:36Z
/openapi.json2200×22026-09-01 06:38:55Z2026-09-03 00:36:00Z
/tools/ai-access.html2200×22026-09-03 04:15:44Z2026-09-03 04:15:46Z
/tools/classify-ua.html2200×22026-09-03 04:15:42Z2026-09-03 04:15:50Z
/tools/example.html2200×22026-09-03 04:15:42Z2026-09-03 04:15:51Z
/tools/index.html2200×22026-09-03 04:15:49Z2026-09-03 04:23:54Z
/tools/index.json2200×22026-09-03 04:15:45Z2026-09-03 04:23:59Z
/tools/robots-allowed.html2200×22026-09-03 04:15:46Z2026-09-03 04:15:48Z
/tools/robots-lint.html2200×22026-09-03 04:15:44Z2026-09-03 04:15:45Z
/tools/verification-methods.html2200×22026-09-03 04:15:47Z2026-09-03 04:15:54Z
/tools/verify-crawler.html2200×22026-09-03 04:15:41Z2026-09-03 04:15:48Z
/tools/whoami.html2200×22026-09-03 04:15:43Z2026-09-03 04:15:47Z
/.well-known/agent-card.json1200×12026-09-03 04:15:33Z2026-09-03 04:15:33Z
/.well-known/agent-permissions.json1200×12026-09-03 04:15:35Z2026-09-03 04:15:35Z
/.well-known/ai-catalog.json1200×12026-09-03 04:15:37Z2026-09-03 04:15:37Z
/.well-known/api-catalog1200×12026-09-01 06:39:00Z2026-09-01 06:39:00Z
/.well-known/api-onboarding1200×12026-09-01 06:39:02Z2026-09-01 06:39:02Z
/.well-known/glama.json1404×12026-09-01 13:05:08Z2026-09-01 13:05:08Z
/.well-known/oauth-authorization-server1404×12026-09-03 04:15:54Z2026-09-03 04:15:54Z
/.well-known/security.txt1200×12026-09-01 06:40:16Z2026-09-01 06:40:16Z
/.well-known/x4021200×12026-09-03 04:16:18Z2026-09-03 04:16:18Z
/a2a.html1200×12026-09-03 04:16:06Z2026-09-03 04:16:06Z
/a2a.json1200×12026-09-03 04:16:48Z2026-09-03 04:16:48Z
/a2a.md1200×12026-09-03 04:16:54Z2026-09-03 04:16:54Z
/a2a/discovery/.well-known/agent-card.json1200×12026-09-03 04:21:45Z2026-09-03 04:21:45Z
/a2a/doctor/.well-known/agent-card.json1200×12026-09-03 04:21:57Z2026-09-03 04:21:57Z
/a2a/example.json1200×12026-09-03 04:20:30Z2026-09-03 04:20:30Z
/a2a/lint/.well-known/agent-card.json1200×12026-09-03 04:21:59Z2026-09-03 04:21:59Z
/a2a/netcheck/.well-known/agent-card.json1200×12026-09-03 04:21:46Z2026-09-03 04:21:46Z
/a2a/robots/.well-known/agent-card.json1200×12026-09-03 04:21:55Z2026-09-03 04:21:55Z
/a2a/score/.well-known/agent-card.json1200×12026-09-03 04:21:58Z2026-09-03 04:21:58Z
/a2a/triage/.well-known/agent-card.json1200×12026-09-03 04:21:55Z2026-09-03 04:21:55Z
/about.html1200×12026-09-01 06:38:58Z2026-09-01 06:38:58Z
/ai-crawler-logs/1200×12026-09-03 04:17:39Z2026-09-03 04:17:39Z
/ai-crawler-logs/data.json1200×12026-09-03 04:20:20Z2026-09-03 04:20:20Z
/ai-crawler-logs/index.html1200×12026-09-03 04:18:47Z2026-09-03 04:18:47Z
/ai-crawler-logs/index.json1200×12026-09-03 04:21:35Z2026-09-03 04:21:35Z
/ai-crawler-logs/index.md1200×12026-09-03 04:20:09Z2026-09-03 04:20:09Z
/ai-crawler-robots/1200×12026-09-03 04:17:40Z2026-09-03 04:17:40Z
/ai-crawler-robots/data.json1200×12026-09-03 04:20:03Z2026-09-03 04:20:03Z
/ai-crawler-robots/index.html1200×12026-09-03 04:21:36Z2026-09-03 04:21:36Z
/ai-crawler-robots/index.json1200×12026-09-03 04:21:59Z2026-09-03 04:21:59Z
/ai-crawler-robots/index.md1200×12026-09-03 04:20:11Z2026-09-03 04:20:11Z
/api.html1200×12026-09-01 06:38:49Z2026-09-01 06:38:49Z
/apis.json1200×12026-09-01 06:38:59Z2026-09-01 06:38:59Z

1294 further paths are in the JSON.

What it asked for that did not exist

path it asked forwhat it gotfirst askedsince then
/.well-known/glama.json404×12026-09-01 13:05:08Zstill absent — on purpose: No, and 404 is the honest 'unclaimed'
/.well-known/oauth-authorization-server404×12026-09-03 04:15:54Zstill absent — on purpose: No

What it got, and what it sent

Status codes
200×1833, 404×2
Accept headers
*/*, text/html,application/xhtml+xml,application/xml;q=0.9,image/avif,image/webp,image/apng,*/*;q=0.8,application/signed-exchange;v=b3;q=0.9, text/plain
Bytes served
11388383
Attributed to a channel
pypi-registry (5), mcp-registry-official (3), glama-mcp (1), mcpservers-org (1), thecolony (1), agentlist (1)
Our instrument classed it
crawler×1029, agent×806 — that is our classifier's label from the user-agent, not the operator's, and it has been wrong before.

How to verify it is really them

published IP ranges The published ranges we mirror for it are at /ip-ranges/openai-gptbot.html.

From the crawler record at /crawler/gptbot.html, which cites https://platform.openai.com/docs/bots — one source for both pages.

What it publishes about you afterwards

Not observed. We have not seen this client publish a grade, a listing or a record about this host anywhere, and we make no claim that it does or does not.

Its own documentation

It names https://openai.com/gptbot in its own user-agent. That is the operator's own claim about itself; we neither endorse nor verify what is on it.

This page as data

curl -s https://www.pathwren.workers.dev/bot/gptbot.json
curl -s https://www.pathwren.workers.dev/data/observed-clients.json   # every client, one request

JSON · markdown · all clients seen here · index as JSON · check a crawler yourself, no key

Method, and what this page will not say

Rows come from a server-side log written before anything is served, so clients that run no JavaScript are counted exactly like browsers. The window is the whole life of the log, 2026-09-01 05:23:47Z to 2026-09-03 06:00:46Z UTC, and it is stated on every number because a count without a window is not a fact. Addresses are stored as salted hashes and counted, never printed. Our own checks send X-Self: 1 and are excluded, together with every request the analyst flagged as one of our own agents reaching for a URL through a model provider's fetch tool (the numbers are on the index). Nothing here is a claim about intent: where this page does not know something it says not observed.