What this client asked www.pathwren.workers.dev for, when, and what it got.
Every number below is from this host's own server-side request log, 2026-09-01 05:23:47Z to 2026-09-01 05:28:35Z UTC.
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.4; +https://openai.com/gptbot) Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; robots.txt; +https://openai.com/searchbot
3 request(s) from 2 distinct address(es), 3 distinct path(s), first seen 2026-09-01 05:23:47Z, last seen 2026-09-01 05:28:35Z UTC. 1 separate visit(s), counting a gap of more than 30 minutes as a new one.
It describes itself, inside its own user-agent, as GPTBot/1.4
. That is the client's own words, quoted; this page makes no claim about what it is for.
The crawler catalogue on this site has a record for it: /crawler/gptbot.html — what it is for, and what blocking it costs. This page is only what it did here.
| user-agent | requests | addresses | first seen | last seen |
|---|---|---|---|---|
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.4; +https://openai.com/gptbot) | 2 | 1 | 2026-09-01 05:23:47Z | 2026-09-01 05:28:35Z |
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; robots.txt; +https://openai.com/searchbot | 1 | 1 | 2026-09-01 05:23:47Z | 2026-09-01 05:23:47Z |
First request to each path, oldest first.
| # | path | at (UTC) | status |
|---|---|---|---|
| 1 | /robots.txt | 2026-09-01 05:23:47Z | 200 |
| 2 | /icon.png | 2026-09-01 05:23:47Z | 200 |
| 3 | /sitemap.xml | 2026-09-01 05:28:35Z | 200 |
| path | requests | status codes | first | last |
|---|---|---|---|---|
/icon.png | 1 | 200×1 | 2026-09-01 05:23:47Z | 2026-09-01 05:23:47Z |
/robots.txt | 1 | 200×1 | 2026-09-01 05:23:47Z | 2026-09-01 05:23:47Z |
/sitemap.xml | 1 | 200×1 | 2026-09-01 05:28:35Z | 2026-09-01 05:28:35Z |
Nothing. Every path it asked for existed at the time it asked.
published IP ranges The published ranges we mirror for it are at /ip-ranges/openai-gptbot.html.
From the crawler record at /crawler/gptbot.html, which cites https://platform.openai.com/docs/bots — one source for both pages.
Not observed. We have not seen this client publish a grade, a listing or a record about this host anywhere, and we make no claim that it does or does not.
It names https://openai.com/gptbot in its own user-agent. That is the operator's own claim about itself; we neither endorse nor verify what is on it.
curl -s https://www.pathwren.workers.dev/bot/gptbot.json curl -s https://www.pathwren.workers.dev/data/observed-clients.json # every client, one request
JSON · markdown · all clients seen here · index as JSON
Rows come from a server-side log written before anything is served, so clients that
run no JavaScript are counted exactly like browsers. The window is the whole life of the log,
2026-09-01 05:23:47Z to 2026-09-01 05:28:35Z UTC, and it is stated on every number because a count without a window is
not a fact. Addresses are stored as salted hashes and counted, never printed. Our own checks
send X-Self: 1 and are excluded, together with every request the analyst flagged
as one of our own agents reaching for a URL through a model provider's fetch tool
(the numbers are on the index). Nothing here is a claim about
intent: where this page does not know something it says not observed
.