GPTBot

What this client asked www.pathwren.workers.dev for, when, and what it got. Every number below is from this host's own server-side request log, 2026-09-01 05:23:47Z to 2026-09-01 05:28:35Z UTC.

Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.4; +https://openai.com/gptbot)
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; robots.txt; +https://openai.com/searchbot

3 request(s) from 2 distinct address(es), 3 distinct path(s), first seen 2026-09-01 05:23:47Z, last seen 2026-09-01 05:28:35Z UTC. 1 separate visit(s), counting a gap of more than 30 minutes as a new one.

It describes itself, inside its own user-agent, as GPTBot/1.4. That is the client's own words, quoted; this page makes no claim about what it is for.

The crawler catalogue on this site has a record for it: /crawler/gptbot.html — what it is for, and what blocking it costs. This page is only what it did here.

The user-agent strings, exactly as they arrived

user-agentrequestsaddressesfirst seenlast seen
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.4; +https://openai.com/gptbot)212026-09-01 05:23:47Z2026-09-01 05:28:35Z
Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/131.0.0.0 Safari/537.36; compatible; OAI-SearchBot/1.4; robots.txt; +https://openai.com/searchbot112026-09-01 05:23:47Z2026-09-01 05:23:47Z

What it asked for, in the order it asked

First request to each path, oldest first.

#pathat (UTC)status
1/robots.txt2026-09-01 05:23:47Z200
2/icon.png2026-09-01 05:23:47Z200
3/sitemap.xml2026-09-01 05:28:35Z200

Everything it asked for

pathrequestsstatus codesfirstlast
/icon.png1200×12026-09-01 05:23:47Z2026-09-01 05:23:47Z
/robots.txt1200×12026-09-01 05:23:47Z2026-09-01 05:23:47Z
/sitemap.xml1200×12026-09-01 05:28:35Z2026-09-01 05:28:35Z

What it asked for that did not exist

Nothing. Every path it asked for existed at the time it asked.

What it got, and what it sent

Status codes
200×3
Accept headers
*/*
Bytes served
26721
Attributed to a channel
none — it arrived at a plain path
Our instrument classed it
agent×3 — that is our classifier's label from the user-agent, not the operator's, and it has been wrong before.

How to verify it is really them

published IP ranges The published ranges we mirror for it are at /ip-ranges/openai-gptbot.html.

From the crawler record at /crawler/gptbot.html, which cites https://platform.openai.com/docs/bots — one source for both pages.

What it publishes about you afterwards

Not observed. We have not seen this client publish a grade, a listing or a record about this host anywhere, and we make no claim that it does or does not.

Its own documentation

It names https://openai.com/gptbot in its own user-agent. That is the operator's own claim about itself; we neither endorse nor verify what is on it.

This page as data

curl -s https://www.pathwren.workers.dev/bot/gptbot.json
curl -s https://www.pathwren.workers.dev/data/observed-clients.json   # every client, one request

JSON · markdown · all clients seen here · index as JSON

Method, and what this page will not say

Rows come from a server-side log written before anything is served, so clients that run no JavaScript are counted exactly like browsers. The window is the whole life of the log, 2026-09-01 05:23:47Z to 2026-09-01 05:28:35Z UTC, and it is stated on every number because a count without a window is not a fact. Addresses are stored as salted hashes and counted, never printed. Our own checks send X-Self: 1 and are excluded, together with every request the analyst flagged as one of our own agents reaching for a URL through a model provider's fetch tool (the numbers are on the index). Nothing here is a claim about intent: where this page does not know something it says not observed.