search.marginalia.nu

What this client asked www.pathwren.workers.dev for, when, and what it got. Every number below is from this host's own server-side request log, 2026-09-11 14:41:58Z to 2026-09-11 14:44:07Z UTC.

search.marginalia.nu
search.marginalia.nu, search.marginalia.nu

106 request(s) from 1 distinct address(es), 104 distinct path(s), first seen 2026-09-11 14:41:58Z, last seen 2026-09-11 14:44:07Z UTC. 1 separate visit(s), counting a gap of more than 30 minutes as a new one.

The user-agent strings, exactly as they arrived

user-agentrequestsaddressesfirst seenlast seen
search.marginalia.nu10412026-09-11 14:41:58Z2026-09-11 14:44:07Z
search.marginalia.nu, search.marginalia.nu212026-09-11 14:42:02Z2026-09-11 14:42:04Z

What it asked for, in the order it asked

First request to each path, oldest first, first 40 shown.

#pathat (UTC)status
1/2026-09-11 14:41:58Z200
2/robots.txt2026-09-11 14:42:00Z200
3/feed.xml2026-09-11 14:42:02Z200
4/icon.png2026-09-11 14:42:03Z200
5/sitemap.xml2026-09-11 14:42:04Z200
6/crawler/2026-09-11 14:42:05Z200
7/operator/2026-09-11 14:42:06Z200
8/policy/2026-09-11 14:42:08Z200
9/ip-ranges/2026-09-11 14:42:09Z200
10/status.html2026-09-11 14:42:10Z200
11/data/2026-09-11 14:42:11Z200
12/api.html2026-09-11 14:42:12Z200
13/mcp.html2026-09-11 14:42:13Z200
14/llms.txt2026-09-11 14:42:15Z200
15/tools/2026-09-11 14:42:16Z200
16/documents.json2026-09-11 14:42:17Z200
17/changes2026-09-11 14:42:19Z200
18/openapi.json2026-09-11 14:42:20Z200
19/.well-known/agent-card.json2026-09-11 14:42:23Z200
20/mcp2026-09-11 14:42:25Z200
21/a2a2026-09-11 14:42:26Z200
22/register2026-09-11 14:42:27Z200
23/pricing2026-09-11 14:42:28Z200
24/reference2026-09-11 14:42:29Z200
25/feed.json2026-09-11 14:42:30Z200
26/crawler/gptbot.html2026-09-11 14:42:32Z200
27/crawler/oai-searchbot.html2026-09-11 14:42:33Z200
28/crawler/claudebot.html2026-09-11 14:42:35Z200
29/crawler/claude-searchbot.html2026-09-11 14:42:36Z200
30/crawler/google-extended.html2026-09-11 14:42:37Z200
31/crawler/perplexitybot.html2026-09-11 14:42:38Z200
32/crawler/ccbot.html2026-09-11 14:42:39Z200
33/crawler/bytespider.html2026-09-11 14:42:40Z200
34/crawler/index.html2026-09-11 14:42:42Z200
35/category/search.html2026-09-11 14:42:43Z200
36/category/ai-training.html2026-09-11 14:42:44Z200
37/category/tool.html2026-09-11 14:42:45Z200
38/category/ai-search.html2026-09-11 14:42:46Z200
39/category/dataset.html2026-09-11 14:42:47Z200
40/category/seo.html2026-09-11 14:42:49Z200

Everything it asked for

pathrequestsstatus codesfirstlast
/2200×22026-09-11 14:41:58Z2026-09-11 14:42:01Z
/data/observed-clients.csv2200×22026-09-11 14:43:04Z2026-09-11 14:43:05Z
/.well-known/agent-card.json1200×12026-09-11 14:42:23Z2026-09-11 14:42:23Z
/a2a1200×12026-09-11 14:42:26Z2026-09-11 14:42:26Z
/a2a.html1200×12026-09-11 14:43:36Z2026-09-11 14:43:36Z
/about.html1200×12026-09-11 14:43:28Z2026-09-11 14:43:28Z
/ai-crawler-logs/index.html1200×12026-09-11 14:43:37Z2026-09-11 14:43:37Z
/ai-crawler-robots/index.html1200×12026-09-11 14:43:38Z2026-09-11 14:43:38Z
/api.html1200×12026-09-11 14:42:12Z2026-09-11 14:42:12Z
/blog/ai-crawler-cost.html1200×12026-09-11 14:43:39Z2026-09-11 14:43:39Z
/blog/ai-crawler-traffic-2026-w36.html1200×12026-09-11 14:43:40Z2026-09-11 14:43:40Z
/blog/asked-and-absent-2026-w36.html1200×12026-09-11 14:43:41Z2026-09-11 14:43:41Z
/blog/client-census.html1200×12026-09-11 14:43:42Z2026-09-11 14:43:42Z
/blog/crawler-fleet-fold-2026-w36.html1200×12026-09-11 14:43:44Z2026-09-11 14:43:44Z
/blog/crawler-ua-asn-2026-w36.html1200×12026-09-11 14:43:45Z2026-09-11 14:43:45Z
/blog/fediverse-fanout-2026-w36.html1200×12026-09-11 14:43:46Z2026-09-11 14:43:46Z
/blog/index.html1200×12026-09-11 14:43:47Z2026-09-11 14:43:47Z
/blog/mcp-conformance-2026-w36.html1200×12026-09-11 14:43:48Z2026-09-11 14:43:48Z
/blog/mcp-endpoint-callers-2026-w36.html1200×12026-09-11 14:43:49Z2026-09-11 14:43:49Z
/blog/workers-plan-2026-w36.html1200×12026-09-11 14:43:51Z2026-09-11 14:43:51Z
/bot/1200×12026-09-11 14:43:00Z2026-09-11 14:43:00Z
/bot/amazonbot.html1200×12026-09-11 14:42:54Z2026-09-11 14:42:54Z
/bot/claudebot.html1200×12026-09-11 14:42:57Z2026-09-11 14:42:57Z
/bot/gptbot.html1200×12026-09-11 14:42:58Z2026-09-11 14:42:58Z
/bot/meta-externalagent.html1200×12026-09-11 14:42:59Z2026-09-11 14:42:59Z
/bot/node.html1200×12026-09-11 14:42:53Z2026-09-11 14:42:53Z
/bot/sentineloracle.html1200×12026-09-11 14:42:55Z2026-09-11 14:42:55Z
/c1200×12026-09-11 14:43:52Z2026-09-11 14:43:52Z
/category/ai-search.html1200×12026-09-11 14:42:46Z2026-09-11 14:42:46Z
/category/ai-training.html1200×12026-09-11 14:42:44Z2026-09-11 14:42:44Z
/category/archive.html1200×12026-09-11 14:42:52Z2026-09-11 14:42:52Z
/category/dataset.html1200×12026-09-11 14:42:47Z2026-09-11 14:42:47Z
/category/preview.html1200×12026-09-11 14:42:51Z2026-09-11 14:42:51Z
/category/search.html1200×12026-09-11 14:42:43Z2026-09-11 14:42:43Z
/category/seo.html1200×12026-09-11 14:42:49Z2026-09-11 14:42:49Z
/category/tool.html1200×12026-09-11 14:42:45Z2026-09-11 14:42:45Z
/category/user-fetch.html1200×12026-09-11 14:42:50Z2026-09-11 14:42:50Z
/changelog.html1200×12026-09-11 14:43:53Z2026-09-11 14:43:53Z
/changes1200×12026-09-11 14:42:19Z2026-09-11 14:42:19Z
/changes.json1200×12026-09-11 14:43:31Z2026-09-11 14:43:31Z
/compliance1200×12026-09-11 14:43:54Z2026-09-11 14:43:54Z
/contact1200×12026-09-11 14:43:55Z2026-09-11 14:43:55Z
/crawler/1200×12026-09-11 14:42:05Z2026-09-11 14:42:05Z
/crawler/adsbot-google-mobile-apps.html1200×12026-09-11 14:43:56Z2026-09-11 14:43:56Z
/crawler/adsbot-google-mobile.html1200×12026-09-11 14:43:58Z2026-09-11 14:43:58Z
/crawler/adsbot-google.html1200×12026-09-11 14:43:59Z2026-09-11 14:43:59Z
/crawler/ahrefsbot.html1200×12026-09-11 14:44:00Z2026-09-11 14:44:00Z
/crawler/ahrefssiteaudit.html1200×12026-09-11 14:44:01Z2026-09-11 14:44:01Z
/crawler/ai2bot-dolma.html1200×12026-09-11 14:44:02Z2026-09-11 14:44:02Z
/crawler/ai2bot.html1200×12026-09-11 14:44:04Z2026-09-11 14:44:04Z
/crawler/aihitbot.html1200×12026-09-11 14:44:05Z2026-09-11 14:44:05Z
/crawler/aiwebindex.html1200×12026-09-11 14:44:06Z2026-09-11 14:44:06Z
/crawler/amazonbot.html1200×12026-09-11 14:44:07Z2026-09-11 14:44:07Z
/crawler/bytespider.html1200×12026-09-11 14:42:40Z2026-09-11 14:42:40Z
/crawler/ccbot.html1200×12026-09-11 14:42:39Z2026-09-11 14:42:39Z
/crawler/claude-searchbot.html1200×12026-09-11 14:42:36Z2026-09-11 14:42:36Z
/crawler/claudebot.html1200×12026-09-11 14:42:35Z2026-09-11 14:42:35Z
/crawler/google-extended.html1200×12026-09-11 14:42:37Z2026-09-11 14:42:37Z
/crawler/gptbot.html1200×12026-09-11 14:42:32Z2026-09-11 14:42:32Z
/crawler/index.html1200×12026-09-11 14:42:42Z2026-09-11 14:42:42Z

44 further paths are in the JSON.

What it asked for that did not exist

Nothing. Every path it asked for existed at the time it asked.

What it got, and what it sent

Status codes
200×106
Accept headers
text/html, application/xhtml+xml, text/*;q=0.8, text/*, */*;q=0.9
Bytes served
2672544
Attributed to a channel
none — it arrived at a plain path
Our instrument classed it
agent×104, crawler×2 — that is our classifier's label from the user-agent, not the operator's, and it has been wrong before.

How to verify it is really them

Not observed. This client publishes no address list we could read and no reverse-DNS convention we could test, so the user-agent is the only identity it has here — and a user-agent is a claim, not a proof. Anyone can send this string.

Ask this page a question, with one GET

A question about this page, answered by a call you can make right now — no key, no account, no JSON-RPC, nothing to assemble. The query string is already filled in with this page's own subject.

curl -s 'https://www.pathwren.workers.dev/tools/classify-ua?ua=search.marginalia.nu&s=client-dossiers'

Every parameter above is a value the endpoint reads, not an address it visits: nothing is fetched on your behalf. Called without any parameter, each endpoint returns its own specification instead of an error. All of them.

What it publishes about you afterwards

Not observed. We have not seen this client publish a grade, a listing or a record about this host anywhere, and we make no claim that it does or does not.

Its own documentation

Not observed. It carries no URL and no contact address in its user-agent, so there is nowhere documented to ask what it is.

This page as data

curl -s https://www.pathwren.workers.dev/bot/search-marginalia-nu.json
curl -s https://www.pathwren.workers.dev/data/observed-clients.json   # every client, one request

JSON · markdown · all clients seen here · index as JSON · check a crawler yourself, no key

Method, and what this page will not say

Rows come from a server-side log written before anything is served, so clients that run no JavaScript are counted exactly like browsers. The window is the whole life of the log, 2026-09-11 14:41:58Z to 2026-09-11 14:44:07Z UTC, and it is stated on every number because a count without a window is not a fact. Addresses are stored as salted hashes and counted, never printed. Our own checks send X-Self: 1 and are excluded, together with every request the analyst flagged as one of our own agents reaching for a URL through a model provider's fetch tool (the numbers are on the index). Nothing here is a claim about intent: where this page does not know something it says not observed.