YandexBot

What this client asked www.pathwren.workers.dev for, when, and what it got. Every number below is from this host's own server-side request log, 2026-08-31 20:58:12Z to 2026-09-01 06:16:32Z UTC.

Mozilla/5.0 (compatible; YandexBot/3.0; +http://yandex.com/bots)

71 request(s) from 65 distinct address(es), 66 distinct path(s), first seen 2026-08-31 20:58:12Z, last seen 2026-09-01 06:16:32Z UTC. 4 separate visit(s), counting a gap of more than 30 minutes as a new one, median 120.0 minutes between them.

It describes itself, inside its own user-agent, as YandexBot/3.0. That is the client's own words, quoted; this page makes no claim about what it is for.

The crawler catalogue on this site has a record for it: /crawler/yandexbot.html — what it is for, and what blocking it costs. This page is only what it did here.

The user-agent strings, exactly as they arrived

user-agentrequestsaddressesfirst seenlast seen
Mozilla/5.0 (compatible; YandexBot/3.0; +http://yandex.com/bots)71652026-08-31 20:58:12Z2026-09-01 06:16:32Z

What it asked for, in the order it asked

First request to each path, oldest first, first 40 shown.

#pathat (UTC)status
1/988171cd4c390c0650e32d6b4ac6cfaf.txt2026-08-31 20:58:12Znot recorded
2/robots.txt2026-09-01 02:59:29Z200
3/operator/microsoft.html2026-09-01 02:59:30Z200
4/crawler/archive-org-bot.html2026-09-01 03:00:42Z200
5/category/preview.html2026-09-01 03:00:42Z200
6/operator/cohere.html2026-09-01 03:00:43Z200
7/crawler/ahrefsbot.html2026-09-01 03:00:43Z200
8/snippet/caddy.html2026-09-01 03:00:43Z200
9/crawler/petalbot.html2026-09-01 03:00:43Z200
10/operator/google.html2026-09-01 03:00:44Z200
11/crawler/seznambot.html2026-09-01 03:00:44Z200
12/operator/mistral.html2026-09-01 03:00:44Z200
13/operator/perplexity.html2026-09-01 03:00:45Z200
14/a2a.html2026-09-01 03:00:45Z200
15/category/tool.html2026-09-01 03:00:46Z200
16/changelog.html2026-09-01 03:00:46Z200
17/operator/openai.html2026-09-01 03:00:47Z200
18/category/user-fetch.html2026-09-01 03:00:47Z200
19/operator/amazon.html2026-09-01 03:00:48Z200
20/snippet/apache-htaccess.html2026-09-01 03:00:48Z200
21/crawler/timpibot.html2026-09-01 03:00:49Z200
22/crawler/anthropic-ai.html2026-09-01 03:00:49Z200
23/ip-ranges/openai-gptbot.html2026-09-01 03:00:50Z200
24/status.html2026-09-01 06:15:17Z200
25/category/ai-training.html2026-09-01 06:15:17Z200
26/policy/maximum-ai-visibility.html2026-09-01 06:15:17Z200
27/bot/measure-mcp-schema.html2026-09-01 06:15:17Z200
28/crawler/index.html2026-09-01 06:15:17Z200
29/bot/curl.html2026-09-01 06:15:18Z200
30/crawler/ai2bot.html2026-09-01 06:15:18Z200
31/bot/mcp-observatory.html2026-09-01 06:15:19Z200
32/bot/agentreputationbot.html2026-09-01 06:15:20Z200
33/ip-ranges/perplexity-user.html2026-09-01 06:15:20Z200
34/bot/aetherlink-public-agent-card-policy-check.html2026-09-01 06:15:21Z200
35/operator/laion.html2026-09-01 06:15:21Z200
36/bot/mozilla.html2026-09-01 06:15:22Z200
37/operator/timpi.html2026-09-01 06:15:22Z200
38/bot/mcp-schema-archive.html2026-09-01 06:15:23Z200
39/crawler/scrapy.html2026-09-01 06:15:23Z200
40/snippet/python-classify.html2026-09-01 06:15:24Z200

Everything it asked for

pathrequestsstatus codesfirstlast
/robots.txt6200×62026-09-01 02:59:29Z2026-09-01 06:15:16Z
/988171cd4c390c0650e32d6b4ac6cfaf.txt1not recorded×12026-08-31 20:58:12Z2026-08-31 20:58:12Z
/a2a.html1200×12026-09-01 03:00:45Z2026-09-01 03:00:45Z
/about.html1200×12026-09-01 06:16:24Z2026-09-01 06:16:24Z
/api.html1200×12026-09-01 06:16:28Z2026-09-01 06:16:28Z
/bot/402explorer.html1200×12026-09-01 06:16:28Z2026-09-01 06:16:28Z
/bot/a2a-registry-smoke.html1200×12026-09-01 06:16:29Z2026-09-01 06:16:29Z
/bot/aetherlink-public-agent-card-policy-check.html1200×12026-09-01 06:15:21Z2026-09-01 06:15:21Z
/bot/agentreputationbot.html1200×12026-09-01 06:15:20Z2026-09-01 06:15:20Z
/bot/archive-org-bot.html1200×12026-09-01 06:16:28Z2026-09-01 06:16:28Z
/bot/curl.html1200×12026-09-01 06:15:18Z2026-09-01 06:15:18Z
/bot/deno.html1200×12026-09-01 06:16:23Z2026-09-01 06:16:23Z
/bot/exaforce-mcprep.html1200×12026-09-01 06:16:30Z2026-09-01 06:16:30Z
/bot/gf-agent-toll-outbound.html1200×12026-09-01 06:16:26Z2026-09-01 06:16:26Z
/bot/guzzlehttp.html1200×12026-09-01 06:16:29Z2026-09-01 06:16:29Z
/bot/mcp-observatory.html1200×12026-09-01 06:15:19Z2026-09-01 06:15:19Z
/bot/mcp-schema-archive.html1200×12026-09-01 06:15:23Z2026-09-01 06:15:23Z
/bot/measure-mcp-schema.html1200×12026-09-01 06:15:17Z2026-09-01 06:15:17Z
/bot/mozilla.html1200×12026-09-01 06:15:22Z2026-09-01 06:15:22Z
/bot/python-httpx2.html1200×12026-09-01 06:15:24Z2026-09-01 06:15:24Z
/bot/sentineloracle.html1200×12026-09-01 06:16:25Z2026-09-01 06:16:25Z
/bot/telegrambot.html1200×12026-09-01 06:16:31Z2026-09-01 06:16:31Z
/bot/undici.html1200×12026-09-01 06:16:24Z2026-09-01 06:16:24Z
/category/ai-training.html1200×12026-09-01 06:15:17Z2026-09-01 06:15:17Z
/category/preview.html1200×12026-09-01 03:00:42Z2026-09-01 03:00:42Z
/category/seo.html1200×12026-09-01 06:15:25Z2026-09-01 06:15:25Z
/category/tool.html1200×12026-09-01 03:00:46Z2026-09-01 03:00:46Z
/category/user-fetch.html1200×12026-09-01 03:00:47Z2026-09-01 03:00:47Z
/changelog.html1200×12026-09-01 03:00:46Z2026-09-01 03:00:46Z
/crawler/ahrefsbot.html1200×12026-09-01 03:00:43Z2026-09-01 03:00:43Z
/crawler/ai2bot.html1200×12026-09-01 06:15:18Z2026-09-01 06:15:18Z
/crawler/amazonbot.html1200×12026-09-01 06:16:23Z2026-09-01 06:16:23Z
/crawler/anthropic-ai.html1200×12026-09-01 03:00:49Z2026-09-01 03:00:49Z
/crawler/archive-org-bot.html1200×12026-09-01 03:00:42Z2026-09-01 03:00:42Z
/crawler/googlebot-image.html1200×12026-09-01 06:16:25Z2026-09-01 06:16:25Z
/crawler/index.html1200×12026-09-01 06:15:17Z2026-09-01 06:15:17Z
/crawler/oai-searchbot.html1200×12026-09-01 06:15:27Z2026-09-01 06:15:27Z
/crawler/petalbot.html1200×12026-09-01 03:00:43Z2026-09-01 03:00:43Z
/crawler/scrapy.html1200×12026-09-01 06:15:23Z2026-09-01 06:15:23Z
/crawler/seznambot.html1200×12026-09-01 03:00:44Z2026-09-01 03:00:44Z
/crawler/storebot-google.html1200×12026-09-01 06:15:26Z2026-09-01 06:15:26Z
/crawler/timpibot.html1200×12026-09-01 03:00:49Z2026-09-01 03:00:49Z
/ip-ranges/bing-bingbot.html1200×12026-09-01 06:16:23Z2026-09-01 06:16:23Z
/ip-ranges/google-user-triggered.html1200×12026-09-01 06:16:26Z2026-09-01 06:16:26Z
/ip-ranges/openai-gptbot.html1200×12026-09-01 03:00:50Z2026-09-01 03:00:50Z
/ip-ranges/perplexity-user.html1200×12026-09-01 06:15:20Z2026-09-01 06:15:20Z
/mcp-doctor.html1200×12026-09-01 06:16:27Z2026-09-01 06:16:27Z
/operator/amazon.html1200×12026-09-01 03:00:48Z2026-09-01 03:00:48Z
/operator/cohere.html1200×12026-09-01 03:00:43Z2026-09-01 03:00:43Z
/operator/commoncrawl.html1200×12026-09-01 06:16:32Z2026-09-01 06:16:32Z
/operator/google.html1200×12026-09-01 03:00:44Z2026-09-01 03:00:44Z
/operator/laion.html1200×12026-09-01 06:15:21Z2026-09-01 06:15:21Z
/operator/microsoft.html1200×12026-09-01 02:59:30Z2026-09-01 02:59:30Z
/operator/mistral.html1200×12026-09-01 03:00:44Z2026-09-01 03:00:44Z
/operator/openai.html1200×12026-09-01 03:00:47Z2026-09-01 03:00:47Z
/operator/perplexity.html1200×12026-09-01 03:00:45Z2026-09-01 03:00:45Z
/operator/scrapy.html1200×12026-09-01 06:16:30Z2026-09-01 06:16:30Z
/operator/timpi.html1200×12026-09-01 06:15:22Z2026-09-01 06:15:22Z
/policy/block-all-ai.html1200×12026-09-01 06:15:26Z2026-09-01 06:15:26Z
/policy/maximum-ai-visibility.html1200×12026-09-01 06:15:17Z2026-09-01 06:15:17Z

6 further paths are in the JSON.

What it asked for that did not exist

Nothing. Every path it asked for existed at the time it asked.

What it got, and what it sent

Status codes
200×70, not recorded×1
Accept headers
text/html,application/xhtml+xml,application/xml;q=0.9,*/*;q=0.8, */*
Bytes served
423546
Attributed to a channel
none — it arrived at a plain path
Our instrument classed it
crawler×71 — that is our classifier's label from the user-agent, not the operator's, and it has been wrong before.

How to verify it is really them

reverse DNS

From the crawler record at /crawler/yandexbot.html, which cites https://yandex.com/support/webmaster/robot-workings/check-yandex-robots.html — one source for both pages.

What it publishes about you afterwards

Not observed. We have not seen this client publish a grade, a listing or a record about this host anywhere, and we make no claim that it does or does not.

Its own documentation

It names http://yandex.com/bots in its own user-agent. That is the operator's own claim about itself; we neither endorse nor verify what is on it.

This page as data

curl -s https://www.pathwren.workers.dev/bot/yandexbot.json
curl -s https://www.pathwren.workers.dev/data/observed-clients.json   # every client, one request

JSON · markdown · all clients seen here · index as JSON

Method, and what this page will not say

Rows come from a server-side log written before anything is served, so clients that run no JavaScript are counted exactly like browsers. The window is the whole life of the log, 2026-08-31 20:58:12Z to 2026-09-01 06:16:32Z UTC, and it is stated on every number because a count without a window is not a fact. Addresses are stored as salted hashes and counted, never printed. Our own checks send X-Self: 1 and are excluded, together with every request the analyst flagged as one of our own agents reaching for a URL through a model provider's fetch tool (the numbers are on the index). Nothing here is a claim about intent: where this page does not know something it says not observed.