Google-Extended

Google · AI training crawlers · json

User-agent: Google-Extended
Disallow: /
robots.txt tokenGoogle-Extended
User-agent contains(control token only — no crawler)
OperatorGoogle
CategoryAI training crawlers
robots.txtcontrol token only — no crawler
Verify byno published verification method

What it is

Not a crawler. A robots.txt token that tells Google whether pages Googlebot already fetched may be used to train and ground Gemini. You will never see it in an access log; disallowing it changes what Google does with content it fetched under a different name.

What blocking it costs you

You are excluded from Gemini grounding and Gemini training. Google Search ranking and indexing are explicitly unaffected. This is the cleanest 'no training, keep my search traffic' lever that exists.

Full user-agent string

(none: Google-Extended never appears as a user-agent)

Allow it instead

User-agent: Google-Extended
Allow: /

Operator documentation: https://developers.google.com/search/docs/crawling-indexing/overview-google-crawlers
Machine copies: json · markdown
Policies that name this crawler: allow-all · block-ai-training · block-all-ai · maximum-ai-visibility