UnsourcedSee who crawls your site →

AI TrainerGoogle

Google-Extended: IP ranges, verification & how to handle it

Last verified: 2026-07-16 · maintained by Unsourced

Google-Extended is the control Google uses for whether your content informs Gemini and its AI features. It uses Googlebot's infrastructure; opting out here doesn't affect normal Search indexing.

What Google-Extended does, and how it differs

Google-Extended isn't a crawler with traffic of its own — it's a robots.txt token telling Google whether your pages may be used to train and ground Gemini and other Google AI products. The fetching still happens over Googlebot's normal infrastructure, so a Google-Extended disallow changes AI usage only and leaves ordinary Search ranking and indexing untouched. It's the lever to opt out of AI training without paying for it in organic visibility.

How to verify Google-Extended

Google-Extended can be checked two independent ways, and a real request passes both. Its source IP has to fall inside the 585 ranges Google publishes, and it has to forward-confirm by reverse DNS: the IP resolves to google.com, googlebot.com, and that hostname resolves back to the same IP. Anything carrying the Google-Extended user-agent that fails either test is an impostor, whatever the header says.

Should you allow or block Google-Extended?

Recommended: keep. It's a permission token, not a traffic source: a disallow opts you out of Gemini training and grounding only, and leaves ordinary Google Search ranking and indexing completely untouched.

If you do choose to act in robots.txt (which crawlers honour but don't enforce):

# Google-Extended: recommended to ALLOW — blocking can cost you AI visibility
User-agent: Google-Extended
Disallow:

Official sources

Common questions about Google-Extended

Does blocking Google-Extended hurt my Google Search ranking?

No. Google-Extended is only a permission token for Gemini and Google's AI products — disallowing it stops AI training and grounding use while leaving ordinary Search crawling and ranking completely untouched.

Is Google-Extended a real crawler?

Not in the usual sense — it has no traffic of its own. The fetching happens over Googlebot's normal infrastructure; Google-Extended is just the robots.txt switch for whether those pages may feed Google's AI.

How do I opt out of Google's AI without losing Search?

Add a User-agent: Google-Extended disallow. That removes you from Gemini training and grounding only — Googlebot keeps crawling and indexing you for Search exactly as before.

Related crawlers

See who is really crawling your site — Google, or an impostor.

Unsourced checks each crawler against published ranges andreverse DNS, and shows where AI search cites you instead of a competitor.

Check your site free →

10-day free trial · no card required · cancel anytime