Google-Extended

Operated by GoogleAI Trainingrobots.txt control token

Not a crawler: a robots.txt token that controls whether Google uses your content for Gemini training and grounding.

What Google-Extended does

Google-Extended is a standalone product token: no separate crawler visits your site with this name. Adding it to robots.txt tells Google not to use content Googlebot has already fetched for training Gemini models or for grounding in Gemini apps and the Gemini API. It does not affect Google Search indexing or AI Overviews eligibility.

Should I opt out of Google-Extended?

This is the cleanest separation in the ecosystem: you can block Gemini training while keeping full Google Search and AI Overviews visibility. Disallowing Google-Extended does not change how your site appears in Search, but it does limit what Gemini models learn from your content.

How to control Google-Extended

Google-Extended is a robots.txt control token, not a crawler: no separate agent visits your site under this name. Add the rule below to opt out; omitting it (the default) leaves the content usable.

Opt out

User-agent: Google-Extended
Disallow: /

There is no Google-Extended user agent in server logs; the token is honored by Google's crawl infrastructure on content fetched by Googlebot.

Google-Extended FAQ

What is Google-Extended?

Google-Extended is not a separate crawler and has no HTTP user-agent string of its own; it is a standalone robots.txt product token that controls whether content Google crawls from your site may be used to train future generations of Gemini models and for grounding in Gemini Apps and on Vertex AI. Crawling itself is done with existing Google user agents.

How do I block Google-Extended?

Add a robots.txt group with "User-agent: Google-Extended" followed by "Disallow: /". This opts your content out of Gemini training and grounding use but does not stop Google from crawling your pages, and per Google's documentation it does not impact your site's inclusion in Google Search and is not used as a ranking signal.

About AI Training Crawlers

Crawlers that collect content to train foundation models. Data gathered today shapes what future model versions know about your brand.

Blocking these doesn't affect live citations, but it limits what future models learn about you from your own site, leaving third-party sources to fill the gap.

More bots from Google

Googlebot

Google's main indexer: the crawl behind Search, AI Overviews, and AI Mode.

GoogleOther

Google's generic crawler for product teams: research and development fetches outside Search.

Google-Agent

Google agent that fetches pages on behalf of generative experiences.

Google-NotebookLM

Fetches sources a user adds to NotebookLM so it can summarize and cite them.

Google-CloudVertexBot

Crawls site-owner-requested pages to build Vertex AI Search agents.

Googlebot-Image

Indexes images for Google Image Search.

Googlebot-Video

Indexes video content for Google Video Search.

Googlebot-News

Indexes articles for Google News.

Storebot-Google

Crawls shopping content for all Google Shopping surfaces.

Chrome Lighthouse

Lighthouse page-experience and performance auditing.

AdsBot-Google-Mobile

Checks mobile ad landing-page quality for Google Ads.

AdsBot-Google

Checks ad landing-page quality for Google Ads.

Google-Display-Ads-Bot

Verifies site eligibility during the AdSense approval process.

Mediapartners-Google

Crawls participating sites to serve relevant AdSense/AdMob ads.

Google Stackdriver

Google Cloud uptime checks and availability monitoring.

Google-InspectionTool

Search Console URL Inspection and Rich Results Test.

Google-Safety

Abuse-specific crawling such as malware discovery.

Google Site Verifier

Fetches Search Console verification tokens.

Google Image Proxy

Image caching proxy used by Gmail and other Google services.

Chrome Prefetch Proxy

Fetches traffic-advice to enable privacy-preserving prefetch hints.

Google Feedfetcher

Crawls RSS/Atom feeds for Google News and PubSubHubbub.

Google Publisher Center

Fetches feeds publishers supply for Google News landing pages.

Google Read Aloud

Fetches and reads out web pages using text-to-speech on request.

APIs-Google

Delivers push-notification messages for Google APIs.

See which bots visit your site

Temso helps marketers track every AI crawler and bot that visits your site, and runs regular audits to make sure your website is always performing at its full potential.

Explore Website Audits
Website audit dashboard showing bot visits, citation-bot visits, and a table of AI bots crawling the site