Google-Extended user agent
Operated by Google · Model training · Learning from you
Google-ExtendedGoogle-Extended is a robots.txt control token, not a crawler. No request ever arrives with Google-Extended in its User-Agent header, so grepping your access logs for it will always return nothing. Google reads the token from your robots.txt and applies it to content its ordinary crawler has already fetched.
Allow or block Google-Extended in robots.txt
Robots.txt applies the most specific matching group, so a group naming Google-Extended overrides your User-agent: * rules for this bot alone.
User-agent: Google-Extended Allow: /
User-agent: Google-Extended Disallow: /
What blocking costs you: Won't train Gemini or feed AI Overviews. Does NOT remove you from Search — Googlebot is separate.
To see which AI crawlers your robots.txt allows right now, run your domain through the AI crawler access checker. To check whether AI answers actually cite you, use the AI citation checker.
Google's other crawlers
Blocking one Google bot does not block the others — each token is a separate group.
- GoogleOther — Research
Google's own documentation: https://developers.google.com/search/docs/crawling-indexing/overview-google-crawlers#google-extended
FAQ
What is the Google-Extended user agent?
Who operates Google-Extended and what does it collect?
How do I block Google-Extended?
Should I block Google-Extended?
Can Google-Extended be spoofed?
Paste a full User-Agent line and identify any crawler with the AI bot user agent list and lookup, or see every token side by side.