SEO

Google-Extended

Also called Google-Extended robots.txt token

A robots.txt control that lets a site opt out of Google using its content for Gemini products, without leaving Search.

Quick facts: Google-Extended

Category
SEO
Also called
Google-Extended robots.txt token
Level
Advanced
Affects
Use of your content in Gemini products, publisher licensing decisions
Where to see it
robots.txt, Search Console, your site's meta robots tags
In this article4
  1. How Google-Extended works
  2. Why Google-Extended matters
  3. Common mistakes with Google-Extended
  4. What to do about it

How Google-Extended works

Google-Extended is not a crawler. It is a token you can name in robots.txt to tell Google whether content already fetched by Googlebot may be used to help improve Gemini apps and the grounded answers in Google’s AI developer products. No separate bot arrives at your server under that name, and no line appears in your access log for it.

That is the entire mechanism, and it explains why the control behaves oddly compared with the crawler tokens people are used to. You are not turning off a visit. You are declining one particular use of something the ordinary search crawler has already collected.

Why Google-Extended matters

It exists so a publisher can stay in Google Search while opting out of a single downstream use. Before it, the only lever available was blocking Googlebot, which meant leaving the index altogether — a price almost nobody would pay. Separating the two was a genuine improvement for anyone with a licensing position to protect.

For most small businesses the honest answer is that it changes very little. If your aim is to be found, quoted and recommended, you have no reason to reach for it. Its value belongs to publishers whose content is itself the product.

Common mistakes with Google-Extended

The big one is believing it controls AI Overviews. It does not. AI Overviews are a feature of Google Search, so they follow the ordinary search controls — whether Googlebot can crawl the page, whether it is indexed, and the snippet directives you set with a meta robots tag. Disallowing Google-Extended in the hope of disappearing from AI Overviews is a wasted change.

The second is fearing a penalty. Using it has no effect on crawling, indexing or ranking in Google Search, because it has no relationship with Googlebot at all.

The third is confusing it with other companies’ crawlers. A rule for Google-Extended says nothing about GPTBot or any other operator’s agent. Every vendor’s control is separate, and there is no single switch for all of them.

What to do about it

Ask one question: is my content itself the thing I sell? A news archive, a paid research library or a licensed database has a real reason to decline. A services business that wants to be recommended does not, and using the control gains it nothing at all.

If you do want it, add a user-agent group named Google-Extended to your robots.txt and disallow the paths concerned, remembering that a group takes only the rules written inside it. Then keep the snippet controls separate in your mind. Those are what govern how much of your page can be displayed in Search and its AI features, and those are the ones to review if display, rather than model training, is your actual concern.

Do and do not

Do

  • Use it only if your content is the product
  • Give it its own robots.txt user-agent group
  • Use snippet controls to limit what Search displays

Do not

  • Expect it to remove you from AI Overviews
  • Fear a ranking penalty from using it
  • Assume it covers other companies' crawlers

Questions people ask about this

Does Google-Extended stop me appearing in AI Overviews?

No. AI Overviews are part of Google Search and follow the same controls as any other result: whether Googlebot can crawl the page, whether the page is indexed, and the snippet directives you have set. Google-Extended governs a different use of your content inside Google's Gemini products. Blocking it changes nothing about what Search displays.

Will using Google-Extended hurt my rankings?

No. It is not a crawler and it has no relationship with Googlebot, so crawling, indexing and ranking are all unaffected. The only thing that changes is whether Google may use your content to help improve certain Gemini products. If a supplier calls it a ranking risk, they have confused it with blocking the search crawler.

Where do I put the Google-Extended rule?

In the robots.txt file at the root of your domain, as its own user-agent group with the excluded paths listed underneath. Crawlers and controls read only the rules inside the group that names them, so do not rely on your general wildcard rules carrying over. Load the file from your live domain afterwards to confirm it is served as written.

Related terms

Found this useful?

Share it, or ask an AI to summarise it

Back to the glossary

Knowing the term is the easy part

Applying it to your own site and budget is the work. Book a call and I will tell you what actually applies to you.