Which search index does each AI assistant use?

Google's AI features use Google's index, Copilot uses Bing, Perplexity uses its own, ChatGPT uses third-party providers plus its own crawler, and Claude is reported to use Brave. Each has its own crawler rules, so being indexed in Google alone does not make you retrievable everywhere.

In short

  • Google's documentation says AI Overviews and AI Mode need nothing beyond a page being indexed and eligible for a snippet in Google Search.
  • OpenAI's documentation says sites that disallow OAI-SearchBot "will not be shown in ChatGPT search answers", while GPTBot only governs training.
  • Microsoft's February 2026 AI Performance report covers Copilot, AI summaries in Bing and some partner products, which ties Copilot visibility to the Bing index.
  • Claude's web search is reported to run on Brave Search, based on one analysis of Anthropic's supplier list, which Anthropic's product documentation does not confirm.
  • In Ahrefs' September 2025 study of short keywords, Perplexity matched Google's top 10 for 65% of its citations, against 10% for ChatGPT.

Each assistant searches a different copy of the web. Google’s AI features use Google’s index. Copilot uses Bing’s. Perplexity built its own. ChatGPT combines third-party search providers with its own crawler. Claude is reported to use Brave Search. If your pages are missing from one of those indexes, that assistant cannot cite you, whatever your Google rankings are.

How AI search chooses sources has a short version of this table. This lesson goes through each assistant: what is officially documented, what is only reported, which crawler to allow and what to check.

The summary

Assistant Index it retrieves from Search crawler to allow How sure
Google AI Overviews and AI Mode Google’s index Googlebot High. Google documentation
Gemini Google Search, through a grounding feature Googlebot High for the API, lower for the consumer app
ChatGPT search Third-party providers plus OpenAI’s own crawl OAI-SearchBot, and Bingbot High for the official part, medium for the Google report
Microsoft Copilot Bing’s index Bingbot Medium to high. Microsoft’s own reporting ties the two together
Perplexity Its own index PerplexityBot High. Perplexity documentation
Claude Brave Search, reportedly Brave’s crawler, and Claude’s own agents Medium. One secondary analysis

Google AI Overviews, AI Mode and Gemini

Google’s documentation on AI features says there are no extra requirements. A page needs to be indexed and eligible to show with a snippet in Google Search. It also says you do not need new machine-readable files, AI text files or markup. Google’s March 2025 announcement of AI Mode also names the Knowledge Graph and shopping data as sources.

For Gemini, Google’s API documentation describes “grounding with Google Search”: the model writes search queries, runs them on Google Search and returns an answer with source annotations. That is confirmed for developers. That the consumer Gemini app works the same way is widely assumed, and we have not verified it in Google’s documentation.

What to check: ordinary Google indexing in Search Console. Nothing else is needed.

OpenAI’s crawler documentation, read on 4 October 2026, lists separate agents with separate jobs.

  • OAI-SearchBot is used to surface websites in ChatGPT’s search features. Sites that disallow it “will not be shown in ChatGPT search answers”.
  • GPTBot collects content for training models. Disallowing it has no effect on search answers.
  • ChatGPT-User fetches a page when a user’s request needs it. OpenAI says robots.txt rules may not apply, because the action is started by a user.

On where results come from, OpenAI says it uses third-party search providers as well as its own crawl. Bing was the launch partner. The Information reported in August 2025 that OpenAI also used Google results obtained through the scraping service SerpApi, and Google sued SerpApi on 19 December 2025. OpenAI’s documentation does not confirm the Google part, so treat it as a report.

What to check: that robots.txt allows OAI-SearchBot, and that your key pages are indexed in Bing as well as Google.

Microsoft Copilot

Copilot draws on Bing. The clearest evidence is Microsoft’s own reporting. Its February 2026 blog post introduced an AI Performance report inside Bing Webmaster Tools that counts citations across “Microsoft Copilot, AI-generated summaries in Bing, and select partner integrations”, and lists the grounding queries used to retrieve the cited pages.

What to check: verify the site in Bing Webmaster Tools, submit a sitemap and look at the AI Performance report. Many sites that are careful about Google have never opened it. Links in Bing, Yandex, Baidu and Naver explains how Bing weighs links.

Perplexity

Perplexity’s documentation describes two agents. PerplexityBot builds its index. Perplexity-User fetches pages live when a user asks, and the documentation says it generally ignores robots.txt for that reason. Perplexity has offered its index to developers through a search API since September 2025, which VentureBeat reported at launch.

What to check: that PerplexityBot is not blocked. Peec AI’s March 2026 research notes that Facebook blocks Perplexity’s bots, which is why Facebook pages are absent from its citations.

Claude

Anthropic’s supplier list has included Brave Search since 19 March 2025, according to an analysis by the agency Xponent21. The same analysis says a vector and full-text search store called TurboPuffer was added under web search on 6 May 2026, and that its role is not confirmed. It also says Brave only crawls pages that Googlebot is allowed to crawl. All of this comes from one secondary source, so it is reported and not documented.

What to check: search for your brand and key pages in Brave Search. If Googlebot can crawl them, Brave’s crawler should be able to as well.

Search bots and training bots differ

The most common mistake is to block the wrong agent.

Purpose Examples Effect of blocking
Building a search index Googlebot, Bingbot, OAI-SearchBot, PerplexityBot You drop out of that assistant’s answers
Collecting training data GPTBot, Google-Extended Your content is not used for training. Search answers are unaffected
Fetching for a user in real time ChatGPT-User, Perplexity-User, Claude-User May not follow robots.txt at all

Check both robots.txt and any firewall or CDN rules. A bot-blocking setting at the CDN can stop a search crawler that robots.txt allows.

What this means for citations

Because the indexes and merge steps differ, the same prompt yields different sources on each assistant.

  • In Ahrefs’ August 2025 study of 15,000 long-tail prompts, the share of cited URLs found in Google’s top 10 was 8% for ChatGPT, 8% for Gemini, 8.6% for Copilot and 28.6% for Perplexity.
  • In the September 2025 short-keyword study, ChatGPT’s figure was 10% and Perplexity’s 65%. Domain-level overlap between ChatGPT and Google was 31.8%.
  • Semrush reported in July 2025 that about 90% of the time, ChatGPT cited pages ranking at position 21 or lower for related queries.

These are vendor studies and they describe a moving target. The practical reading is that a strong Google position helps most on Perplexity and on Google’s own surfaces, and least on ChatGPT, where breadth of presence across many sites matters more. The hidden searches behind each answer are covered in query fan-out.

A short checklist

  1. Robots.txt allows Googlebot, Bingbot, OAI-SearchBot and PerplexityBot.
  2. CDN and firewall rules do not block those agents.
  3. The site is verified in both Search Console and Bing Webmaster Tools, with sitemaps submitted.
  4. Key pages appear when you search for them in Google, Bing and Brave.
  5. Decide on training bots separately. That is a content policy choice, not a visibility one.
  6. Skip llms.txt for this purpose. SE Ranking’s November 2025 study of 300,000 domains found no relationship between the file and AI citation frequency.

Being in every index is only the entry condition. What gets you cited once you are in is coverage on the third-party pages assistants retrieve, which is the subject of link building for AI visibility and of digital PR.

Where to go next

To see which assistants name you today, read how to measure AI visibility. The AI visibility tools comparison lists trackers by the assistants they cover, and the glossary defines terms such as index tier and crawlable link.

Common questions

Does ChatGPT use Google or Bing?

OpenAI's documentation says ChatGPT search uses third-party search providers and its own crawler, OAI-SearchBot. Bing was the launch partner. The Information reported in August 2025 that Google results obtained through a scraping service were also used, which OpenAI has not confirmed in its documentation.

What search engine does Claude use?

Brave Search, according to one analysis of Anthropic's published list of suppliers, where Brave has appeared since March 2025. That is a secondary source, so treat it as reported.

Does blocking GPTBot remove my site from ChatGPT answers?

No. OpenAI's documentation says GPTBot controls use of content for model training. Appearing in ChatGPT search depends on OAI-SearchBot, which is a separate agent with its own robots.txt rule.

Do I need to submit my site to each AI assistant?

No. There is no submission process. You need to be crawlable by each assistant's search bot and indexed by the search engines they draw on, which in practice means Google and Bing at minimum.

Does Perplexity have its own index?

Yes. Perplexity's documentation describes its own crawler, PerplexityBot, and the company has offered its index through a search API since September 2025.