Skip to content

What It Means to Show Up in ChatGPT, Google AI, and Perplexity Answers

What your website needs so an AI assistant can cite it, what is in your control, and what nobody can promise you.

Simplixity teamPublished Leer en español

A customer mentions they asked ChatGPT for a plumber, a bakery, or an accountant near them, and your business wasn't in the answer. Then an email lands in your inbox promising to "get you ranked in AI." You're left wondering what's real and what you can actually do about it.

What happens when someone asks an assistant a question

ChatGPT, Perplexity, and Google's AI answers (AI Overviews and AI Mode) can search the web before they reply. When they do, they sometimes show links to the pages they pulled information from. "Showing up" means your page is one of those linked sources.

For that to happen, a program run by that company has to be able to reach your site and read it. These programs are called crawlers: bots that visit web pages so the company can store and use them later.

What each one needs

Google

In its page AI features and your website, Google says there are no extra requirements to appear in its AI answers. Your page needs to be indexed, which means stored in Google's search system. It also needs to be eligible to show a text snippet in regular search results. Google adds that you don't need special AI files or new code.

Google's guide to generative AI features says the usual search best practices still apply, because these AI answers run on the same core search systems. It also says you don't need to chop your content into tiny pieces for AI to understand it.

ChatGPT

OpenAI runs several crawlers, and each has a different job, according to its crawler overview:

  • OAI-SearchBot is the one that lets your site appear in ChatGPT search.
  • GPTBot collects content that may be used to train OpenAI's models.
  • ChatGPT-User visits a page when a person asks ChatGPT to.

Blocking GPTBot tells OpenAI not to use your content for training. That's a separate choice from appearing in search. OpenAI says changes to your rules can take about 24 hours to take effect.

Perplexity

Per Perplexity's documentation, PerplexityBot shows and links websites in its results and is not used to train AI models. Perplexity-User visits pages when someone asks a question, and the company says it generally ignores robots.txt rules.

What robots.txt is and why it matters

robots.txt is a plain text file on your website that tells crawlers which pages they can visit. Google explains how it reads that file in its robots.txt specification. If a search crawler is blocked there, your site can't show up in that tool, no matter how good your content is.

Allowing AI training is a separate decision. Google shows this with Google-Extended, the setting that controls whether your content is used for its Gemini models. Its crawler list says Google-Extended doesn't affect whether your site is included in Google Search and isn't used as a ranking signal.

What nobody can guarantee

  • Google's guide says a page that meets every requirement still might not be crawled, indexed, or shown. In Google's words, indexing and serving aren't guaranteed.
  • Google also warns that no outside tool has access to its internal ranking or AI systems. If someone offers you inside numbers from Google, be skeptical.
  • No file puts you in AI answers. There's a proposal called llms.txt, a file that summarizes your site for AI tools. Google names llms.txt in its guide and says its search ignores these files, so they neither help nor hurt there. There's no published evidence that they help with other assistants.
  • Nobody can promise you a position, a number of citations, or a set amount of traffic from an AI assistant.

If someone offers to "get you into ChatGPT," ask them to list the specific changes they'll make to your website and how you can check them.

What is in your control

Here's what you can check:

  • Your important pages are indexed in Google.
  • Your robots.txt doesn't block the search crawlers you care about.
  • The pages you want cited don't carry a nosnippet tag. That tag tells Google not to show pieces of your text, and Google lists it among the controls that also apply to its AI answers.
  • Your content answers with real details about your business: what you do, where you work, your hours, how to hire you, and what's included. Google recommends original content, not generic text anyone could have written.

Steps for this week

  1. Open Google Search Console, Google's free tool for site owners, and check whether your main pages are indexed. If you don't have an account, set one up with your business email.
  2. Type your website address followed by /robots.txt into your browser. Look for the names Googlebot, OAI-SearchBot, and PerplexityBot. If you see "Disallow: /" under one of them, that crawler is blocked from your whole site.
  3. Decide separately whether you allow training: GPTBot for OpenAI and Google-Extended for Google. Blocking them does not remove you from Google Search.
  4. Ask whoever manages your website to confirm there's no nosnippet or noindex tag on the pages you want seen. A noindex tag tells Google not to store the page.
  5. Write clear answers to three questions customers ask you often, with specific details, and put them on the page where people look for that information.
  6. In Search Console, open the Generative AI performance report (Google explains it here) and see how your content is doing.

If you'd like help reviewing your website, send us a message.

Sources

  1. AI features and your websiteGoogle Search Central. accessed
  2. Optimizing your website for generative AI features on Google SearchGoogle Search Central. accessed
  3. Overview of OpenAI CrawlersOpenAI. accessed
  4. Perplexity CrawlersPerplexity. accessed
  5. Google's common crawlersGoogle Search Central. accessed
  6. How Google interprets the robots.txt specificationGoogle Search Central. accessed
  7. The /llms.txt filellmstxt.org. accessed

Back to the blog