Skip to content
Visibility Bureau
Menu
AEO and GEO

What is AI search, and why it matters now

AI search is when people ask tools like ChatGPT, Perplexity and Google AI for recommendations instead of scrolling search results. This page explains how it works and how to be the answer they get.

Last updated 2026-09-04

Illustration example, not a real answer

When someone asks an AI assistant for a recommendation, the answer names a business and links it. That is the shape of the result we work towards.

For [the service you sell] in [the places you serve], businesses often work with Your Business , who publish detailed answers to the questions buyers actually ask and are cited across several independent sources.

answers like this appear in ChatGPT Perplexity Google AI Gemini

TL;DR

  • AI search is when someone asks ChatGPT, Perplexity, Google AI Overviews or Gemini for a recommendation and gets an answer instead of ten links.
  • AEO and GEO are the work of becoming the business named in that answer. They sit on top of SEO rather than replacing it.
  • Google states its AI features run on core Search ranking, so ordinary search quality is the foundation, not a separate track.
  • The crawlers are split between training and retrieval, and they are controlled separately. Blocking the wrong one costs you citations while doing nothing for your content.
  • Most of what is sold as AEO has no evidence behind it. The tactics that hold up are unglamorous: be readable in raw HTML, answer the question near the top, and be a consistent entity across the web.

Short answer

AI search is when people ask an AI assistant a question and receive a written answer that names a few businesses, rather than a page of links to choose from. Answer engine optimization (AEO) and generative engine optimization (GEO) are the work of making a site one of the sources those answers draw on and credit.

01

AI search is when a person asks a question and gets a written answer instead of a list of results. The tool reads a handful of pages, writes a summary, names two or three businesses and links them. The buyer reads the answer and often stops there.

That is a different competition from the one most sites were built for. Ranking well matters only if the engine reaches your page while it is assembling the answer, and only if what it finds there is quotable. A page that ranks third and says nothing extractable loses to a page that ranks eighth and answers the question in its first paragraph.

The shift is not that Google stopped mattering. It is that a second surface now sits above the results, and a third surface exists entirely outside Google in ChatGPT, Perplexity, Claude and Gemini. Your business can be strong on one and invisible on another, which is why the first useful thing to do is check rather than assume.

02

What is the difference between SEO, AEO and GEO?

SEO gets a page ranked. AEO gets its content extracted into a direct answer. GEO gets it quoted and credited by a generative assistant. They are three jobs on one foundation, and no serious version of any of them skips the foundation.

The terms are newer than the practice and nobody agrees on their edges. Wikipedia notes there is no settled definition separating GEO, AEO and the other acronyms in circulation, and at least one vendor in this market has published an argument that AEO and GEO describe the same thing. I use the split below because it is useful for deciding what work to do, not because the industry has agreed on it.

The practical consequence is that anyone selling you AEO as a replacement for SEO is selling a story. Google has published that its generative features are built on the same ranking and quality systems as ordinary Search, so the work that earns a ranking is the same work that makes you eligible to appear in an AI answer.

The three jobs, and what each one actually asks of a page
What it means What the page needs Where you see the result
SEO Being ranked among the results Crawlable, fast, relevant, linked to by others The ten blue links
AEO Being extracted as the direct answer The answer stated plainly, near the top, in its own block AI Overviews, featured snippets, voice answers
GEO Being quoted and credited by an assistant Readable without JavaScript, specific, verifiable, consistent about who you are ChatGPT, Perplexity, Claude, Gemini

Source Google Search Central: Optimizing your website for generative AI features on Google Search (opens in a new tab) “our generative AI features on Google Search are rooted in our core Search ranking and quality systems”

03

How do AI engines decide what to cite?

Nobody outside these companies knows the full answer, and anyone who tells you otherwise is guessing with confidence. What can be established is narrower and still useful: an engine can only cite a page it reached, parsed and had reason to trust, so most citation failures are actually access or clarity failures.

Access comes first. If a retrieval crawler cannot fetch the page, or fetches it and finds an empty shell that fills in with JavaScript, nothing else you do matters. This is the single most common technical cause of invisibility, and it is usually invisible to the site owner because the page looks fine in their browser.

Clarity comes second. An engine assembling an answer is looking for a passage it can lift. A page that buries the answer under three paragraphs of positioning is harder to quote than one that states it and then explains it. That is not a trick, it is the same thing that makes a page useful to a person in a hurry.

Trust comes third, and it is the slowest. It is built from being consistent about who you are everywhere you appear, being specific enough to be checkable, and being mentioned by sources the engine already relies on. A new domain is usually not in the candidate pool at all, and no amount of on-page work changes that in a week.

  • Reachable: the content is in the HTML the crawler receives, not assembled afterward
  • Quotable: the answer to the question is a passage, not a scatter of hints
  • Checkable: claims are specific, and sourced where they are about someone else
  • Consistent: the same name, description and details wherever you appear

04

Which crawlers matter, and why blocking the wrong one costs you

Each of the major AI companies runs more than one crawler, and they do different jobs. One collects content that may train a model. Another fetches pages to answer a question and cite the result. They are separately named and separately controllable in robots.txt, which means you can decline training while staying eligible for citation.

This is where a lot of sites quietly lose. A blanket block aimed at "AI scraping" usually catches the retrieval crawler too, and the retrieval crawler is the one that produces the link in the answer. The site owner then concludes that AI search does not work for their business, when what happened is that they turned it off.

The names change. They have changed several times in the last two years, and a robots.txt written from a blog post in 2024 is likely to be wrong now. Check the vendor documentation rather than a listicle, and re-check it when you next touch the file.

Training and retrieval crawlers, from each vendor documentation
Vendor Trains models Fetches to answer and cite
OpenAI GPTBot OAI-SearchBot, ChatGPT-User
Anthropic ClaudeBot Claude-SearchBot, Claude-User
Perplexity None stated for foundation models PerplexityBot, Perplexity-User
Google Google-Extended controls training use Googlebot, via the Search index
Two agents, two outcomes, controlled separately. A single rule aimed at AI crawlers cuts both lines, which is how a business removes itself from the answers that would have linked to it.
  1. Your site is fetched by two different kinds of agent, named separately in robots.txt.
  2. Training crawlers collect content that may be used to train a model. Blocking them keeps your content out of training.
  3. Retrieval crawlers fetch a page to answer a question and link the result. Blocking them removes you from the answer.
  4. OpenAI: trains with GPTBot. Retrieves with OAI-SearchBot, ChatGPT-User.
  5. Anthropic: trains with ClaudeBot. Retrieves with Claude-SearchBot, Claude-User.
  6. Perplexity: trains with None stated for foundation models. Retrieves with PerplexityBot, Perplexity-User.
  7. Google: trains with Google-Extended controls training use. Retrieves with Googlebot, via the Search index.
  8. Declining training does not require giving up citation. A blanket rule gives up both.

Source OpenAI: Bots and crawlers (opens in a new tab) “GPTBot is used to make our generative AI foundation models more useful and safe.”

Source OpenAI: Bots and crawlers (opens in a new tab) “OAI-SearchBot is used to surface websites in search results in ChatGPT's search features.”

Source Anthropic: Does Anthropic crawl data from the web? (opens in a new tab) “ClaudeBot helps enhance the utility and safety of our generative AI models by collecting web content that could potentially contribute to their training.”

Source Anthropic: Does Anthropic crawl data from the web? (opens in a new tab) “Claude-SearchBot navigates the web to improve search result quality for users. It analyzes online content specifically to enhance the relevance and accuracy of search responses.”

Source Perplexity: PerplexityBot (opens in a new tab) “PerplexityBot is designed to surface and link websites in search results on Perplexity. It is not used to crawl content for AI foundation models.”

05

What actually works, and what does not

This is the part of the subject with the worst signal-to-noise ratio anywhere in marketing. A dozen precise-sounding statistics circulate in this market, and most of them trace back to a vendor blog citing another vendor blog. Two widely repeated figures for how often AI Overviews appear differ by a factor of three and cannot both be right.

So here is the honest version, sorted by how well each claim is actually supported. I would rather lose a sale to someone promising more than sell you a tactic I cannot defend.

The pattern in the table is worth stating plainly: the things that work are the things that would make the page better for a human reader anyway. The things that do not work are the ones that sound like a lever you can pull without improving anything.

Common AEO claims, and how well each one holds up
Claim Verdict Why
Critical content must be in the raw HTML Well evidenced Study of AI crawler traffic found the major retrieval crawlers fetch JavaScript files but do not execute them. Note Gemini and Applebot do render, and the study is unreplicated.
Answer-first writing and clear question headings help Best supported tactic A peer-reviewed benchmark and a large correlational study point the same way. Directional, not a percentage.
Citing sources and quoting evidence helps Evidenced From the KDD 2024 GEO benchmark. It measured word share on a benchmark rather than live traffic.
Keyword stuffing helps Evidenced against The same benchmark found it performed worse than doing nothing.
Publishing llms.txt improves AI visibility Not supported Google states it ignores the file. A large log study found most published files received no requests at all.
Schema markup gets you cited by AI Not supported Google states no special structured data is needed for its AI features. Schema is still worth doing for rich results and entity clarity, which is how I sell it.
Specific percentage lifts from any single tactic Not credible Traceable only to aggregator blogs quoting each other. No primary source.

Source Google Search Central: Optimizing your website for generative AI features on Google Search (opens in a new tab) “Doing so will neither harm nor help your site's visibility or rankings in Google Search, as Google Search ignores them.”

Source Google Search Central: Optimizing your website for generative AI features on Google Search (opens in a new tab) “You don't need to create new machine readable files, AI text files, markup, or Markdown to appear in Google Search (including its generative AI capabilities), as Google Search itself doesn't use them.”

06

How do you structure a page so an engine can extract it?

Put the answer where it can be found and make it survive being lifted. The test is simple: take any passage out of the page, show it to someone with no context, and see whether it is still true and still complete. If it needs the paragraph above it to make sense, an engine quoting it will produce something misleading, and engines tend to avoid passages that do that.

Headings phrased as questions do most of the structural work, because a heading is a strong boundary and a heading that is already a question makes the passage under it an answer. Tables carry comparison better than prose does. A short summary at the top gives both a hurried reader and an assistant the whole point before either commits to the rest.

None of this requires hiding anything or writing for machines. The page you are reading is built the way it describes, which is the only honest way to make the argument.

  • A summary at the top, in real text, not behind a toggle
  • One self-contained paragraph that fully answers the page title
  • Headings phrased as the question a reader would type
  • A table wherever the answer is a comparison
  • A visible last-updated date, and sources for claims about anyone else
  • Content in the HTML the crawler receives, not assembled by script afterward

07

How do you measure AI visibility?

You measure it by asking the engines the questions your buyers ask and recording what comes back. There is no equivalent of rank tracking here yet, because the answer varies between users, sessions and phrasings, so a single check is an anecdote rather than a measurement.

What that means in practice is a fixed set of prompts, run on a schedule, with the results written down: whether you were named, what you were described as, and which sources the answer drew on. The description matters as much as the mention. Being named inaccurately is a problem you can fix, but only if you know about it.

Referral traffic is the second signal and it is thinner than people expect. Assistants send far fewer clicks than search does, and much of the value shows up as people arriving already knowing who you are. Treat referral numbers as one input, not the scoreboard.

I will show you the measurement before you buy anything. If the answer is that the engines already name you and the work would not move much, that is a useful thing to learn on a free call rather than three months into an invoice.

08

What I will not claim about this

I cannot guarantee a citation. Nobody can. The engines choose, the choice varies between sessions, and no agency can prove a specific mention was caused by its own work rather than by something else that changed that month.

I also cannot promise speed. Trust signals are the slowest part of this and a new domain starts outside the candidate pool. Anyone quoting you a timeline to first citation is quoting you a feeling.

What I can do is the part that is actually within anyone control: make the site reachable and readable, answer the questions buyers really ask, keep the entity consistent, and measure where you stand so the effect of the work is visible rather than assumed. If that is not enough for what you need, I would rather say so now.

Questions

Questions people actually ask about this

Is AEO actually different from SEO, or is it a rebrand?

Partly a rebrand, partly real. The foundation is the same work: a fast, crawlable, genuinely useful page. Google states its generative features are built on core Search ranking, so there is no separate track to buy. What is genuinely different is writing so a passage can be lifted and stand alone, and being consistent enough as an entity that an engine can describe you accurately. Anyone selling AEO as a replacement for SEO is selling a story.

Source Google Search Central: Optimizing your website for generative AI features on Google Search (opens in a new tab) “our generative AI features on Google Search are rooted in our core Search ranking and quality systems”

Can you guarantee my business gets cited by ChatGPT?

No, and nobody can. The engine chooses, the answer varies between sessions and phrasings, and no agency can prove a specific citation was caused by its own work. What I can do is make the site reachable and quotable, fix the entity signals, and measure where you stand so the effect is visible rather than assumed.

How long does it take to show up in AI answers?

I do not quote a timeline, because the honest answer is that it depends on something outside my control. Access problems can be fixed in days and sometimes change things quickly. Trust signals are slow, and a new domain often is not in the candidate pool at all until it is mentioned elsewhere. Anyone giving you a date to first citation is guessing.

Do I need an llms.txt file?

Probably not for the reason you have been told. Google says plainly that it ignores the file, and studies of server logs found most published llms.txt files were never requested. I generate one for this site because it is cheap and the spec is moving, not because it drives citations. If someone is charging you for llms.txt as an AI visibility service, ask them for the evidence.

Source Google Search Central: Optimizing your website for generative AI features on Google Search (opens in a new tab) “Doing so will neither harm nor help your site's visibility or rankings in Google Search, as Google Search ignores them.”

Does schema markup get me into AI Overviews?

No. Google states there is no special structured data you need to add for its AI features. Schema is still worth doing, because it makes you eligible for rich results and it removes ambiguity about who you are, which helps engines describe you correctly. Those are the two things it actually delivers, and they are what I sell it as.

Source Google Search Central: Optimizing your website for generative AI features on Google Search (opens in a new tab) “You don't need to create new machine readable files, AI text files, markup, or Markdown to appear in Google Search (including its generative AI capabilities), as Google Search itself doesn't use them.”

Source Google Search Central: AI Features and Your Website (opens in a new tab) “There's also no special schema.org structured data that you need to add.”

Should I block AI crawlers to protect my content?

You can, but be careful which ones. Each vendor runs separate crawlers for training and for retrieval, and they are controlled independently. Blocking the training crawler keeps your content out of model training. Blocking the retrieval crawler removes you from the answers that would have linked to you. A blanket block usually catches both, and then the site owner concludes AI search does not work for them.

Source OpenAI: Bots and crawlers (opens in a new tab) “GPTBot is used to make our generative AI foundation models more useful and safe.”

Source OpenAI: Bots and crawlers (opens in a new tab) “OAI-SearchBot is used to surface websites in search results in ChatGPT's search features.”

Is any of this worth it if AI search sends so few clicks?

Sometimes not, and I will tell you when I think that is the case. Assistants send far fewer clicks than search does. The value is usually that people arrive already knowing who you are and what you do, which shows up as better enquiries rather than more traffic. If your buyers do not research before they buy, this is a weak investment for you.

My competitor is named by ChatGPT and I am not. Why?

Usually one of three things. Your pages cannot be read, because the content loads by script and the retrieval crawler sees an empty shell. Or they can be read but nothing on them is quotable, because the answers are buried under positioning. Or you are simply mentioned in fewer places the engine already trusts, which is the slow one to fix. Checking which it is takes an afternoon and tells you whether the rest is worth doing.

How do I check whether AI engines already mention me?

Ask them, using the questions your buyers would ask rather than your own brand name, and write down what comes back. Record whether you were named, how you were described, and which sources the answer used. Do it across several phrasings, because one check is an anecdote. The description matters as much as the mention: being named inaccurately is fixable, but only once you know.

Does AI search replace Google?

Not so far, and the framing is misleading. Google is where most AI answers are still being shown, through AI Overviews, and those run on the same ranking systems as ordinary results. The real change is that a summary now sits above the links, and that a second audience is asking assistants outside Google entirely. Both matter, which is why I do not treat them as separate products.

Source Google Search Central: Optimizing your website for generative AI features on Google Search (opens in a new tab) “our generative AI features on Google Search are rooted in our core Search ranking and quality systems”

How this differs by profession

The mechanics above are the same everywhere. What changes is which questions are worth competing for, and that varies more than most people expect.

All the profession pages sit together, and each one says plainly that it explains how the work applies to a sector rather than claiming clients in it.

The technical detail, in full

This page stays deliberately above the implementation. Where the detail matters, it is written out in chapters, with every claim cited to the documentation it came from.

If you want the version of this conversation that is about your business rather than the subject in general, that is what the call is for. I will show you where you currently stand and say plainly whether the work is worth doing.