Skip to content

GEO in 2026: Get Your Brand Cited by ChatGPT, Gemini and AI Mode

A practical GEO playbook for 2026: how ChatGPT, Perplexity, Gemini and Google AI Mode pick sources, plus crawler rules, llms.txt and GEO tracking tools.

GEO8 min read
By the AI App Hunters editors

To get cited by ChatGPT, Perplexity, Gemini and Google AI Mode, you need three things: crawlers that can reach your pages, content that answers specific questions in plain text, and brand facts that match everywhere they appear. There is no secret file or markup. Then measure it with a prompt tracking tool, because your Google rankings will not tell you what AI engines say about you.

Key takeaways

  • Google says there are no extra requirements or special optimizations for AI Overviews or AI Mode. If a page is indexed and eligible for a snippet, it can be used.
  • Search bots and training bots are separate. You can allow OAI-SearchBot, PerplexityBot and Claude-SearchBot while blocking GPTBot and ClaudeBot.
  • Google-Extended does not affect Google Search inclusion or ranking. Blocking it will not take you out of AI Overviews.
  • llms.txt is a proposal, not a standard. Google says you do not need AI text files to appear in its AI features.
  • Track prompts, not keywords. Tools like Peec AI, Otterly.AI and Profound show how often you are mentioned and cited.

How do AI answer engines pick their sources?

Each engine works a little differently, but the pattern is the same: the model rewrites your question into several searches, pulls candidate pages, and cites the ones that answer cleanly.

Google AI Overviews and AI Mode. Google's documentation says both "may use a 'query fan-out' technique, issuing multiple related searches across subtopics and data sources" to build a response. That is why AI Mode often links to a wider set of pages than the classic blue links. The entry ticket is plain: a page "must be indexed and eligible to be shown in Google Search with a snippet."

ChatGPT search. OpenAI's help center says ChatGPT may "rewrite your query into one or more targeted queries" and send them to search partners, naming Microsoft Bing. Results are ranked "using multiple factors intended to help users find relevant, reliable information." To be eligible, OpenAI says to allow OAI-SearchBot and make sure your host or CDN does not block its published IP ranges. Placement is not guaranteed.

Perplexity. PerplexityBot exists "to surface and link websites in search results on Perplexity." Perplexity recommends allowing it in robots.txt if you want to be indexed.

Claude. Anthropic runs Claude-SearchBot to improve search results for Claude users. Disallowing it reduces your visibility in those results.

The practical read: query fan-out rewards depth. Clear pages for each sub-question can be picked several times in one answer.

What content structure gets cited?

Write pages that a model can quote without rewriting. In practice, that means:

  • Lead with the answer. Put a two or three sentence direct answer at the top. Models pull from passages that stand on their own.
  • Use questions as headings. Match the way buyers phrase things: "How much does X cost?", "Is X better than Y for small teams?"
  • Keep facts in text. Google lists "making sure that important content is available in textual form" as a best practice. Prices locked in images or tabs that only render on click are easy to miss.
  • Add a comparison table when readers are choosing between options. Tables are compact and easy to extract.
  • Link internally. Google also calls out "making your content easily findable through internal links." Connect your pillar page to each sub-question page and back.
  • Date and update. Pricing, plan names and feature lists go stale fast. Show a last-updated date and keep it honest.

Why does entity consistency matter?

AI engines assemble an answer from many sources. If your homepage says one thing, your LinkedIn page another, and a three-year-old directory listing a third, the model has to guess. It often guesses wrong, or skips you.

Fix the basics first:

  • One canonical company name, plus any alternate names you actually use.
  • The same one-line description of what you do on your site, social profiles, review sites and directories.
  • Current pricing and plan names everywhere, or no price at all on third-party pages you cannot update.
  • A clear "About" page with who you are, where you are based and what you sell.

Third-party mentions matter because engines like Google and ChatGPT pull from the open web, not just your site. Review sites, directories, partner pages and credible comparison articles all shape what the model believes about you.

Which structured data should you add?

Structured data will not get you cited on its own. Google is direct about it: "There's also no special schema.org structured data that you need to add" for AI features. But it still helps search engines understand who you are.

Start with Organization markup on your homepage. Google says it helps "disambiguate" your organization from others. The useful properties are name, url, logo, description, alternateName and sameAs, which points to your profiles on other sites. That last one directly supports entity consistency.

One rule from Google matters more than any schema type: "Making sure your structured data matches the visible text on the page." Markup that claims something the page does not show is a liability.

Should you publish an llms.txt file?

Be clear on its status. llms.txt is "a proposal to standardise on using an /llms.txt file to provide information to help agents use a website," written by Jeremy Howard in 2024. It is a Markdown file at your site root with a title, a short summary and lists of links to key pages. The spec says it is open for community input.

It is not a standard Google uses. Google's AI features guide says: "You don't need to create new machine readable files, AI text files, or markup to appear in these features."

Our take: it is cheap to add, but do not expect it to move your AI search visibility. Fix your actual pages first.

How should you configure robots.txt for AI bots?

This is where most teams make a mistake: they block "AI bots" as a group and quietly remove themselves from AI search. Training bots and search bots are separate user agents. Here is what each one does, per the vendors' own documentation:

User agent Owner What it does Obeys robots.txt
GPTBot OpenAI Crawls content that may be used to train foundation models Yes
OAI-SearchBot OpenAI Surfaces sites in ChatGPT search results Yes
ChatGPT-User OpenAI Visits pages on a user's request Rules "may not apply"
PerplexityBot Perplexity Surfaces and links sites in Perplexity search, not used for model training Yes
Perplexity-User Perplexity Fetches pages for a user's question Generally ignores it
ClaudeBot Anthropic Collects content that could contribute to training Yes
Claude-SearchBot Anthropic Improves Claude search results Yes
Claude-User Anthropic Fetches pages when users ask Claude Yes, blocking reduces visibility
Google-Extended Google Controls use of content for Gemini training and grounding Yes, a product token

OpenAI says each setting is independent, so you can allow search and block training. A common setup for brands that want visibility but not training use:

User-agent: OAI-SearchBot
Allow: /

User-agent: PerplexityBot
Allow: /

User-agent: Claude-SearchBot
Allow: /

User-agent: GPTBot
Disallow: /

User-agent: ClaudeBot
Disallow: /

Two notes on Google. First, Google-Extended "does not impact a site's inclusion in Google Search nor is it used as a ranking signal." AI Overviews and AI Mode run on normal Googlebot crawling, so blocking Google-Extended does not opt you out of them. Second, if you want to limit what Google shows from a page in AI features, Google points to the existing controls: nosnippet, data-nosnippet, max-snippet or noindex.

Also check your CDN and firewall. Bot protection rules can block these crawlers even when robots.txt allows them. OpenAI publishes IP lists for its bots, and Anthropic notes that blocking by IP alone can stop its crawler from reading your robots.txt at all.

How do you measure GEO?

Start with what you already have. Google says traffic from AI Overviews and AI Mode is included in Search Console's Performance report under the "Web" search type. It is not broken out separately, so pair it with analytics to see conversions.

For the rest, you need a prompt tracker: you list the questions buyers ask, the tool runs them across engines on a schedule, and it reports mentions, citations and competitors. Here is how the main options compare (prices as of October 2026, from each vendor's site):

Tool Engines covered Starting price Good fit
Otterly.AI ChatGPT, AI Overviews, Perplexity, Copilot, with Gemini, AI Mode and Claude as add-ons Lite at $29/month for 15 prompts Small teams starting out
Peec AI Choose 3 models on standard plans, including ChatGPT, AI Mode, AI Overviews, Gemini, Copilot Starter includes 50 prompts, see site for price Marketing teams tracking competitors
Rankscale ChatGPT, Gemini, Perplexity, Claude, Copilot and more Essentials at $17/month (yearly only), Pro at $99/month Agencies, with page audits and white label
Ziptie 7 engines incl. ChatGPT, AI Mode, AI Overviews, Gemini, Copilot Usage based, about $0.01 per check Teams who want to pay per check
Profound ChatGPT, Perplexity, Claude, Gemini, Copilot, DeepSeek, AI Overviews Contact sales Enterprise brands that want agents to act on findings
Mangools AI Search Grader Up to 7 models Free, with limits A quick one-off check
Cito ChatGPT, Perplexity, Gemini, AI Overviews Free self-hosted (MIT license), Growth at $599/month Technical teams who want open source

A few specifics worth knowing. Profound adds Agent Analytics, which tracks "how bots view, cite, and refer to your pages," and an AI marketer agent that drafts work for approval. Ziptie integrates with Search Console and scores mentions, citations and sentiment. Otterly.AI and Peec AI both offer 15% off annual billing and free trials. Rankscale includes page audits on every plan, from 50 on Pro.

What to look for: engine coverage that matches where your buyers search, citation tracking (not just mentions), competitor share of voice, and exports you can put in a monthly report.

How to choose your first GEO moves

If you do nothing else this quarter, do this, in order:

  1. Audit robots.txt and your CDN rules. Make sure the search bots above can reach you.
  2. Fix your entity basics: one name, one description, current pricing, Organization markup with sameAs.
  3. Rewrite your top ten commercial pages to open with a direct answer and use question headings.
  4. Build pages for the sub-questions buyers ask, and link them to your pillar page.
  5. Set up a tracker with 20 to 50 real buyer prompts and check it monthly.

Skip the shortcuts. No file, schema type or prompt trick replaces being the clearest, most current source on the question.

Frequently asked questions

What is GEO (generative engine optimization)?+

GEO is the work of getting your brand mentioned and cited in answers from AI engines such as ChatGPT, Perplexity, Gemini and Google AI Mode. It builds on SEO basics: crawlable, indexable, clearly written pages, plus consistent facts about your brand across the web.

Do I need llms.txt to appear in AI answers?+

No. llms.txt is a community proposal by Jeremy Howard, not an adopted standard. Google states you do not need new machine readable files or AI text files to appear in AI Overviews or AI Mode.

Does blocking Google-Extended remove my site from AI Overviews?+

No. Google says Google-Extended does not affect inclusion in Google Search and is not a ranking signal. It controls whether crawled content is used for Gemini training and grounding in Gemini apps and Vertex AI.

Which robots.txt user agents matter for AI search visibility?+

For search-style visibility, allow OAI-SearchBot (ChatGPT search), PerplexityBot (Perplexity search) and Claude-SearchBot (Claude search). GPTBot and ClaudeBot relate to model training and can be blocked separately without removing you from those search features.

How do I measure GEO results?+

Use a prompt tracking tool such as Peec AI, Otterly.AI, Rankscale or Profound to see how often you are mentioned and cited. Google counts AI Overviews and AI Mode traffic inside the Web search type of the Search Console Performance report.

Sources, checked 1 Oct 2026

  1. AI features and your website, Google Search Central
  2. Google's common crawlers (Google-Extended)
  3. Organization structured data, Google Search Central
  4. Overview of OpenAI crawlers
  5. ChatGPT search, OpenAI Help Center
  6. Perplexity crawlers
  7. Does Anthropic crawl data from the web, Claude Help Center
  8. The /llms.txt file
  9. Profound
  10. Peec AI pricing
  11. Otterly.AI pricing
  12. Rankscale pricing
  13. Ziptie
  14. Mangools AI Search Grader
  15. GetCito

Keep reading

App of the Week6 min read

App of the Week: Attio, the AI-native CRM

Attio is an AI-native CRM with agents, workflows and an MCP server. What it does, who it suits, October 2026 pricing, limits and the best alternatives.

Attio logo
App of the Week5 min read

App of the Week: HeyGen, AI Avatar Video

HeyGen makes avatar videos from text and translates video into 175+ languages. What it does, October 2026 pricing, limits and the best alternatives.

HeyGen logo
App of the Week4 min read

App of the Week: Profound, AI Search Visibility

Profound tracks how your brand shows up in ChatGPT, Claude, Perplexity and Gemini, then runs agents to improve it. Features, pricing, limits, alternatives.

Profound logo