AI Visibility Explained: Citations vs Mentions, and Why Most AI SEO Advice Is Wrong

Ishant

Ishant

Published : September 12, 2026 at 4:46 pm

Updated : September 12, 2026 at 4:46 pm

Ask ChatGPT to name the best agency in your category. If your name does not come back, you have an AI visibility problem. Ask it about your brand by name and watch it describe you perfectly. Now you have two problems, because those are not the same problem, and almost every article on this subject treats them as one.

This guide explains what is actually happening inside these systems, in plain language, with the evidence attached. It covers the difference between a mention and a citation, how a brand ends up in an AI answer step by step, what the data says works, and the popular advice that measurably does nothing.

We test this on our own agency and on client accounts every month. We have taken a client from zero AI presence to citations across ChatGPT, Gemini and Google AI Overviews alongside 2.4x organic traffic, and grown another 117 percent in organic clicks while building AI visibility from scratch. Some of what follows contradicts what you will read elsewhere. Where it does, the source is linked so you can check it yourself.

What is AI visibility, in one sentence?

AI visibility is how often your brand is named, linked, or recommended inside an AI generated answer.

That is it. Not a score. Not a ranking position. It is presence inside a paragraph of text that a machine wrote for a person who asked a question. If you want the wider framework around this, we cover it in our complete guide to generative engine optimization.

The reason it matters is that ranking first on Google no longer guarantees you appear. Semrush found that only 44.3 percent of pages in Google’s top 10 also appear in AI generated answers. Ahrefs went further and measured how many AI cited URLs rank in Google’s top 10 for the original query at all:

PlatformCited URLs that rank in Google’s top 10
Perplexity28.6%
Gemini8.6%
Copilot8.2%
ChatGPT8.0%
Average11.9%

Roughly 80 percent of AI citations do not rank anywhere in Google for the question that produced them. That single table is why “just do SEO” is an incomplete answer, even though, as you will see later, it is closer to the truth than most of the alternatives.

What is the difference between an AI mention and an AI citation?

This is the concept everything else depends on, and it is the one nobody explains properly.

The definitions

MentionCitation
What it isYour brand name appears in the words of the answerYour URL appears as a linked source
Where it comes fromThe model’s internal knowledge, or from somebody else’s page it readRetrieval. A page was fetched, split up and used as evidence
What it looks like“For Shopify stores, agencies like Acme are often recommended”A little numbered link, or a source card under the answer
Commercial valueHigh. This is the recommendationLower on its own, but it brings traffic and it proves the claim
What moves itHow often your name shows up next to your category on other people’s contentCrawlability, page structure, passage quality, being in the index

They come apart constantly, and the direction flips by engine

Semrush and Kevin Indig ran 115 prompts across 14 countries and logged 3,981 domain appearances. Orbit Media tracked 13,184 citations across four models over 12 weeks. Put their numbers side by side and the picture is strange:

EngineCitation rateMention rateWhat that means
ChatGPT87%20.7%Links to sources constantly, rarely names brands
Gemini21.4%83.7%Names brands constantly, rarely links
Claude40%70%Names more than it links
Perplexity44%67%Names more than it links

ChatGPT and Gemini are close to mirror images of each other.

Now hold that next to the single most useful number in the whole study:

61.7 percent of all appearances were “ghost citations”: the page was cited, but the brand was never named. 25.1 percent were the reverse: mentioned with no link at all. Only 13.2 percent were both.

Why this happens, mechanically

Ghost citation, cited but not named. Your page answered a factual sub-question well. The model used it as evidence for a sentence, attached your link, and moved on. It never needed to say who you were. You supplied the raw material and somebody else got recommended in the same paragraph.

Mentioned but not cited. The model already knew your name, either from training or from a third party page it read, and named you without linking. No traffic. But the buyer now has your name, and the next thing they do is search it.

Which one should you chase? If nobody in your market has heard of you, chase mentions. Mentions are the recommendation. Citations are evidence supply. Most agencies sell “citations” because they are easier to screenshot.

The practical consequence almost everyone misses

These two metrics have different levers.

  • Citations respond to things on your website: can it be crawled, is the passage self-contained, is the page fast and in the index.
  • Mentions respond to things on other people’s websites: how often your brand name appears next to your category across the open web.

Ahrefs measured this across 75,000 brands and 76.7 million AI Overviews:

FactorCorrelation with AI Overview brand visibility
Branded web mentions0.664
Branded anchors0.527
Branded search volume0.392
Domain Rating0.326
Referring domains0.295
Organic traffic0.274
Backlinks0.218

Mentions of your name beat links to your site by three to one. That is the whole strategy in one row, and it is the reason an unlinked mention of your brand can be worth more than a link. If you are still weighing link types, our guide to nofollow links and when they still matter covers the related ground.

How does a brand actually end up in an AI answer?

Here is the full pipeline, step by step, with the evidence for each stage.

Step 1: The AI decides whether to search at all

Most people assume every question triggers a web search. It does not. It depends on intent.

Prompt typeHow often ChatGPT runs a live search
Commercial (“best X for Y”, “who should I hire”)86.5%
Informational (“what is a negative keyword”)0.9%
All prompts mixed31% to 34.5%

That gap is enormous and it changes your strategy completely. Commercial questions are answered from the live web. That is the part you can influence this month. Informational questions are answered mostly from training data, which is frozen until the next model release and which nobody outside the labs controls.

If you sell something, your buyers are typing commercial prompts. That is good news.

Step 2: Your one question becomes many questions

Google calls this query fan-out and describes it in its own words: “AI Mode uses our query fan-out technique, breaking down your question into subtopics and issuing a multitude of queries simultaneously on your behalf.”

So “best PPC agency for Shopify” quietly becomes a dozen searches: pricing, Performance Max expertise, feed management, reviews, alternatives, case studies, contract terms, and so on.

The measurable result is dramatic. Ahrefs tracked 863,000 SERPs and 4 million AI Overview URLs:

Where AI Overview citations came fromJuly 2025January 2026
Top 10 for the original query~76%38%
Positions 11 to 10031.2%
Beyond position 10031.0%

What this means for you: you can be pulled into an answer by ranking for a sub-question you never targeted. You can also rank number one for the main query and be left out entirely. Ranking is an input, not a ticket.

Step 3: Passages get selected, not pages

This is the technical heart of it. Retrieval works on chunks, not documents. Perplexity says so plainly: “sub-document units are individually surfaced and scored against the original query parameters.”

Anthropic published the clearest explanation of why chunking breaks content. Their example of a bad chunk:

“The company’s revenue grew by 3% over the previous quarter.”

Which company? Which quarter? Standing alone, that sentence is useless, so it never gets retrieved. Anthropic’s fix, adding context back into each chunk, cut retrieval failures by 35 percent. Adding keyword matching on top took it to 49 percent. Adding a reranking step took it to 67 percent.

Two lessons fall out of that, and the second one is unpopular:

  1. Every section of your page must stand on its own. If a paragraph only makes sense because of the heading three screens above it, it dies in chunking. Name the service, the entity, the location and the timeframe inside the paragraph itself.
  2. Exact match keywords still matter. The 14 point jump came from adding BM25, which is old fashioned keyword matching. Anyone telling you keywords are dead has not looked at how retrieval is built.

This is the part that will change how you think about citations.

A teardown of ChatGPT’s retrieval stack published by Search Engine Land found that of 61,332 retrieved URLs, only 5,032 became lead citations and only 759 pages were fully opened. And:

An opened page is cited 74 percent of the time. A page that was retrieved but never opened is cited 7 percent of the time.

A ten times difference. Being retrieved is not the goal. Being opened is the goal.

It gets stranger. An arXiv study of 14 models found deep research agents cite sources that do not actually support the claim between 23 and 75 percent of the time, while link validity stays above 94 percent and topical relevance above 80 percent. The links work. The links are on-topic. The links often do not back up the sentence they are attached to.

Citations are partly a presentation layer. Treat them as a signal, not as proof of influence.

Step 5: Each engine reads a different internet

ProductWhere its index comes fromThe thing you need to know
ChatGPTIts own index, plus live scraped resultsOnly 1.5 percent of its URLs appear in Bing’s top 20 for the same queries. “ChatGPT is just Bing” is out of date advice. See our strategies to rank in ChatGPT search
Google AI Overviews and AI ModeThe Google Search index plus a custom Gemini modelYou cannot opt out of AI Overviews without opting out of Google Search
PerplexityIts own index, hundreds of billions of pages, updated tens of thousands of times per secondLeans on Reddit, LinkedIn and YouTube
ClaudeBrave Search (observed, never officially confirmed)Claude never cited Reddit once across 13,184 citations. It leans on review platforms like Clutch and G2
GeminiGoogle Search groundingNames brands far more often than it links
CopilotBing indexBing indexing is fast and underused

This table is why a single “AI content strategy” fails. A Reddit heavy programme produces zero Claude visibility. A review-site programme produces little Perplexity movement. You need to know which engine your buyers use before you pick a lever.

And here is the number that should end the “one AI score” reporting habit: across four models asked the same question, all four cited the same domain only 1.7 percent of the time.

Step 6: The crawlers, and why so many sites block themselves

There is not one AI bot. There are eleven that matter, and they do different jobs.

BotOwnerWhat it does
GPTBotOpenAITraining
OAI-SearchBotOpenAISearch index inclusion, not training
ChatGPT-UserOpenAILive fetch when a user asks
OAI-AdsBotOpenAIChatGPT ads
ClaudeBotAnthropicTraining
Claude-UserAnthropicLive user fetch
Claude-SearchBotAnthropicSearch quality
PerplexityBotPerplexityIndexing for citation. Not used for training
Perplexity-UserPerplexityLive user fetch
GooglebotGoogleThe Search index, which is what feeds AI Overviews
Google-ExtendedGoogleGemini training and grounding only. Does not affect Google Search or AI Overviews

Most robots.txt files know about one of these. A survey of 72 companies found GPTBot named by 18, ChatGPT-User by 15, OAI-SearchBot by 11 and OAI-AdsBot by none. Eight sites blocked GPTBot and four of those never mentioned OAI-SearchBot, which means they blocked training and left citation wide open, almost certainly by accident.

Here is a starting robots.txt you can copy. It allows retrieval and citation everywhere and lets you decide separately about training:

# Search and citation bots - allow these if you want to be cited
User-agent: OAI-SearchBot
Allow: /

User-agent: ChatGPT-User
Allow: /

User-agent: Claude-User
Allow: /

User-agent: Claude-SearchBot
Allow: /

User-agent: PerplexityBot
Allow: /

User-agent: Perplexity-User
Allow: /

User-agent: Googlebot
Allow: /

# Training crawlers - your call. Blocking these does not stop citation
User-agent: GPTBot
Allow: /

User-agent: ClaudeBot
Allow: /

User-agent: Google-Extended
Allow: /

User-agent: *
Disallow: /wp-admin/
Disallow: /wp-login.php
Disallow: /*?s=
Disallow: /*replytocom=

Sitemap: https://example.com/sitemap.xml

The trap nobody warns you about: your CDN may be blocking AI for you

This is the single most common silent killer of AI visibility right now, and it is not in your robots.txt.

Cloudflare has an “AI Scrapers and Crawlers” toggle that prepends its own rules and overrides whatever your site says. One practitioner auditing sites in 2026 reported that roughly half of the Cloudflare sites they audited had it switched on without the owner knowing. From 15 September 2026 Cloudflare begins blocking agent traffic by default on ad-carrying pages for new sites and free tier accounts.

You have to allow these at CDN level. Your robots.txt will look perfect while nothing gets through.

Go and check this now, before you read the rest of this article. It takes two minutes and it is worth more than every other tip here combined.

Three more technical failures worth checking while you are in there, and all three are in our technical SEO checklist of 105 checks built for AI search:

  • JavaScript. GPTBot, ClaudeBot and PerplexityBot do not execute JavaScript. They pull the HTML document and leave. If your key content renders client side, those crawlers see an empty page.
  • Page weight. ChatGPT rejects pages over 4MB with an HTTP 400 and reads nothing at all. No error reaches you.
  • JSON-LD. In ChatGPT’s pipeline, HTML is converted to Markdown and that process strips scripts, iframes and JSON-LD. Your schema may never reach the model at all through that route. It still does useful work for Google, which is why we publish copy-paste markup in guides like restaurant schema markup and Magento schema markup.

Step 7: It all changes again next month

Profound tracked around 80,000 prompts per platform and measured how many cited domains were completely different one month later:

PlatformMonthly citation drift
Google AI Overviews59.3%
ChatGPT54.1%
Copilot53.4%
Perplexity40.5%

Over six months the drift reaches 70 to 90 percent.

This has a brutal implication for reporting. If between 40 and 60 percent of cited domains change on their own every month, then any month on month change smaller than about 50 percent is indistinguishable from noise. Almost every “AI visibility report” being sold right now fails that test. If your agency shows you a tidy 12 percent improvement, ask them what their margin of error is.

What actually works? The evidence ranked

Only one properly controlled academic study exists on this, from Princeton, Georgia Tech and Allen AI, run across 10,000 queries. Here is what they found when they changed content and measured the visibility difference:

Change made to the contentVisibility lift
Adding quotations from credible people+41%
Adding specific statistics+30%
Improving fluency and readability+28%
Citing credible sources+27%
Adding technical terms+18%
Writing more authoritatively+10%
Simplifying language+13%
Adding unique words+6%
Keyword stuffingminus 8%

Keyword stuffing was the only tactic in the entire study with a negative result. The oldest trick in SEO is the single worst thing you can do here.

One honest caveat, because nobody else gives it: this study was run on GPT-3.5 in 2023. The authors themselves warned the methods “may require adaptation as generative engines evolve”. The +41 percent figure gets quoted with far more confidence than the paper supports. Use it as direction, not as a guarantee.

Separately, Semrush tested adding a structured summary, a TL;DR or a comparison table at the top of long content and measured a 16.1 percent citation increase across models.

The practical version

Do thisWhy
Put a direct quote from a named person with a real job title in every important pageHighest scoring change in the only controlled study
Replace adjectives with numbers. “We improved ROAS” becomes “we took ROAS from 2.1 to 4.8 over five months on a $40k monthly budget”Statistics scored second
Open every page with a two or three sentence answer, then a comparison tableStructured summaries lifted citations 16.1%
Make every section self-contained. Repeat the entity name, service and location inside the paragraphChunking destroys context-dependent paragraphs
Use question-shaped headings that match how people actually askQuery fan-out matches sub-questions to passages
Cite your sources with real linksScored +27%, and it is the right thing to do anyway
Publish original data nobody else hasIt cannot be copied, and it is the most quotable asset type

We go deeper on the on-page side of this in 10 advanced SEO techniques that win rankings and AI citations, and on the ecommerce side in our ecommerce SEO checklist for product pages and AI search.

Where the industry is getting AI visibility wrong

This is the section other guides will not write, because most of them are selling one of these things.

Myth 1: “Add an llms.txt file”

This is the most confidently repeated bad advice in the category. The evidence against it is overwhelming.

EvidenceFinding
Ahrefs, 137,000 domains28% publish an llms.txt. 97% received zero requests in May 2026. Only 1.1% of the requests that did arrive came from AI retrieval bots. Slackbot fetched llms.txt more often than PerplexityBot
John Mueller, Google, March 2026“AI companies have had a long time to do that, and nothing has happened regarding LLMS.txt support. I suspect the main users are SEO tools & companies curious to see what their competitors claim to be doing.”
John Mueller, earlierCompared it to the keywords meta tag: “this is what a site-owner claims their site is about… why not just check the site directly?”
Gary Illyes, Google, July 2025Google does not support llms.txt and is not planning to
Two independent agency server log samples“Less than 0.1 percent” of crawler hits, and “maybe 2 out of 4k”
Cyrus Shepard’s synthesis of 54 experiments and patentsScores llms.txt 2 out of 10, “no credible evidence”, against URL accessibility at 9.5

Our position: keep it if you already have one, it costs nothing and does no harm. Do not pay anyone to build you one. Do not let it near your priority list.

Myth 2: “Add FAQ schema and you will get cited”

Schema is useful. It is not a citation trigger, and the causal claim has never been demonstrated.

  • Google’s own AI optimization guide says directly: “Structured data isn’t required for generative AI search, and there’s no special schema.org markup you need to add.”
  • John Mueller: “Your ‘best geo insurance comparison site’ isn’t going to rank better by adding insurance markup.”
  • In ChatGPT’s pipeline, JSON-LD is stripped during the HTML to Markdown conversion, so it may never reach the model at all.
  • Google deprecated FAQ rich results in May 2026, removing the SERP-side reason to use FAQPage markup.
  • Ahrefs tracked 1,885 pages adding schema and AI citations barely moved.

The nuance worth keeping, and Gianluca Fiorelli made this point well: entity schema and content schema are different things. Organization and Person markup with a complete sameAs array is identity infrastructure. It helps machines know which “Acme” you are. FAQPage markup on a blog post is a citation lever that has never been shown to work. Do the first, stop obsessing over the second. Our schema guides show the difference in practice.

Myth 3: “GEO is a brand new discipline”

Every platform that has commented on this says otherwise:

WhoWhat they said
Gary Illyes, Google“To rank in AI Overviews, use normal SEO practices. You don’t need GEO, LLMO, or anything else”
Danny Sullivan, Google“Good SEO is good GEO, or AEO, AI SEO, LLM SEO, or LMNOPEO”
Nick Fox, SVP, Google“The way to optimize for AI Search is very similar, the same as how to perform well in traditional search”
Krishna Madhavan, Bing“Be skeptical of shortcuts. Fundamentals of SEO are still critical”
Google Search Central, May 2026“From Google Search’s perspective, optimizing for generative AI search is optimizing for the search experience, and thus still SEO”
Aleyda Solis“I don’t really care what optimizing for AI Search is called”

The SEO community has been blunter. A thread titled “GEO/AIO is essentially just a scam” drew 85 upvotes in r/bigseo. One commenter called it “just the next iteration of voice seo, aeo, geo, schema seo… fear based overtly technical seo approaches that cost businesses money but get little in the way of results.” Another, in a thread about what to deliver when a client asks for GEO: “That’s the whole point of ‘GEO’! You can do nothing besides what you do for SEO and charge twice as much.”

But here is the nuance that neither camp will give you. The “GEO is just SEO” quotes come almost entirely from Google and Microsoft, whose AI products run on their own search indexes. For them the statement is architecturally true. It is much less true for ChatGPT, which runs its own index with 1.5 percent overlap with Bing’s top 20, and Perplexity, which runs its own. Remember that only about 12 percent of AI cited URLs rank in Google’s top 10.

The defensible position: the inputs are the same, the retrieval systems and the measurement are genuinely different. Both extremes are overclaiming.

Myth 4: “Publish a listicle ranking yourself number one”

This one is actively backfiring, and the data is specific.

Lily Ray tested 100 B2B “best [category] software” queries. Self-promotional listicles were cited 323 times. In 224 of those cases, 69 percent, the brand’s own page was cited and the AI recommended a competitor listed on that same page. You wrote the page, you paid for the page, and the page sold your competitor.

She also found Google appears to have started demoting these from mid-January 2026: an $8bn B2B brand with 191 self-promo listicles lost 49 percent visibility, a SaaS with 228 lost 43 percent, another with 267 lost 38 percent.

The fix is not to stop making comparison pages. It is to make honest ones. Publish your scoring criteria. Include competitors who genuinely beat you at specific things and say so. Add a “who should not hire us” section. Date it. An honest list with your name at number four is a better citation object than a dishonest one with your name at number one, because the model reads the whole page and recommends whoever the page makes look correct. That is the standard we hold ourselves to in how to choose a digital marketing agency and our top Google Ads agencies list.

Myth 5: “Inject text aimed at the AI”

It works in laboratory conditions. Researchers at ETH Zurich made a manipulated camera 2.5 times more likely to be recommended by Bing Copilot, and a completely fictional camera hit a 59.4 percent recommendation rate, beating real Nikon and Fujifilm listings at 57.9 percent.

Three reasons not to do it anyway:

  1. Google’s spam policy now opens by defining spam as including “attempting to manipulate generative AI responses in Google Search.” Hidden text and cloaking are already covered.
  2. It is fragile. An RL-trained snippet rewriter worked against several models and failed completely against GPT-5-mini.
  3. It destroys itself at scale. In the same ETH study, when the number of attacking pages went from one to four, recommendation rates dropped for everybody including the attackers. It is a prisoner’s dilemma with no winner.

Myth 6: “Our tool gives you an accurate AI visibility score”

Practitioners who have actually bought these tools are not impressed.

WhoWhat they said
Paul Dyer, CEO of /prompt“If you use three different tools and give them the same prompts, you get three different answers”
Ryan Mason, President of Markacy“There’s really not much an AI tool can do or tell you to do. It’s just a benchmarker in my mind”
Heather Physioc, VMLHer agency tested around 17 different tools; results are “point in time”, not trended

Digiday reported $227 million of venture capital flowing into AI visibility tracking, with one vendor selling “30+ articles on autopilot” for $299 a month, and experts calling much of the category snake oil.

The strongest argument against these tools is not commercial, it is mathematical: at 40 to 60 percent monthly drift and 1.7 percent agreement between models, a precise-looking score describes a distribution that has already reshuffled by the time the report lands.

Myth 7: “AI search volume” data

Prompt volume numbers are built from paid panels and browser extension data representing far less than 1 percent of an estimated 2.5 billion daily inputs. AI does not return a predictable list, it returns a paragraph, and the same prompt asked three times can give three different answers.

Two more numbers that get repeated everywhere and should not be:

  • “70 percent of AI traffic hides in Direct.” No primary source exists for this. The mechanism is real, because mobile AI apps strip the referrer, but the 70 percent figure is invented.
  • “82 to 95 percent of AI citations come from earned media.” Traced to vendor outlets with undisclosed methodology. The Ahrefs and Orbit Media figures say something similar with disclosed samples. Use those.

Myth 8: “We can get you into the training data”

Training weights change when a model is retrained or released, on a schedule nobody outside the labs controls or can observe. Retrieval caches refresh in minutes, and ChatGPT’s cache runs on roughly a 30 minute freshness window.

So publishing today can change a grounded answer within hours. It cannot change what the model knows from memory until the next training run. Anyone selling the second thing is selling something unobservable.

Myth 9: “AI loves fresh content”

Half true, and false for Google.

PlatformFreshness preference
ChatGPTCites content 458 days newer than organic. Strongest freshness pull
AI assistants overall25.7% fresher than organic on average
Google AI OverviewsCites content 16 days older than organic. No preference at all
The average cited page across platformsAbout 500 days old

Changing publish dates without changing content carries Google penalty risk and buys you nothing. Ahrefs said so directly.

Myth 10: “AI traffic converts better” (or worse)

This is the biggest genuinely unresolved argument in the field, and anyone stating it confidently is guessing.

PositionEvidence
Converts worse973 ecommerce sites, $20bn combined revenue, 12 months: ChatGPT was about 0.2 percent of sessions and only paid social converted worse
Converts betterPractitioners with law firm, SaaS and B2B industrial clients consistently report the opposite, at 1 to 3 percent of traffic
Both are trueThe vertical is the confounder. High consideration B2B services behave differently from impulse ecommerce

There is also a mechanism that explains why AI may be getting under-credited: people use AI at the top of the funnel, then go back to Google in a new tab to buy. The AI did the work. Google got the attribution.

How do you actually measure AI visibility?

Track four things, separately

MetricWhat it is
Mention ratePercentage of your prompt set where your brand is named in the answer
Citation ratePercentage where your URL appears as a source
Position and framingWhere you appear in the list, and how you are described
Factual accuracyDoes it get your pricing, location, team size and services right

That last one is a real and separate failure mode. A model confidently quoting your 2023 pricing is a different problem from not being mentioned, and it is fixable at the source.

Build a fixed prompt set and keep it fixed

Pick 40 to 60 prompts across: discovery (“best X for Y”), comparison (“X vs Y”), problem-led (“my ROAS dropped, who can help”), qualification (“what should I ask before hiring”), local, and vertical-specific.

Keep them short. This matters more than people realise. Short conversational queries produce 30 to 50 times more brand mentions than long structured prompts. And short is what real people type: two surveys found a median prompt of eight words, and one practitioner reported that around 60 percent of real prompts were just search terms typed into a chatbot.

Run the set monthly. Fresh sessions, no logged-in history, consistent country setting. Log mention and citation separately, per engine. Never average ChatGPT and Gemini together, because their rates are inverted.

Set up GA4 properly

Create a custom channel group matching this regex:

chatgpt\.com|openai\.com|perplexity\.ai|gemini\.google\.com|bard\.google\.com|copilot\.microsoft\.com|claude\.ai|anthropic\.com|you\.com|phind\.com

Then build a second segment for direct traffic landing on pages three or more folders deep. Nobody arrives at /services/ppc/ecommerce/shopify/ by typing it. That is almost always AI or a chat app that stripped the referrer.

Add Bing Webmaster Tools and IndexNow while you are there. Bing indexes far faster than Google and it feeds Copilot.

If your GA4 setup is not clean to start with, none of this will read correctly. Start with our complete guide to using GA4 for marketing, then fix the data layer with Shopify server side tracking, server side GTM hosting and Google Enhanced Conversions and Meta CAPI setup.

Be honest about the ceiling. ChatGPT appends utm_source=chatgpt.com inconsistently, mobile apps strip referrers entirely, and OpenAI only shares citation data with licensed publishers. Nobody can measure this cleanly. Saying so builds more trust than inventing a number.

What should you actually do first? A priority order

PriorityActionTime
1Check your CDN and firewall are not blocking AI crawlers. Cloudflare’s toggle overrides your robots.txt10 minutes
2Check your key pages render in server side HTML with JavaScript off, and are under 4MB1 hour
3Build a fixed prompt set and take a baseline across ChatGPT, Gemini, Perplexity, Claude and AI Mode1 day
4Claim and complete every relevant directory and review profile, then push for review volume1 to 2 weeks
5Add complete Organization and Person schema with a sameAs array linking every profile you own1 day
6Rewrite your top pages: answer first, comparison table near the top, real numbers, named quotes, self-contained sections2 to 4 weeks
7Publish real pricing. “How much does X cost” is one of the most common qualification prompts and pages that answer it get quoted1 day
8Publish original data nobody else hasongoing
9Build presence where AI actually reads: YouTube, LinkedIn, Reddit, review platforms6 to 12 months
10Fix wrong facts about you at the source, quarterlyongoing

Notice what is not on that list: llms.txt, FAQ schema sprinkling, AI-targeted hidden text, and buying a visibility score.

Where AI actually reads, so you know where to spend

DomainShare of AI Overview brand mentions
YouTube22.9%
Reddit18.5%
Facebook10.1%
Google properties8.8%
Instagram5.6%
Quora4.7%
Wikipedia4.0%
Forbes0.8%
New York Times0.5%

User generated content is over 65 percent of the top 50 mention share. Forbes, the outlet everyone pays to get into, is at 0.8 percent.

If you sell products, there is a second surface most brands miss entirely. AI assistants pull a large share of product recommendations straight from Google Shopping data, which we cover in how ChatGPT, Claude and AI platforms pick products from Google Shopping and AI feed optimization. Your product feed is an AI visibility asset, not just an ads asset.

One important correction to the Reddit advice though: Claude never cited Reddit once across 13,184 citations. Reddit is a Perplexity, Gemini and AI Overviews lever. Review platforms like Clutch and G2 are the Claude lever. Pick based on where your buyers are.

And a warning on Reddit specifically: every major buyer subreddit bans vendor self-promotion outright, most require account age and karma thresholds, and AutoMod on r/marketing removes posts that merely name an agency. There is no shortcut. The only compliant route is a real personal account answering real questions for months before any mention is credible.

Frequently asked questions

What is the difference between AEO, GEO and AI visibility? AEO is answer engine optimization, GEO is generative engine optimization, and AI visibility is the outcome those two are trying to produce. In practice the acronyms describe the same work. Google’s own documentation calls all of it “still SEO”. Do not pay a premium for the label.

How long does it take to show up in AI answers? Directory and review work can appear within weeks because those pages are already crawled frequently. Your own content can be picked up within days on Bing and Copilot if you use IndexNow, and within a few weeks elsewhere. Community presence on Reddit, YouTube and LinkedIn is a 6 to 12 month project. Training data is not a timeline you control at all.

Why is my brand not showing up in ChatGPT? Work through it in this order. Are you blocked at CDN level, which is the most common cause. Does your content render without JavaScript. Are your pages under 4MB. Are you named on third party pages that ChatGPT reads. Is your brand name distinctive enough to be an entity, or is it colliding with something better known. Most “invisible in AI” cases are one of the first three.

Do I need an llms.txt file? No. 97 percent of published llms.txt files received zero requests, Google has said it does not support the format and is not planning to, and only 1.1 percent of the requests that do arrive come from AI retrieval bots. Keep one if you already have it. Do not build a project around it.

Does schema markup help with AI citations? Entity schema does a different job from content schema. Organization and Person markup with a complete sameAs array helps machines identify who you are, which is genuinely useful. FAQPage markup as a citation trigger has never been demonstrated, Google says no special markup is required, and ChatGPT’s pipeline strips JSON-LD during processing.

Should I block AI crawlers? Separate the decision by bot. Blocking training crawlers like GPTBot and ClaudeBot does not stop you being cited. Blocking retrieval crawlers like OAI-SearchBot and PerplexityBot does stop you being cited. In one study, the median domain that blocked GPTBot had 0.003 citations per Google ranking versus 0.417 for non-blocking domains, though the authors were clear this is correlation, not causation.

How do I know if an agency actually knows how to do this? Ask them to show you a mention rate on a fixed prompt set, per engine, over at least three runs, with mentions and citations reported separately. Ask what their margin of error is. Most cannot answer either question. Our guide on hiring a Google Ads agency covers the wider vetting process, and what a PPC agency should cost covers the pricing side.

Is AI traffic worth anything? It depends entirely on your vertical, and the honest answer is that the data disagrees with itself. Aggregate ecommerce data says AI traffic converts worse than Google organic. B2B, SaaS and professional services practitioners consistently report the opposite. Measure it for your own business before believing either side.

Is AI search replacing Google? Not so far. Similarweb found 95 percent of ChatGPT’s users also use Google in the same period, search reaches about 3.3 billion monthly uniques against 655 million for AI chatbots, and only 6.8 percent of ChatGPT answers included external links as of May 2026. AI search is layering on top of search, not replacing it. What has changed is that fewer searches produce a click: zero-click searches went from 60.45 percent in 2024 to 68.01 percent in 2026.

Can I pay to appear in AI answers? Not for organic mentions. ChatGPT has begun running ads, and OpenAI now operates a dedicated OAI-AdsBot crawler, so paid placement inside AI surfaces is arriving as its own channel. Anyone offering to pay their way into an organic recommendation is describing manipulation, which Google’s spam policy now explicitly covers.

The short version

AI visibility is not one thing and it is not one number. It is two separate outcomes with two separate levers.

Citations come from your website being technically reachable, structurally clean and genuinely useful at the passage level. That is an SEO job, and it is mostly the same SEO job you already know.

Mentions come from your brand name appearing next to your category across other people’s content. That is a PR, reviews, community and original-research job, and branded web mentions correlate with AI visibility three times more strongly than backlinks do.

Everything else being sold right now is either one of those two things with a new name on it, or it is llms.txt.

We do this work for clients as an AI SEO service, alongside SEO, PPC management, ecommerce PPC and white label PPC for other agencies. Recent results include AI citations across ChatGPT, Gemini and Google AI Overviews with 2.4x organic traffic, 117 percent organic click growth with AI visibility built from zero, and an SEO and AEO programme that scaled a pet supply store past AED 6.5m in revenue.

If you want to know where your brand currently stands across ChatGPT, Gemini, Perplexity, Claude and Google AI Mode, we run a fixed prompt set, log mentions and citations separately per engine, and tell you which of the two you actually have a problem with. Claim your free $500 audit and we will send you the baseline along with it.

Ishant

Ishant Sharma is the Founder and CEO of Hustle Marketers, a Google Partner digital marketing agency. With 12+ years of experience in Google Ads, Meta Ads, SEO, and e-commerce PPC, he has helped 2500+ brands generate $780M+ in trackable revenue. Upwork Top Rated Plus with 99% Job Success Score. Ishant Sharma is the digital marketing specialist, not the Indian cricketer of the same name.

I hope you enjoy reading this blog post. If you want my team to just do your marketing for you, click here.
Scroll to Top