Why AI Still Doesn’t Mention Your Business
You added the structured data. You wrote the FAQ page. You put an llms.txt file at the root of your site because someone said the AI crawlers read it. Months later you ask ChatGPT what the best options in your category are, and your name does not come up.
You are probably not doing it wrong. Most of the advice is aimed at the wrong problem.
Here is what we have measured, on our own site and from published research, written as plainly as we can manage.
Three things you were told to do that do very little
1. The llms.txt file
This is a file you put on your site to tell AI systems what your pages are about. Ahrefs looked at 137,000 sites that had one. In 97% of cases, nothing ever requested the file. SE Ranking looked at around 300,000 domains and found no link between having it and being mentioned in AI answers. Google has said it does not support it. No other major AI company has committed to reading it either.
It takes ten minutes to create, so keep yours if you like it. It is not a plan.
2. Schema markup — at least for the reason you were given
Schema is the hidden code that describes your page to machines. A controlled test across seven AI platforms, run between December 2025 and March 2026, found that only one of them — Google’s Gemini — could read it at all. When the testers put a fact ONLY in the schema and then asked each platform about that fact, not one of them could answer.
ChatGPT is the clearest case. When it converts your page into something the model can read, it strips the schema out first.
Keep your schema. It still earns rich results in Google, and it still helps machines tell your company apart from a similarly named one. It is simply not how you get quoted.
3. “Just publish more content”
Sometimes true. Often not, for a reason we will come back to at the end.
The rule underneath all three. If a fact matters, it has to appear in the visible words on your page. Your price. What you actually do. Who it is for. If those live only in your schema or only in your meta description, an AI cannot quote you on them.
We failed this ourselves. Our own homepage used the word “software” exactly once, and that once was inside the hidden code. In the visible text, we never said what kind of thing we were. We are in the business of catching that, and we did it to ourselves.
There is no single “AI,” and the ones people use do not read the same web
This is the part that explains most of the confusion, and almost nobody says it plainly.
There are a lot of assistants now, and more arriving. But there are fewer indexes behind them than there are names on the front, and that is the thing worth understanding.
These are the ones people go to with a question and get an answer with sources attached. Where each looks, as far as anyone has published:
- ChatGPT built its own index of the web, alongside other sources. Researchers found only about 1.5% of the pages in that index appear in Bing’s top results for the same question.
- Claude looks things up through Brave Search. Anthropic documents Brave as the search service its web search calls.
- Gemini, and Google’s AI Overviews, use Google’s index.
- Perplexity runs its own crawler and its own index, along with some third-party sources. The mix has changed over the product’s life.
- Copilot uses Bing.
- Grok has its own web search, plus search across X, which is why people reach for it on anything current. Who supplies the underlying web results is not published.
- Meta AI has drawn on Google and Bing, and has been reported to be building its own. It is used inside Meta’s apps more than as a place people go to look something up.
One kind of assistant is deliberately missing from that list, and the distinction matters because it changes what you can do about it. A different sort of assistant lives inside a single company’s product — the one on a large retailer’s site, the one inside a messaging app, the support bot on a software vendor’s dashboard. Those mostly answer from their host’s own catalogue or their own documentation, not from the open web. If your customers meet you through one of those, none of this applies and a different discipline does.
Two things follow from the list above.
The first is that one index can cost you several assistants. Brave is not only behind Claude. Mistral’s Le Chat, you.com and Kagi build on the same search service. Being absent there is not a one-assistant problem.
The second is scope, and we would rather tell you than have you find out. We measure four: ChatGPT, Claude, Gemini and Perplexity. Copilot, Grok, Meta AI and others matter too, and we do not measure them. If a tool tells you it covers every AI, ask which ones, and ask where each one looks.
For the four we do measure: when researchers compared which sources they cite for the same questions, the overlap between any two of them ran from 16% to 59%. And only about 12% of the pages ChatGPT cites are in Google’s top ten for that query.
So being first on Google tells you very little about whether ChatGPT or Claude will mention you.
It also means you can be invisible to one assistant for a reason that has nothing to do with your website. We are the example. In early September, three of the four found us by name on ten attempts out of ten. Claude managed three. Same site, same day. Later we found out why: our site is not in Brave’s index, and Brave is where Claude looks. No amount of work on our pages would ever have changed that. We only found out because we checked the index directly instead of assuming.
Three gates, and they happen in order
Everything above fits into a simple sequence. An assistant has to get through three gates before it mentions you.
- Is your site in the index this particular assistant looks in? If not, nothing else matters.
- Does it pull your page in when somebody asks this question?
- Having read your page, does it name you and link to you?
Nearly all the advice you will read is about gate three. Write better answers. Add structure. Improve the page. That advice is fine. It is useless if you are failing gate one.
So the first useful question is not “how do I get cited.” It is “which gate am I failing.” You can find that out in about an hour, and we will show you how below.
What an AI actually reads of your page
Far less than you would think.
Research published in 2026 into how ChatGPT retrieves pages found that in its fast mode, the only body text the model sees is roughly the first 200 characters after your main heading. It is the same 200 characters no matter what was asked. Your meta description is ignored completely.
On most sites, those 200 characters are spent on a breadcrumb trail, a date, a byline and a cookie notice.
Two more findings from the same work. The fetcher does not run JavaScript, so anything your site draws after the page loads is invisible to it. And pages larger than 4MB are rejected outright rather than trimmed.
One number worth carrying away: when a page is merely retrieved, it gets cited about 7% of the time. When it is actually opened and read, that rises to about 74%.
The uncomfortable part: the answer is mostly built from other people’s pages
When somebody asks “what are the best tools for X,” the assistant is mostly not reading vendor websites. It is reading the pages that write about vendors. Roundups. Reddit threads. Review sites. Industry blogs.
Ahrefs studied 75,000 brands. The strongest thing associated with being mentioned by AI was how often the brand was mentioned on other people’s web pages. Backlinks mattered far less. Brands in the top quarter for mentions averaged 169 AI mentions. The next quarter down averaged 14.
A separate analysis of 30 million cited sources found the most-cited sites were Reddit, YouTube and LinkedIn, followed by Wikipedia and Forbes.
Here is what that means in practice, and it is the thing we resisted longest. There is no setting you can change on your own website that puts you inside somebody else’s article. Someone else has to write about you. Almost every task on the usual checklist is something you do to your own property, which is exactly why the checklist can be completed perfectly and change nothing.
Our own number for this is still zero.
Across 200 answers to category questions, measured five times each across four assistants, our domain appears in none of the cited sources. Not ranked last. Absent from the material the answers are built from.
How to check all of this yourself, free, without any tool
You do not need us for this part, and we would rather you did it than took our word for anything.
- Search
site:yourdomain.comon Google, on Bing, and on Brave. Brave matters because it is what Claude uses. If one of them returns nothing, that is gate one, on that assistant. - Look at what your page sends before any code runs on it. Instructions are in the box below. If a sentence you care about is not in there, the fetcher cannot see it.
- Read the first sentence after your main heading. Does it say what you are, in the words a buyer would use? If it is a slogan, an AI has nothing to quote.
- Look for your price and your main claim in the visible words on the page, not only in the code.
- Open four or five of your own URLs and compare them to your homepage. If several return the same document, you have one page, not five.
How to see what the fetcher sees.
The simplest way takes one keystroke and changes no settings.
Open the page you want to check, then press Ctrl+U on Windows, or Cmd+Option+U on a Mac. A new tab opens showing the raw page exactly as your server sent it, before any code ran. That is close to what an AI fetcher receives.
Now press Ctrl+F and search for a distinctive sentence from your page. Your price. Your main claim. The line under your heading. If you cannot find it in there, it is not in what the fetcher sees, no matter how clearly it appears in a normal browser.
The thorough version, if you want it: in Chrome, Edge or Brave, press F12 to open developer tools, then Ctrl+Shift+P (Cmd+Shift+P on a Mac), type “javascript”, and choose Disable JavaScript. Reload the page. What remains is roughly what the fetcher gets. Close developer tools when you are done and everything goes back to normal.
That last one caught us. Four of our own public pages, including our pricing page, were serving the homepage to everybody who was not logged in. The page returned a normal “OK” response the entire time, so nothing looked broken. We build a check for exactly this failure, and we still missed it on our own site for eleven days. We found it by running our own tool on ourselves, which is the only reason we can tell you about it.
Can you do all this yourself?
Yes. And for the five checks above, you should. They cost nothing, they take about an hour, and you will learn more from doing them than from reading about them.
The full measurement is a different kind of job, and it is worth being precise about why, because it is not expertise. It is arithmetic.
A proper measurement is twelve questions your buyers actually ask, put to four assistants, five times each. That is 240 separate questions, asked one at a time, with every answer read and every cited source written down. Then the same 240 again next month, worded identically, or the comparison means nothing.
Five times each is not padding. One answer from one assistant is a coin flip. Ask the same question twice and you will often get two different answers, and neither one is the truth on its own. The pattern is the measurement.
That is where doing it by hand tends to come apart. Not at the start. At question ninety, when you ask once instead of five times. When you quietly drop the run that errored instead of excluding it properly. When you reword a question slightly next month because the old wording felt clumsy, and the comparison you were building stops working without telling you.
Our site analysis takes about ninety seconds. The four-assistant sweep takes a few minutes and costs us about $2.70 in engine fees — measured, from a run of ours on 18 September that came to $0.54 for 48 answers. It asks the same questions the same way every month, keeps every transcript, and prints how many runs each number came from.
A few of the checks are simply too dull for a person to do reliably. Comparing five of your own URLs by their exact byte length. Extracting the precise first two hundred characters after your heading. Looking up whether your domain sits in each of four separate search indexes. None of that is difficult. All of it is the kind of thing a person does correctly the first time and approximately by the fourth.
We have been building this since March 2026, and most of what is in this article we found by getting it wrong on our own site first.
So if you have the patience, do it yourself. We would rather you measured badly than not at all. What we sell is the repetition and the discipline, not a secret.
What we do about it
Briefly, because the sections above matter more than this one.
We check which index each assistant’s search actually finds you in, and we report it one engine at a time rather than as a single blended score, because one number hides the one assistant you are invisible to.
We check whether the facts you care about appear in the visible words on your page, not only in the code.
We ask the questions your buyers ask, across four assistants, several times each. We keep the transcripts. We print how many runs each number came from. We keep answers that came from a live search separate from answers written from the model’s memory, because those are different claims about the world. If a run errors out, we exclude it rather than counting it as a “no”.
The part we care about most
Here is a problem you would never think to worry about, and it is the one that worries us.
A tool can be completely broken and still look completely fine. When a check quietly stops working, it does not show you an error. It shows you a reassuring result. Everything comes back green, and you have no way of knowing that the green meant nothing.
So we test every check twice. Once against something we know is broken, to be sure it says so. Once against something we know is fine, to be sure it does not raise a false alarm. If a check cannot correctly spot a problem we planted on purpose, it does not ship.
That is how we caught one of our own. We had built a way to look up whether a website was in a particular search index, and it appeared to work. Then we ran it against a web address we made up, a domain that does not exist and could not possibly be in any index. It reported that one as present too. It had never been checking anything at all. We deleted it and started again.
Had it shipped, it would have told every customer their site was fine. They would have believed it, because there would have been nothing to see.
This sounds like a technicality. It is the thing that decides whether any number anyone gives you means anything at all. Ask any tool you are considering whether they test their own checks this way, and what happened the last time one of them failed.
What nobody can do
We cannot make an assistant quote you. Nobody can, and anyone telling you otherwise is selling something.
What can be done is finding out which gate you are failing, fixing the part that is genuinely in your control, and measuring again to see whether it moved.
Our own category number is still zero. When it changes, we will publish that too.
Related reading
- I scored zero on my own tool — the dated narrative this primer grew out of.
- Reading isn’t citing — why crawler traffic is not the same as being mentioned.
- How it works — the method, step by step.
- What you actually get — the five layers of the action plan.
- AEO tools: an honest comparison — with dates on every claim.
AEO Analyzers (aeoanalyzers.com) was created and is solely maintained by Lindsay Hiebert, founder of PIGENAI LLC. It is unaffiliated with any similarly named browser extension, plugin, or tool.