Methodology
How citelity measures AI visibility
Every number in this product comes from somewhere, and some sources are better than others. This page says exactly where each one comes from, what the badge next to it means, and what we will not claim.
The short version: AI answers are generated per session. Ask the same question twice, from two cities, and you can get two different answers citing two different sets of pages. Any tool telling you it captures the complete picture is describing something that does not exist.
Where each number comes from
Every engine is included on every plan. There are no per-engine add-on fees, so nothing about the list below depends on what you pay.
ChatGPTlive sample
Real ChatGPT Search responses via a third-party provider
OpenAI publishes no citation API for consumer ChatGPT. Responses are genuine; the sample is collected rather than exhaustive.
Google AI Overviewssearch data
Google search results, including asynchronously generated AI Overviews
Google generates some AI Overviews after page load. We explicitly request those too — omitting them would understate how often you are exposed.
Geminiofficial API
Google's own Gemini API
First-party access. Grounding sources come back as structured data rather than being parsed out of a page.
Claudeofficial API
Anthropic's own API
First-party access, with web search enabled so the answer reflects live sources.
Perplexityofficial API
Perplexity's Sonar API
First-party access. Citations are returned as part of the response.
Microsoft Copilot—
No source available
No official or reliable citation data exists. We show a dash instead of an estimate.
What the badges mean
Every metric carries a badge describing how it was obtained. The badge is part of the number, not decoration around it.
official API
The engine's own API returned the answer and its sources. The most reliable class of data we have: nothing is inferred, nothing is parsed out of a rendered page.
live sample
Real responses from an engine that publishes no citation API, collected through a third-party provider. The answers are genuine. The sample is a set of collected responses, not a complete feed.
search data
Taken from live search results, including AI Overviews that Google generates on the fly. Reflects what a searcher in the queried location would have seen at that moment.
A dash means not measured. It never means zero. Where we have no trustworthy source, you get a dash and an explanation rather than an estimate dressed up as a measurement.
How a citation is counted
An engine typically looks at more pages than it ends up citing. We count only the sources the answer actually cites. Counting everything the engine merely consulted would inflate your citation rate and make the number useless for deciding what to fix next.
Citation position is the order your page appears among the cited sources for that answer. Where an engine attaches sources to individual passages, position follows the order of the answer itself, because the sources behind the opening lines are the ones a reader encounters first.
Mentioned without a link is tracked separately. Being named in an answer without a citation is a different situation from being cited, and it calls for a different fix, so the two are never merged into one figure.
When an engine returns no AI answer for a prompt, that is recorded as a result, not an error. Plenty of queries never trigger one.
How often we check
A full scan runs weekly across every tracked prompt and every engine. On Pro and Scale you can star individual prompts for a daily pulse on ChatGPT and Google AI Overviews.
Daily checks on everything would cost more and tell you less. LLM answers fluctuate from one day to the next without anything having changed on your site or theirs. Weekly depth plus a daily pulse on what matters gives you the trend without turning normal noise into a stream of alarms.
How we measure whether a fix worked
After a page is created or updated, we take a snapshot of its metrics at that moment and measure against it three times.
Day 14 — first checkpoint
Indexation and early signal. Never a verdict.
Day 30 — trend
A direction. Still not a verdict.
Day 60 — result
The only stage where we use the word result.
The staging is not caution for its own sake. Content changes usually begin moving positions at four to six weeks and settle between eight and twelve. Judged at two weeks, most legitimate work would look like failure — measured by a clock rather than by the work.
If day 60 shows no improvement, that is what your timeline will say. You would see the same thing in your own Search Console, and a tool that hides it is only a slower way of finding out.
What we don't claim
We do not publish an accuracy percentage, and we are wary of anyone in this category who does. Doing so honestly would require a ground truth to compare against — a complete record of every AI answer served to every user. No such record exists outside the engines themselves, and they do not publish it.
So we do not tell you how close our figures are to a total we cannot see. We tell you precisely how each figure was obtained and let you judge it. That is the honest ceiling in this category today, and claiming more would mean inventing the missing part.
We also make no promises about outcomes or timelines. What is under our control is the process: finding the prompts and pages where AI answers without you, writing the fix, and measuring what changed.
Questions
Why don't citelity's numbers match another AI visibility tool?
Because there is no single correct number to match. Every tool asks its own prompts, from its own location, at its own moment, through its own access method. AI answers are generated per session, so two tools querying the same prompt minutes apart can legitimately get different answers with different sources. A tool reporting a different figure is not necessarily wrong, and neither are we — the figures describe different samples.
Does citelity see every mention of my site in AI answers?
No, and no tool can. What we report is a floor, not a total: these prompts, on these engines, at these times, produced these citations. Real coverage is at least what we show and almost certainly more. We would rather report a verifiable floor than an unverifiable total.
What does the 'live sample' badge mean?
That the data comes from real ChatGPT Search responses collected through a third-party provider, because OpenAI publishes no citation API for consumer ChatGPT. The answers are genuine; the sample is a set of collected responses rather than an exhaustive feed.
Why is Microsoft Copilot shown as a dash?
Because no official or reliable source of Copilot citation data exists. Rather than estimate it or quietly leave it out, we show a dash. A dash means not measured. It never means zero.
How long before a fix shows up in the numbers?
Longer than most people expect. Content changes typically start moving positions at four to six weeks and stabilise between eight and twelve. That is why measurement runs in three stages — day 14, day 30, day 60 — and why nothing is called a result before day 60.
See what AI answers about your topics right now
Every source labelled, every gap named, and the fix written against the answer AI is giving today.
See plans →