Most AI visibility numbers come from one logged-in screenshot. Ours come from the questions your buyers actually ask, run in clean sessions across seven AI engines, repeated, dated and scored on seven metrics. Here is the full method, so you can check our work.
To measure AI visibility, run the questions your buyers ask through each AI engine in a clean, non-personalized session, repeat every prompt, and track how often you are named, where you rank, who wins instead and which sources the engine cites. Then tie it to Google Search Console so you can see AI impressions, not just answers.
AI engines personalize answers using your chat history, your location and the wording you use. Run a prompt about your own category from your own account and the engine has already learned you care about your brand. We tested it on ourselves.
Each metric answers a different question and points to a different fix. We report all seven per engine and per prompt, so you can see exactly where you are winning and losing.
The percentage of tracked AI answers in your category that name your brand. The closest thing to market share in AI search. In our sample client report, the client reached 31%, the highest of eight brands.
How often you are named across repeated runs of one buyer prompt. Tells you which questions you own and which you lose.
Named first and named fifth are different outcomes. Buyers read the top of the shortlist.
How often the engine quotes your own website as a source. High citation share means AI trusts your pages, not only what others say about you.
Who gets named when you do not. This is your real competitive set in AI search, and it is often not the list your sales team expects.
The third-party sites AI cites in your category, ranked by how often. Roundups, review sites, forums and media. This becomes the GEO placement plan.
How often you appear in Google AI Overviews and AI Mode, straight from your own Search Console. The one AI metric Google reports directly.
A blended score from 0 to 100 looks tidy on a dashboard and hides everything you need to act on. A brand can score well by being named often but always fifth, or by winning prompts no buyer asks. If a score cannot tell you what to fix next, it is a vanity metric. We show the parts.
The same process runs in every audit and every monthly report, so results compare cleanly over time.
We map how your buyers search: their job, their problem, the alternatives they consider. Then we agree on five money prompts, the questions asked right before a demo, plus a wider tracked set.
Every prompt is run on every engine before we change anything on your site. Without a real before, there is no honest after.
Logged out or temporary chat, location stated in the prompt, US Google settings, each prompt repeated. One answer is an anecdote. A rate across runs is a measurement.
Share of answers, mention rate, position, citation share, competitor wins and the source map, per engine and per prompt.
We read your Search Console, including the Generative AI report, and your analytics. We split branded from non-branded and filter out paid traffic, so AI wins show up as real demand.
A monthly report that leads with a one-line takeaway per section, lists what shipped, and sets next month's targets. Full re-run on every engine at day 90.
The prompt decides the result. We never test branded prompts like "Is Arobis AI good?". We use the category words your buyers use, across five intent types. Examples below use a project management tool.
Every result in our reports states these conditions. If a condition changes, we say so in the report.
| Condition | What we do | Why |
|---|---|---|
| Session | Logged out, or a temporary chat with memory off | Removes personalization from your own history |
| Location | Country named inside the prompt; Google set to US, English, non-personalized | Our Oct 1, 2026 test: changing Google's country setting did not change the AI Mode shortlist. Naming the country in the prompt did. |
| Repetition | Each money prompt run several times; results reported as rates | AI answers vary from run to run |
| Wording | Buyer category words, no brand names, no leading phrasing | Leading prompts produce the answer you wanted |
| Dating | Every capture saved with engine, date and session type | Engines update often; undated results cannot be checked |
| Traffic | Paid AI traffic, like your own ChatGPT ads, filtered out by UTM | In one client account, ads made up about 200 of 230 monthly "ChatGPT" visits |
Each engine exposes different data, so we measure each one the way that engine allows.
| Engine | How we measure | Cadence |
|---|---|---|
| ChatGPT | Daily tracking plus logged-out manual checks | Daily |
| Google AI Overviews | Daily tracking plus Search Console Generative AI report | Daily |
| Google AI Mode | Manual non-personalized checks plus Search Console impressions | Monthly |
| Gemini | Manual clean-session checks | Audit + day 90 |
| Perplexity | Manual checks; session type noted where sign-in is required | Audit + day 90 |
| Claude | Manual clean-session checks | Audit + day 90 |
| Microsoft Copilot | Manual checks plus AI citation data in Bing Webmaster Tools | Audit + day 90 |
Every engine is fully re-run at day 90. AI Search Growth clients get weekly re-checks.
Search Console now has a Generative AI report. It is the only first-party AI data Google gives you, and it is easy to misread.
Rules are only useful if they cost you something. These cost us prettier reports.
Founder and CEO of Arobis AI. Leads strategy and measurement on every client program. More about Ran
We run the questions your buyers ask through ChatGPT, Google AI Overviews, Google AI Mode, Gemini, Perplexity, Claude and Copilot in clean, non-personalized sessions, repeat each prompt, and score seven metrics: share of AI answers, prompt mention rate, average position, citation share, competitor win rate, the source map and Google AI impressions from Search Console.
AI share of voice, which we call share of AI answers, is the percentage of tracked AI answers in your category that name your brand. If your brand appears in 310 of 1,000 tracked answers, your share is 31%. It is the closest thing to market share for AI search.
A single score hides what you need to act on. A brand can be named often but always fifth, or rank first on prompts no buyer asks. We report the separate metrics, per prompt and per engine, so every number points to a specific fix.
Because personalization flatters your brand. In our own test on September 10, 2026, the same prompt named Arobis AI first in a logged-in ChatGPT account and did not name it at all in a temporary chat. Buyers do not share your history, so we only report clean sessions.
The AI Visibility Audit tests 40 buyer prompts. In a monthly engagement we agree on five money prompts with you, the questions asked right before a demo, plus a tracked set of about 25 prompts that runs daily.
The tracked prompt set runs daily on ChatGPT and Google AI Overviews. Money prompts are re-checked by hand every month and fully re-run on every engine at day 90. AI Search Growth clients also get weekly re-checks.
Yes. Every result in our reports lists the engine, the date, the session type and the prompt, so you can re-run it in a temporary chat. Answers vary from run to run, which is why we report rates across repeated runs rather than a single screenshot.
Google Search Console including its Generative AI report, your web analytics, Bing Webmaster Tools for Copilot citation data, Ahrefs for authority, ZeroRank for daily AI answer tracking, and manual clean-session checks on every major engine. Tool costs are never billed to clients.
Your free AI Visibility Audit preview uses this exact method: real buyer prompts, clean sessions, dated results.
AI Search Demand Generation for B2B SaaS. Get recommended by ChatGPT, Gemini, Claude, Perplexity, Copilot and Google AI.