Reference · Measuring AI visibility
AEO metrics and KPIs: definitions
AEO and GEO metrics and KPIs defined: mention rate, citation rate, share of voice, accuracy, AI impressions, and AI referrals, with formulas and limits.
By Paul Maxwell, founder of AEO HQ
Published · Updated
The main KPIs for AEO and GEO are the mention rate, which is the share of AI answers to buyer questions that name your company, the citation rate, share of voice, and the accuracy of what the answers say about you. Each is counted separately for each AI assistant, over many repeated runs. Around them sit the counts Google and Microsoft report to site owners, AI referral visits in analytics, and what buyers say when asked how they found you.
This page defines each metric, gives its formula and data source, and states what it cannot show. It is a companion to how to measure AI visibility, which covers how many prompts and runs to collect and how to calculate error ranges, so those steps are not repeated here. Where a platform or a vendor defines a metric differently from AEO HQ, the difference is noted. Sources were checked on September 27, 2026.
Scope and definitions
This page covers assistants that answer questions in sentences and sometimes cite web pages: ChatGPT, Google's AI Overviews and AI Mode, Gemini, Claude, Perplexity, and Microsoft Copilot. It does not cover ordinary search rankings, except where Google or Microsoft report AI features in their own tools.
- A metric is a defined count or rate. A KPI (key performance indicator) is a metric chosen to judge progress toward a stated goal. The same metric can be a KPI for one team and background for another.
- A prompt is a question typed into an assistant. A run is one answer to one prompt, collected once. A prompt panel is a fixed set of prompts run on a schedule.
- An unbranded prompt asks about a need or a category without naming a company. A branded prompt names the company, such as "What does [company] charge?"
- A mention is the company's name in an answer. A citation is a link to one of the company's pages shown with the answer.
- A window is the period a figure covers, such as four weeks.
The metrics at a glance
| Metric | What it counts | Formula | Data source |
|---|---|---|---|
| Mention rate | Answers that name the company | Runs that name the company ÷ all runs | Prompt panel |
| Citation rate | Answers that link to the company's pages | Runs that cite the company ÷ all runs | Prompt panel |
| Prompt coverage | Prompts that named the company at least once | Prompts with at least one mention ÷ prompts tracked | Prompt panel |
| Share of voice | The company's part of all mentions in a named competitor set | Company mentions ÷ mentions of every company in the set | Prompt panel |
| Position | Where the company appears in a list | Order of first mention, averaged over runs | Prompt panel |
| Sentiment | How favorably answers describe the company | Mentions graded positive, neutral, or negative | Prompt panel, graded |
| Accuracy | Branded answers that state the facts correctly | Answers graded correct ÷ answers graded | Prompt panel, graded by a person |
| Composite score | A vendor's combination of several measures | A weighted sum set by the vendor | Tracking tool |
| AI feature impressions | Times links to your site were shown in AI Overviews and AI Mode | Count | Google Search Console |
| Citations and cited pages | Your pages shown as sources in Microsoft's AI answers | Count; average unique pages cited per day | Bing Webmaster Tools |
| Citation Share | Your part of all citations for one grounding query | Your citations ÷ all citations for that query | Bing Webmaster Tools |
| AI referral sessions | Visits that start from a link in an assistant's answer | Count | Google Analytics 4, Microsoft Clarity |
| AI key events | Leads, sign-ups, or purchases in those visits | Count; sessions with a key event ÷ sessions | Google Analytics 4 |
| Self-reported AI source | Buyers who say an AI tool sent them | Answers naming an AI tool ÷ all answers | Form question, CRM |
| AI-influenced pipeline | Deals from buyers who came through AI | Sum of pipeline or revenue | CRM |
| Crawler and fetcher requests | Requests from AI companies' crawlers and fetchers | Count by agent and URL | Server logs |
The sections below take each metric in turn. Figures marked "AEO HQ calculation" are ours.
Metrics from a prompt panel
The first eight metrics come from a prompt panel: fixed prompts run on each assistant, with every answer saved. The measurement guide linked above explains how to build the panel and how many runs to collect, and AEO HQ's recommended measurement design gives default panel sizes.
A worked example
The table below is hypothetical data, made up to show the arithmetic. It describes no real company or result. It has three prompts, two runs of each, one tracked company, and two competitors, X and Y.
| Run | Prompt | Company named | Company's page cited | Answer cited any source | Competitors named |
|---|---|---|---|---|---|
| 1 | P1 | Yes | No | Yes | X |
| 2 | P1 | No | No | Yes | X, Y |
| 3 | P2 | Yes | Yes | Yes | Y |
| 4 | P2 | Yes | No | Yes | X |
| 5 | P3 | No | No | No | None |
| 6 | P3 | No | No | No | X |
From this log (AEO HQ calculation):
- Mention rate: 3 of 6 runs, 50.0%.
- Citation rate: 1 of 6 runs, 16.7%. Counted only among the four answers that cited any source, it would be 1 of 4, 25.0%.
- Prompt coverage: 2 of 3 prompts, 66.7%.
- Share of voice: the company was named 3 times and the competitors 6 times (X four times, Y twice), so 3 of 9 mentions, 33.3%.
Six runs are far too few to act on: the 95% Wilson interval, an error range suited to small samples, runs from 18.8% to 81.2% for 3 of 6 (AEO HQ calculation; the method is in AEO HQ's statistics section).
Mention rate
Definition. The share of runs whose answer names the company. AEO HQ counts a mention when the answer names the company as an option or describes it, using matching rules fixed before the results are read, such as the exact company name, its common variants, and its domain.
How to compute it. Mention rate = runs that name the company ÷ all runs, for one assistant and one window. Keep every run in the denominator, including answers where the assistant did not search.
Why it is the usual primary KPI. Whether a brand is named holds steadier than where it appears. In a study in which 600 volunteers ran 12 prompts a combined 2,961 times, it took about 1 in 1,000 runs to see two brand lists in the same order (opens in a new tab), and the author concluded that "visibility % across dozens to hundreds of prompts run multiple times is a reasonable metric" (opens in a new tab) (vendor study; a co-investigator works for a tracking vendor). In a panel of 102,025 answers, 77.5% of brand, prompt, and engine combinations were always or never mentioned (opens in a new tab) (preprint; the author co-founded, and holds equity in, the platform studied).
Limits.
- It describes the prompts, not all buyers. People word the same need very differently: 142 prompts written by people for one intent had a mean semantic similarity of 0.081 (opens in a new tab). A rate holds for the panel it was measured on.
- It needs an error range and a per-assistant breakdown. Report each assistant separately, with the number of runs and an interval.
- Vendors count mentions differently. HubSpot writes that "Mentions tell you how often your brand appears in AI-generated answers without a link" (opens in a new tab) and treats linked appearances as citations. AEO HQ counts a mention whether or not a link is shown. Ask how a tool counts before comparing its numbers with anyone else's.
Citation rate
Definition. The share of runs whose answer links to at least one of the company's pages.
How to compute it. Citation rate = runs that cite the company ÷ all runs. Group URLs by domain before counting, and decide in advance how redirects, parameters, and subdomains are handled (AEO HQ recommendation).
Variants. Other counts share the name. Some tools report the share of all citations that point to the company, which divides by citations rather than by runs. HubSpot's documentation records a citation when an answer includes a link to a webpage or "a domain name embedded in its answer" (opens in a new tab), so a domain written as text also counts. Bing's first-party counts are described below.
Limits.
- Many answers cite nothing. In one panel, ChatGPT activated web search only for specific queries, leaving 57.8% of its runs with zero citations (opens in a new tab) (preprint; German-language commercial prompts, January–March 2026; one author is affiliated with the company whose data export was used). A rate calculated only among answers that cite something leaves those runs out and overstates visibility. A 2026 review of 45 studies puts it plainly: "Outputs without search, without citations, or with errors are outcomes, not data to be discarded" (opens in a new tab) (preprint).
- Being cited is not being recommended. In a published test of 48 AI answers to questions about which agency to hire, the tester found that one agency, Searchbloom, was the source AI reached for most on the question in its wider tracking, yet was named as an agency once in 48 answers (opens in a new tab) (practitioner test by an agency). Report mention and citation rates side by side.
Prompt coverage
Definition. The share of tracked prompts for which the company was named at least once in a window. HubSpot uses it as one part of its brand visibility score: "Prompt coverage: the percentage of prompts where your brand appears at least once." (opens in a new tab)
How to compute it. Prompt coverage = prompts with at least one mention in the window ÷ prompts tracked.
Limits. Coverage rises with the number of runs even when nothing else changes. If a company has a fixed 10% chance of being named in each answer to a prompt, the chance it is named at least once is 10.0% after one run, 27.1% after three, 52.2% after seven, and 95.8% after 30 daily runs (AEO HQ calculation: 1 minus 0.9 raised to the number of runs, treating runs as independent). Real runs of one prompt tend to repeat each other, as the 77.5% figure above shows, so real coverage climbs more slowly. Compare coverage only between periods, tools, or competitors measured with the same number of runs.
Share of voice
Definition. Share of voice is the company's mentions as a share of all mentions of a named set of companies, the company included. HubSpot's documentation defines it as "the proportion of total brand mentions that belong to your brand across tracked prompts" (opens in a new tab).
How to compute it. Share of voice = company mentions ÷ mentions of every company in the set, counted over the same runs.
Limits. It changes whenever the competitor set changes, so name the set beside the figure and keep it fixed. Two tools can report different shares for the same company if one counts a company once per answer and the other counts every time it is named. It measures how often, not how favorably. Bing's Citation Share, described below, is a different metric.
Position
Definition. Where the company first appears in an answer that lists several options, such as first or third. Several tracking tools report it as a rank or an average position; AEO and AI visibility tools compared notes which.
Limits. Order is the least stable part of an answer: two brand lists in the same order turned up about once in 1,000 runs in the study above, and its author called tracking "ranking position" in AI tools "foolhardy" (opens in a new tab). AEO HQ does not recommend position as a KPI. If you report it, report the share of runs in which the company was named first, with the number of runs.
Sentiment
Definition. How favorably an answer describes the company, usually graded positive, neutral, or negative by a language model or a person. HubSpot's sentiment score shows "how positively or negatively your brand is described in answer engine responses" (opens in a new tab).
Limits. Sentiment moves far more than mentions do: whether a brand was framed positively or negatively flipped about 6.7 times more often than whether it was mentioned (opens in a new tab) (preprint). A change in one window is weak evidence of a real change. AEO HQ's recommended design reports sentiment as description only, not as a KPI (what the design reports).
Accuracy
Definition. The share of answers to branded prompts that state the company's facts correctly, graded by a person against a written fact sheet: what the company sells, its prices, who it serves, its locations, and its founders.
How to compute it. Accuracy = answers graded correct ÷ answers graded. Record every error and the page the answer cited for it, because that page is where the fix starts.
Limits. It needs a person and a current fact sheet, and it covers only the branded prompts you ask. The evidence that assistants misstate facts is covered under "Counting mentions without checking what the answer says" in AEO antipatterns.
Composite visibility scores
Definition. A single number that a vendor builds from several measures. HubSpot's free AI Search Grader, for example, publishes its weights: sentiment up to 40 points, presence quality 20, brand recognition 20, share of voice 10, and market competition 10 (opens in a new tab), scored from what ChatGPT, Perplexity, and Gemini say about a brand "based on their training data" (opens in a new tab).
Limits. A score hides which part moved. The 2026 review of 45 studies says a single score is defensible only when its weights correspond to an explicit objective, and that combining a mention, an accurate citation, and a conversion without a model of their value "merely obscures normative choices" (opens in a new tab). AEO HQ's own automated audit reports a readiness score that a language model writes without a fixed formula, and the methodology page says so. Our recommendation: report the parts separately, and if you use a vendor's score, get its formula and weights in writing.
Metrics from Google's and Microsoft's own reports
Google and Microsoft give site owners first-party counts of their pages in AI answers. OpenAI's publisher FAQ points publishers to analytics platforms such as Google Analytics for tracking referral traffic from ChatGPT (opens in a new tab) and describes no citation report.
AI feature impressions (Google Search Console)
Definition. Search Console's generative AI performance report counts impressions in AI Overviews and AI Mode. Google defines them this way: "Impressions are how many times links to your site were shown to a user in a generative AI feature on Google Search." (opens in a new tab)
How it is counted. If two results from the same site appeared in one generative AI feature, they count as a single impression in the chart total, and the table can group data by page, country, date, or device (opens in a new tab). The report has been available to all websites worldwide since August 31, 2026 (opens in a new tab).
Limits. It counts impressions, not clicks: clicks from AI Overviews and AI Mode are counted in the Performance report under the Web search type and are not shown separately (opens in a new tab). A site may not see the report until it has received enough impressions in these features (opens in a new tab). It covers Google only, and it cannot see an answer that names you without linking to you.
Citations and cited pages (Bing Webmaster Tools)
Definition. Bing's AI Performance report counts citations of your pages across Microsoft Copilot, AI-generated summaries in Bing, and "select partner integrations" (opens in a new tab). Its main figures, in Microsoft's words:
- Total Citations: "the total number of citations that are displayed as sources in AI-generated answers during the selected time frame" (opens in a new tab).
- Average Cited Pages: "the average number of unique pages from your site that are displayed as sources in AI-generated answers per day" (opens in a new tab).
- Grounding queries: "the key phrases the AI used when retrieving content that was referenced in AI-generated answers" (opens in a new tab), shown as a sample.
Limits. The measures Microsoft lists count citations; none counts clicks (opens in a new tab), and Microsoft says its page-level counts reflect how often pages are cited, "not page importance, ranking, or placement" (opens in a new tab). It covers only Microsoft's own products and the partners it does not name.
Citation Share (Bing Webmaster Tools)
Definition. Added in June 2026, Citation Share is "the percentage of citations attributed to your site out of all citations shown across all sites for that same grounding query" (opens in a new tab).
Limits. Microsoft calls it an observational metric, "not a ranking system or a competitive scoreboard," that "does not expose competitor domains, represent traffic share, or assign quality scores to content" (opens in a new tab). It differs from share of voice in three ways: it counts citations rather than mentions, it is calculated for one grounding query at a time, and it compares you with all cited sites rather than a named set of competitors.
Traffic and outcome metrics
AI referral sessions
Definition. Visits that begin with a click on a link in an assistant's answer. Google Analytics 4's AI Assistant channel counts visits from sources like ChatGPT, Gemini, Deepseek, Copilot, or Grok, and excludes Google's AI Overviews and AI Mode (opens in a new tab), which it counts as organic search. Microsoft Clarity's AIPlatform channel captures sessions that begin with a click on "a free, organic link within an AI platform's chat experience," and a separate PaidAIPlatform channel captures clicks on paid ads within those platforms (opens in a new tab).
How to compute it. Count sessions in a channel defined by the assistants' web addresses; the setup is in how to track AI referral traffic in GA4. Share of traffic = AI referral sessions ÷ all sessions.
Limits.
- It is small. In a February 2025 study of 3,000 sites, 0.17% of the average site's visitors came from AI chatbots (opens in a new tab) (vendor study).
- It undercounts. Traffic referred by Claude's native app does not include a Referer header (opens in a new tab) (network measurement), so such visits are recorded as direct.
- Most answers produce no click. Google users who saw an AI summary clicked a link in the summary itself in 1% of visits (opens in a new tab) (browsing data from 900 U.S. adults, March 2025).
AI key events and conversion rate
Definition. Key events are actions you mark as important in Google Analytics 4, such as a form submission or a purchase. GA4's session key event rate is "the percentage of sessions in which a user triggered a key event" (opens in a new tab).
How to compute it. Count key events in AI referral sessions. AI session key event rate = AI referral sessions with a key event ÷ AI referral sessions. Compare it with other channels only when each has enough sessions to be meaningful.
Limits. Whether AI visitors convert better is contested. Across 973 e-commerce sites, ChatGPT referrals converted above paid social but below all other traditional channels (opens in a new tab) (peer-reviewed; the first author is employed by the company that supplied the data). One software company, by contrast, reported that AI search sent 0.5% of its visitors and 12.1% of its sign-ups over 30 days (opens in a new tab) (single site; the company sells SEO software). Counts based on clicks miss buyers who read an answer and arrive later another way, which the next metric addresses.
Self-reported AI source
Definition. The share of new leads or customers who name an AI assistant when asked how they first heard of the company.
How to compute it. Answers naming an AI tool ÷ all answers to the question. Report the response rate beside it, because the question is often optional.
Example. In one agency's records, 189 of 213 leads who answered the question named an AI tool, but first-touch attribution credited AI with only 28 of those 189; 95 were filed under Organic Search and 59 under Direct (opens in a new tab) (single firm; the agency sells AEO services). The agency treats self-report as a lower bound, because only leads who completed an optional field are counted (opens in a new tab).
Limits. People misremember, skip the question, or name the last thing they saw. It does not say which assistant, prompt, or answer they saw unless you ask.
AI-influenced pipeline
Definition. Deals and revenue in the CRM from contacts whose self-reported or recorded source is an AI assistant.
How to compute it. Sum pipeline or revenue for those contacts in a window. Keep contacts identified by self-report apart from those identified by click-based source, because the two methods disagree.
Limits. It shows association, not effect. The assistants' own growth raises AI numbers with no work on your side: in the only controlled field study we found, total ChatGPT referrals to one site grew 5.7 times while its untreated pages grew 3.5 times (opens in a new tab) (preprint; the authors work for the company that owns the site). To judge whether work paid off, compare changed pages or periods with unchanged ones, as that study did. AEO examples collects documented cases, including this one.
Access metrics: crawler and fetcher requests
Definition. Requests to your site from AI companies' agents, counted by agent and URL in server logs. Two kinds matter (crawler vs. fetcher): search crawlers, which build the index an assistant searches, and user-initiated fetchers, which visit a page while an assistant answers a person. OpenAI says ChatGPT-User visits pages when a user's question calls for it and is not used to crawl the web automatically (opens in a new tab); Anthropic says Claude-User accesses websites when people ask Claude questions (opens in a new tab); and Perplexity says Perplexity-User visits a page when a user asks Perplexity a question (opens in a new tab). Each company's agents are listed in the glossary under OpenAI crawlers, Anthropic crawlers, and Perplexity crawlers.
How to compute it. Requests per agent, per URL, per window, and the share answered with HTTP 200. Check that a request claiming to come from a crawler really does: OpenAI publishes IP address lists for its crawlers and fetchers (opens in a new tab), as do Anthropic (opens in a new tab) and Perplexity (opens in a new tab).
Limits. A fetch shows that an answer retrieved your page, not that it cited or recommended you. None of the major AI crawlers rendered JavaScript (opens in a new tab) in Vercel's December 2024 data, so these requests never appear in analytics tools that run in the browser. Some requests cannot be counted by name: Brave's crawler does not advertise a differentiated user agent (opens in a new tab).
Choosing KPIs
These are AEO HQ's recommendations. Pick one KPI for each question the business needs answered, and keep the rest as supporting metrics.
| Question | KPI | Supporting metrics |
|---|---|---|
| Do assistants recommend us to buyers? | Mention rate on unbranded prompts, for each assistant, over four-week windows, with a 95% interval | Share of voice; prompt coverage at a fixed number of runs |
| Do assistants use our pages as sources? | Citation rate | AI feature impressions; Bing citations and Citation Share |
| Do they describe us correctly? | Accuracy on branded prompts | A log of errors and the pages cited for them |
| Is it producing business? | Self-reported AI source and AI-influenced pipeline | AI referral sessions and key events |
| Can assistants reach our pages? | Share of priority pages that return HTTP 200 to AI crawlers and fetchers | Requests per agent and URL |
Five rules keep the numbers honest:
- Report each assistant separately. A pooled figure describes none of them.
- Attach the method to every figure: the prompts, the number of runs, the window, and the interval.
- Mark events on the timeline, such as model releases, panel changes, and site changes.
- Do not combine the KPIs into one score.
- Judge change against a comparison, such as unchanged pages or an earlier period measured the same way.
AEO metrics compared with SEO metrics
| SEO metric | Closest AEO metric | What differs |
|---|---|---|
| Ranking position | Mention rate | Answers rarely repeat the same order (opens in a new tab), so a rate across runs replaces a single position |
| Search impressions | AI feature impressions | Google reports them in a separate report, as impressions only (opens in a new tab) |
| Clicks and click-through rate | AI referral sessions | Links in Google's AI summaries were clicked in 1% of visits (opens in a new tab) |
| Share of voice in search | Share of voice in answers | Counted across runs of fixed prompts against a named competitor set |
| Organic conversions | Self-reported AI source; AI key events | Click-based attribution credited 28 of 189 self-reported AI leads (opens in a new tab) at one agency |
What the evidence shows and does not show
Antipatterns
An antipattern is a practice that looks helpful but fails or backfires. AEO antipatterns covers four measurement mistakes in depth. Seven more concern the metrics themselves:
- Comparing prompt coverage across different numbers of runs. Coverage rises with runs alone, as the calculation above shows. Instead, compare at the same number of runs.
- Dividing citations by cited answers only. It drops the runs where the assistant did not search (opens in a new tab). Instead, divide by all runs.
- Treating a citation as a recommendation. A site can be read often and named rarely (opens in a new tab). Instead, report mention and citation rates separately.
- Tracking position as a KPI. Order almost never repeats (opens in a new tab). Instead, use the mention rate.
- Reporting a score without its weights. A score needs weights tied to a stated objective (opens in a new tab). Instead, report the parts, or publish the formula.
- Crediting all growth to the work. Untreated pages on one site grew 3.5 times (opens in a new tab) over the same period. Instead, compare with unchanged pages or periods.
- Counting fetches as citations. A fetch is retrieval, not a recommendation. Instead, report log data as an access metric.
Checklist
| # | Check | How to verify | Pass when | Basis |
|---|---|---|---|---|
| 1 | Each metric is defined in writing | Read the report's method note | Numerator, denominator, window, and assistant are stated | Vendor definitions differ (opens in a new tab) |
| 2 | Every run is in the denominator | Recount from the raw answers | Runs without search or citations are included | Review of 45 studies (opens in a new tab) |
| 3 | Mentions and citations are reported separately | Read the report | Two columns | Cited but rarely named (opens in a new tab) |
| 4 | Coverage is compared at equal runs | Check runs per prompt in each period | The same number | Coverage calculation |
| 5 | Share of voice names its competitor set | Read the report | The set is listed and unchanged | Glossary definition |
| 6 | Position is not a KPI | Search the report for "rank" and "position" | Not used as a KPI | Order rarely repeats (opens in a new tab) |
| 7 | Any score shows its weights | Ask the vendor | The formula is in writing | Review of 45 studies (opens in a new tab) |
| 8 | First-party reports are connected | Open Search Console and Bing Webmaster Tools | Both are reviewed each window | Search Console (opens in a new tab); Bing (opens in a new tab) |
| 9 | AI referrals have their own channel | Open the GA4 channel settings | A channel matches assistant sources | GA4 channel definition (opens in a new tab) |
| 10 | Self-reported source is collected | Submit a test form | The answer is stored with the lead | Attribution gap (opens in a new tab) |
| 11 | Changes are judged against a comparison | Read the analysis | Unchanged pages or periods are tracked | Field study (opens in a new tab) |
FAQ
What are the most important AEO metrics to track?
Start with the mention rate on unbranded buyer prompts for each assistant, then the citation rate, share of voice, and accuracy on branded prompts. Add Search Console's AI feature impressions, Bing's citations, AI referral sessions, and a "How did you hear about us?" question. This set is AEO HQ's recommendation; no standard list exists.
How do I track AEO?
Use four instruments: a prompt panel run on each assistant, the reports Google and Microsoft give site owners, analytics with a channel for AI assistants, and a question on your forms about how buyers found you. The measurement guide linked at the top of this page sets out the steps.
What is a good mention rate?
No published benchmark exists, because the rate depends on the category, the prompts, and the assistant. As context only, one panel found that first answers named household brands 73% of the time, mid-market brands 44%, and small brands 11% (opens in a new tab) (preprint; the author co-founded the platform studied). Track your own rate against your own earlier windows and your named competitors.
How are AEO KPIs different from SEO KPIs?
SEO KPIs count positions, impressions, and clicks in a results page. AEO KPIs count how often answers name, cite, and correctly describe a company across repeated runs, because answers change from run to run. The comparison table above maps each SEO metric to its closest AEO metric.
Which AEO metrics show ROI?
No metric shows return on investment by itself. Combine self-reported AI source and AI-influenced pipeline with a comparison against unchanged pages or earlier periods, because the assistants' own growth raises AI numbers for everyone. The field study above found untreated pages growing 3.5 times (opens in a new tab) over the same period.
How often should AEO metrics be reported?
AEO HQ recommends rolling four-week windows, reported for each assistant. HubSpot gives similar advice for its own tool: "Because responses change over time, review performance across multiple days or weeks before evaluating trends." (opens in a new tab)
What is the difference between share of voice and Citation Share?
Share of voice counts your mentions against a named set of competitors across your prompt panel. Bing's Citation Share counts your citations against all citations for one grounding query in Microsoft's AI answers, and does not expose competitor domains (opens in a new tab).
Can Google Analytics measure AEO on its own?
No. Analytics records visits, not answers. It misses answers that produce no click and visits that arrive without a referrer, and it cannot separate clicks from AI Overviews and AI Mode, which GA4 counts as organic search (opens in a new tab).
Next steps
AEO HQ's Instant AEO Audit ($499) is an automated check of a site's crawler access, key pages, and schema, with a five-question sample from one AI model. It does not calculate the rates on this page; the methodology page lists what it measures and what it does not.
Change log
- September 28, 2026: First published.
Sources
- Fishkin, R. (2026, January 28). NEW research: AIs are highly inconsistent when recommending brands or products; marketers should take care when tracking AI visibility. SparkToro. https://sparktoro.com/blog/new-research-ais-are-highly-inconsistent-when-recommending-brands-or-products-marketers-should-take-care-when-tracking-ai-visibility/ (opens in a new tab)
- Kumar, P. (2026). Generative engine optimization at scale: Measuring brand visibility across AI search engines (arXiv:2606.20065). arXiv. https://arxiv.org/abs/2606.20065 (opens in a new tab)
- HubSpot. (n.d.). Show up in AI search with answer engine optimization (AEO) [Product page]. Retrieved September 27, 2026, from https://www.hubspot.com/products/marketing/aeo-guide (opens in a new tab)
- HubSpot. (2026, August 27). Set up and analyze AEO [Knowledge base article]. HubSpot Knowledge Base. https://knowledge.hubspot.com/seo/set-up-and-analyze-ai-visibility (opens in a new tab)
- Schulte, J., Bleeker, M., & Kaufmann, P. (2026). Don't measure once: Measuring visibility in AI search (GEO) (arXiv:2604.07585). arXiv. https://doi.org/10.48550/arXiv.2604.07585 (opens in a new tab)
- Martinez, O. (2026). Optimizing visibility in generative engines: A critical survey of generative engine optimization (2023–2026) (arXiv:2607.14035). arXiv. https://doi.org/10.48550/arXiv.2607.14035 (opens in a new tab)
- Hong, A. (2026, September 5). 12 best AI SEO, AEO & GEO agencies in 2026, scored. Tobe Agency. https://www.tobeagency.co/learn/12-best-ai-seo-agencies-in-2026-scored-on-what-they-can-prove (opens in a new tab)
- HubSpot. (n.d.). AI Search Grader [Free tool]. Retrieved September 27, 2026, from https://www.hubspot.com/aeo-grader (opens in a new tab)
- OpenAI. (n.d.). Publishers and developers – FAQ. OpenAI Help Center. Retrieved September 27, 2026, from https://help.openai.com/en/articles/12627856-publishers-and-developers-faq (opens in a new tab)
- Google. (2026). Generative AI performance report (Search) [Search Console Help]. Retrieved September 27, 2026, from https://support.google.com/webmasters/answer/16984139 (opens in a new tab)
- Google. (2025, December 10). AI features and your website. Google Search Central. https://developers.google.com/search/docs/appearance/ai-features (opens in a new tab)
- Madhavan, K., Merchant, M., Canel, F., & Nigam, S. (2026, February 10). Introducing AI Performance in Bing Webmaster Tools public preview. Bing Webmaster Blog. https://blogs.bing.com/webmaster/February-2026/Introducing-AI-Performance-in-Bing-Webmaster-Tools-Public-Preview (opens in a new tab)
- Madhavan, K., Merchant, M., Nigam, S., & Shah, T. (2026, June 16). New AI visibility insights in Bing Webmaster Tools: Intents, topics, citation share, compare. Bing Search Blog. https://blogs.bing.com/search/June-2026/New-AI-Visibility-Insights-in-Bing-Webmaster-Tools-Intents-Topics-Citation-Share-Compare (opens in a new tab)
- Google. (2026). Default channel group [Analytics Help]. Retrieved September 27, 2026, from https://support.google.com/analytics/answer/9756891 (opens in a new tab)
- Microsoft. (2025, September 23). AIPlatform and PaidAIPlatform. Microsoft Learn (Clarity). https://learn.microsoft.com/en-us/clarity/insights/ai-channel-group (opens in a new tab)
- Linehan, L. (2025, February 6). 63% of websites receive AI traffic (new study of 3,000 sites). Ahrefs. https://ahrefs.com/blog/ai-traffic-study/ (opens in a new tab)
- Belson, D., & Rhea, S. (2025, July 1). The crawl before the fall... of referrals: Understanding AI's impact on content providers. Cloudflare Blog. https://blog.cloudflare.com/ai-search-crawl-refer-ratio-on-radar/ (opens in a new tab)
- Chapekis, A., & Lieb, A. (2025, July 22). Google users are less likely to click on links when an AI summary appears in the results. Pew Research Center. https://www.pewresearch.org/short-reads/2025/07/22/google-users-are-less-likely-to-click-on-links-when-an-ai-summary-appears-in-the-results/ (opens in a new tab)
- Google. (2026). [GA4] Traffic acquisition report [Analytics Help]. Retrieved September 27, 2026, from https://support.google.com/analytics/answer/12923437 (opens in a new tab)
- Kaiser, M., & Schulze, C. (2026). ChatGPT referrals to e-commerce websites: How do LLMs compare against traditional channels? Marketing Science, 45(4), 699–715. https://doi.org/10.1287/mksc.2025.0489 (opens in a new tab)
- Stox, P. (2025, June 16). Does AI search traffic convert better than traditional search? For Ahrefs, yes: 0.5% of visitors drove 12.1% of signups. Ahrefs. https://ahrefs.com/blog/ai-search-traffic-conversions-ahrefs/ (opens in a new tab)
- Birkett, A. (2026, August 28). First-touch attribution captures 15% of our AI-sourced leads [Research]. Omniscient Digital. https://beomniscient.com/blog/first-touch-vs-self-reported-attribution-aeo/ (opens in a new tab)
- Watanabe, K., & Nakayashiki, K. (2026). Disentangling answer engine optimization from platform growth: A log-based natural experiment on ChatGPT referral traffic (arXiv:2606.04362). arXiv. https://doi.org/10.48550/arXiv.2606.04362 (opens in a new tab)
- OpenAI. (n.d.). Overview of OpenAI crawlers. OpenAI API documentation. Retrieved September 27, 2026, from https://developers.openai.com/api/docs/bots (opens in a new tab)
- Anthropic. (2026, April 7). Does Anthropic crawl data from the web, and how can site owners block the crawler? Claude Help Center. https://support.claude.com/en/articles/8896518-does-anthropic-crawl-data-from-the-web-and-how-can-site-owners-block-the-crawler (opens in a new tab)
- Perplexity. (n.d.). Perplexity crawlers. Perplexity Docs. Retrieved September 27, 2026, from https://docs.perplexity.ai/guides/bots (opens in a new tab)
- Zecchini, G., Moore, A. A., Ubl, M., & Siddle, R. (2024, December 17). The rise of the AI crawler. Vercel. https://vercel.com/blog/the-rise-of-the-ai-crawler (opens in a new tab)
- Brave. (n.d.). Brave Search crawler. Brave Search Help. Retrieved September 27, 2026, from https://search.brave.com/help/brave-search-crawler (opens in a new tab)
How to cite this page
Maxwell, P. (2026). AEO metrics and KPIs: definitions. AEO HQ. Last updated September 28, 2026. https://www.aeohq.ai/articles/aeo-metrics-and-kpis
More in Measuring AI visibility
Complete guide
How to measure AI visibility
How to measure AI visibility: what to count, how many prompts and runs, error ranges, and the Google, Bing, and GA4 reports that fill the gaps.
Guide
How to track AI referral traffic in GA4
Track AI referral traffic in GA4: what the AI Assistant channel counts, a custom channel for ChatGPT, Claude, Perplexity, and others, and what GA4 cannot see.
Reference
AEO and AI visibility tools compared
AEO and AI visibility tools by kind: what search engine reports, analytics, prompt trackers, graders, and log tools can and cannot measure, from vendors' pages.
Reference
How AEO HQ measures AI visibility, and how our research is done
How AEO HQ grades evidence, pulls search data, and measures AI visibility, and what its $499 automated audit does and does not measure today.