How should awareness be scored?
Weight awareness by query value. A citation on a comparison query where buyers are choosing is worth more than ten citations on definitional queries, and an unweighted count will tell you the opposite.
Semrush found the intent mix shifting underneath everyone: keywords triggering AI Overviews went from 89.03% informational in October 2024 to 57.16% in October 2025, with commercial and navigational queries picking up the difference.
That shift is the reason to reweight now rather than later. The feature is moving into the queries where money changes hands, so an awareness score built on informational appearances is measuring the part of the surface that matters least and shrinking fastest. Score citation rate and mention rate per query group and multiply by a commercial weight you set yourself.
How should traffic be scored?
Score traffic against the control group. And when visibility climbs while clicks fall, flag it for review instead of filing it as damage.
Position matters more than it used to. Amsive measured a 27.04% click-through decline for keywords outside the top three when an AI Overview appeared, against a 15.49% average across the full keyword set.
So a page sitting at position six on an AI Overview query is losing clicks roughly twice as fast as the average, which makes position four-to-ten content the first place to look when traffic drops and the last place worth defending with more of the same content. Track clicks and branded search side by side. Rising impressions with falling clicks and rising branded search is a discovery gain wearing a traffic loss costume.
How should revenue be scored?
Score revenue directionally and say so in the report. The honest method is to correlate movement in affected query groups with pipeline.
Similarweb recorded referrals from generative AI platforms to transactional sites growing 357% year over year as of Q4 2025, with AI referral traffic converting at roughly 7%, matching traditional search on rate while arriving in far smaller volume.
Small volume converting at parity is an early channel. Track assisted conversions and pipeline against your affected query groups, then compare those figures with an unaffected group over the same window. If the affected group holds revenue while losing sessions, your discovery improved and your reporting was the only thing that got worse.
Verification catches harmful visibility
Sample the actual answers. Automated counts tell you that you appeared, and only a human reading the text can tell you whether the appearance helped or hurt.
Accuracy is not a given. Oumi evaluated Google AI Overviews against the SimpleQA benchmark and published on April 14, 2026 that roughly 50% of overviews were untrustworthy, with hallucination rates rising between the Gemini 2 and Gemini 3 versions powering the feature.
A coin-flip accuracy rate on a surface reaching billions of queries means a rising citation count can be a rising misinformation count, and your dashboard would show both as green. Pull 20 to 30 sampled results per quarter and check each one:
-
Is the brand cited and described accurately?
-
Does the framing favor you, and do you appear at all on category-level prompts?
-
Do competitors dominate the recommendation, or do unsupported claims about your product appear?
Trends matter more than snapshots
Judge movement over weeks, never over a single check, because AI Overviews rewrite themselves faster than any Google surface you've tracked before. One bad Tuesday is noise.
Authoritas measured this across 11,203 keywords on Google.com and scored AI Overview ranking volatility at 0.68 against 0.49 for organic rankings, which showed that around 70% of the pages appearing in AI Overviews change over two to three months.
At that rate of churn, monthly rank-tracking habits sample far below the speed of change, and a team reacting to one week's disappearance will rewrite a page that was going to come back on its own. Check priority queries weekly under identical conditions, then judge sustained direction across four to six observations. Read that movement next to your broader organic performance and conversion trend before anyone rewrites the content strategy. Use AI visibility as a complementary signal rather than relying on rankings alone.
Monitor discovery systematically with Snoika
The measurement approach in this article needs a monitoring layer that runs without you, because checking citations by hand at the cadence volatility demands is not sustainable past a few dozen queries. Snoika is an AI visibility platform, founded by Anton Vedeshin and registered in Estonia, that tracks brand mentions and competitor visibility across ChatGPT and Google AI Overviews.
The platform simulates real user prompts on a weekly schedule and reports how often and how favorably your brand appears, which covers the awareness layer of the scorecard. You keep Search Console for impressions and analytics for sessions and engagement, plus your CRM for pipeline. Snoika fills the column those three cannot produce.
Start with the free AI visibility check. Run your 20 highest-value non-branded queries through it and compare them against the same 20 in 30 days.