← All writing
GEO Aug 5, 2026 6 min read

The IAB's new AI visibility rules contain exactly one hard number

Christopher Dorsey

Christopher Dorsey

AI & MadTech Advisor · Enterprise Sales Leader

TL;DR

On August 3 the IAB published “Measuring Visibility in the AI Era,” 36 pages sorting AI visibility metrics into four categories and two quality tiers: directional, for internal briefings, and decision-grade, the tier it says you should use for budget allocation and picking a vendor. Nine dimensions separate the two. Exactly one carries a number: fewer than 50 queries is “exploratory” and cannot characterize a category. Decision-grade query volume is “large, diverse.” Decision-grade platform coverage is “a substantial majority” of consumer AI traffic. No figures for either, no certification program to check, only a “potential future” one with no date. So any of the 20-plus companies selling AI visibility tools can put decision-grade on a slide, and IAB says up front the document “does not rate measurement providers.” The useful part is the disclosure list underneath. Three dimensions are checkable today: coverage of all four query intent types with results segmented by intent, a weekly or faster testing cadence, and per-platform results reported separately with the weighting disclosed. Put those three in your next GEO renewal by name. A vendor selling you a blended monthly score is selling directional data, and IAB just wrote down what directional is for, which is not your budget.

The IAB published its AI visibility guidance on Monday, 36 pages called “Measuring Visibility in the AI Era.” It sorts metrics into four categories: presence (do you show up), prominence (where in the answer), portrayal (in what context, and is it accurate), and persuasion (does any of it drive a click). Then it grades every measurement claim as either directional, which IAB says is fine for competitive awareness and internal briefings, or decision-grade, which is what you’re supposed to require before you allocate budget or pick a provider.

Nine dimensions separate those two tiers. One of them has a number in it. Fewer than 50 queries per measurement program is “exploratory,” a rung below even directional, because query volumes under that floor “cannot meaningfully characterize a category.” That’s the whole quantitative content of the document.

Decision-grade is defined in words nobody can check

Look at what the top tier asks for. Decision-grade query volume is a “large, diverse query set with category and subcategory coverage.” Decision-grade platform coverage is platforms “collectively representing a substantial majority of consumer AI traffic in the target market.” Decision-grade sample size is “enough responses per query to establish a stable distribution.” Large. Substantial. Enough. IAB’s own working group flagged that the acceptable variability ranges are still an open question, which is honest and also means the range you’d test a vendor against doesn’t exist yet.

IAB is candid about the limits, to its credit. The document says plainly that it “does not rate measurement providers, prescribe tools, or build measurement systems.” A provider certification program is described as a “potential future” thing, with no date attached anywhere I could find. So starting this week, any of the more than 20 companies selling AI visibility tools can put “decision-grade” on a slide, and the marketer holding that slide has no mechanism to disagree.

Two things can be true. Caroline Giegerich, IAB’s VP of AI, told AdExchanger this deliberately isn’t a standard, because a standard requires stability and the market is in a mass transition. She’s right, and the document is better than what buyers had a week ago. It’s also going to get used as a badge long before it can function as a bar. Both of those land on the same buyer.

I’ve seen guidance arrive before the audit before

Digital advertising ran this sequence a decade ago with viewability. The industry agreed on what a viewable impression was well before it agreed on who was allowed to count one, and the gap between those two moments ran for years. What filled it was a verification business, and a lot of meetings where a brand and an agency compared two vendors’ numbers for the same campaign and discovered they weren’t comparing anything. I sat in those meetings. Nobody was lying. Everybody was measuring a slightly different thing and calling it the same word.

The setup right now rhymes. Similarweb found only 11% of citations overlap between major AI platforms, and that citation sets churn roughly half every month. Semrush tracked 126 million US prompts from January through April and found 36 brands out of more than 1,200 that showed up in the top-100 most-mentioned on every platform every month. That’s the underlying instability two vendors are both trying to sell you a stable number about. IAB’s own summary puts the buyer’s position bluntly: “Buyers have budgets ready to spend. They have no basis for evaluating what they are buying.”

The scale is why this is worth your Monday. ChatGPT is over 900 million weekly actives, Google’s AI Overviews reach more than 2.5 billion people a month, and only 16% of brands systematically track their visibility inside any of it. McKinsey’s projection, cited in the IAB doc, is that unprepared brands could see traditional search traffic fall 20 to 50%.

Three of the nine dimensions are checkable today

The document earns its keep below the tier labels. Underneath them sits a disclosure list, and disclosure is verifiable even when the threshold isn’t. Three items in particular you can write into a contract this quarter and hold a vendor to without waiting on anybody’s certification.

Query intent coverage. IAB defines four types: informational (“what is X”), comparison (“X vs. Y”), recommendation (“best X for Y”), and transactional (“where to buy X”). Directional requires two. Decision-grade requires all four, with the distribution disclosed and results segmented by intent type. That segmentation is the most useful thing in the framework for anyone running a category, because “best running shoes for flat feet” and “where to buy X” are different problems with different owners and different budgets, and a blended score buries which one you’re losing.

Testing cadence. Monthly or quarterly is directional. Weekly or more frequent is decision-grade. Given citation sets that turn over half their contents in a month, a quarterly report is a photo of something that already left.

Per-platform reporting. Decision-grade providers must report each platform separately, show how much results vary across platforms, and document the weighting. I wrote in July that a blended AI-visibility number averages two unrelated games and hides losing one of them. The IAB has now made the split a condition of its top tier, which is a good outcome and also means your current vendor’s single number is, by IAB’s own definition, directional data.

The prominence category is the other quiet win. Placement inside an answer is now a named metric with ranking order and depth of citation attached, which is the thing I argued was a dial Google turns category by category when it moved recipe links to the top of AI Mode. If your report says “cited” and stops there, it’s answering the 2025 question.

What to do before your next renewal

Stop asking vendors whether they’re decision-grade. They’ll say yes, and under this document they’re entitled to. Ask for the disclosures instead: total query volume and the count per subcategory, the intent-type distribution, responses per query, results broken out by platform with the weighting basis, and the variation range they’ll commit to inside a seven-day window. A provider running a real program produces that in a few days. A provider running a dashboard over a thin query set will negotiate about it, and the negotiation is your answer.

One small thing I noticed and can’t unsee. The only outside voice IAB put in its own press release is a senior product manager at Microsoft Clarity, a company that ships an AI visibility product. It doesn’t make the guidance wrong. It does tell you who was paying attention on day one, and how fast the word “decision-grade” is going to travel.

Send your GEO vendor one email this week asking for last quarter’s results segmented by the four intent types and broken out by platform. Whatever comes back, and however long it takes, you’ll know which tier you’ve been buying.

Share this post

About the author

Christopher Dorsey

Christopher Dorsey

Enterprise Sales Leader · AI Go-To-Market · Startup Advisor · Denver, CO

Fifteen years selling technology to Fortune 500 brands across AI, advertising, and data infrastructure — most recently at Zeta Global, Oracle, and Fastly. Currently advising founders and sales leaders on AI go-to-market and Generative Engine Optimization.

Questions, pushback, or just want to compare notes?