What An AI SEO Agency Should Report Every Month
Refusals matter. A report that only contains successes is either describing a suspiciously easy month or omitting the parts that did not work, and the omitted parts are usually where the useful information is.
What Brands Usually Get Wrong in Response The instinctive response is to publish more brand content, which addresses none of the above. The second instinct is to try to displace the review site, which is not achievable and would not help if it were.
What the Evidence Actually Is The figure quoted most often comes from Opollo, which reported assistant referred traffic converting at 14.2 percent against 2.8 percent from conventional search. The sample was 312 business to business brands, attributed through UTM parameters, covering the third quarter of 2024 through the first quarter of 2025.
The Structural Reason A system composing a recommendation needs to weigh several options against each other. A review site has already done that. A brand site argues for one option and has an obvious interest in the conclusion.
One scheduling detail improves comparability more than it should. Run on roughly the same date each month rather than whenever somebody remembers. Retrieval behaviour and the freshness of competing sources both vary over a month, and a series taken at irregular intervals introduces variation that looks like a trend.
Results Split by Intent, With Run Counts Not one number. Mention rate reported as a fraction with the run count visible, generative Engine Optimization broken out by prompt tier, so buying intent is never blended with definitional questions.
This section sounds procedural and it is the foundation of everything after it. A prompt set quietly edited between runs makes every trend line in the document meaningless, and it is the easiest way to manufacture improvement without doing anything.
Log the conditions with every run, including which assistant, which mode, whether web access was enabled and the date. When a result moves sharply, the conditions log is usually what tells you whether the world changed or your setup did.
The defensible version states the mechanism, cites the available evidence with its sample sizes, presents your own segmented data however thin, and is explicit that most of the channel's value is not measurable through referrals at all.
Good versions read like this: mention rate on evaluation prompts rose from two in fifteen to six in fifteen, which we attribute to the three directory corrections completed in week two, though a competitor also stopped publishing during the same period.
The Prompt Set, Unchanged The report opens with the prompt set used, versioned and dated, and a statement that it is identical to last month's. If it changed, the change is listed explicitly with a reason, and the previous series is kept alongside so comparisons remain honest.
Category Costs Rise as Coverage Fills In Influencing the third party sources assistants cite is easiest while those sources are thin. A category with two mediocre comparison articles is inexpensive to influence. The same category in three years, once somebody has built the definitive resource that every assistant settles on quoting, is not.
Run a commercial prompt in almost any category and look at what gets cited. Review platforms, roundups and comparison sites appear first and most often, and the brands being discussed appear well down the list if at all.
Corroboration Beats Assertion The single clearest pattern in observed behaviour is that independent agreement outweighs self description. A claim made only on your own site is treated as a claim. The same claim appearing on a review platform, in a trade publication and in a forum thread is treated as a fact about the world.
How to Use This Honestly in a Business Case Do not build a return calculation on a borrowed conversion rate. Applying somebody else's percentage to an estimated mention volume produces a confident looking number resting on two guesses, and it will not survive the first person who asks where the inputs came from.
What Ranking Does and Does Not Buy You Ranking still helps, because the retrieval step usually starts with a search. But it buys far less than people assume. Ahrefs examined 15,000 long-tail prompts across four assistants in July 2025 and found roughly 80 percent of cited pages did not rank for the original query at all, with about 12 percent in the top ten.
This is also why review volume and recency show up so consistently in what gets cited. A platform with forty recent accounts of working with you is more informative than your own page saying customers love you, and it is treated accordingly.
This claim circulates constantly and it is usually presented with more confidence than the evidence supports. It is also probably directionally true, for reasons that are structural rather than mysterious.
What Padding Looks Like Screenshots of favourable answers with no indication of how many runs produced them. Industry news summaries that could have been written without opening your account. A rising score with no methodology. Traffic charts from unrelated channels included to fill space.