From Blue Links To Answers: How Search Changed
Finally, pay attention to how they talk about their existing clients. Somebody who describes a client's category accurately, names the specific constraint that made the work difficult, and mentions something that did not work has actually done the job. Somebody who describes every engagement as a success in identical language has either been unusually lucky or is describing a template.
Assistant measurement is not there yet. There is no console reporting how often you were named, answers vary between sessions and accounts, and referral traffic is attributed inconsistently across assistants. The honest approach is a fixed prompt set run on a schedule, with the raw answers kept, and any tool metric attributed to the tool that produced it.
Where to Get Real Language Four sources, all of which you already own. Sales call notes, where prospects describe their problem before anyone corrects their terminology. Support tickets, where customers describe things going wrong in their own words.
The other practical difference is in how quickly work shows up. A ranking change takes weeks to settle and then holds reasonably steady. A citation can appear within days of publishing and disappear just as quickly when a fresher source arrives. Planning that assumes search-like stability will read normal volatility here as failure, which is how sound programmes get cancelled in their second quarter.
And read the raw text periodically rather than only the tallies. Changes in how you are described, from hedged to definite or from generic to specific, often precede changes in whether you appear at all, and no counting method will surface that. ai search optimization
Keep a small number of deliberately hostile prompts in the set permanently. Questions asking whether you are expensive, slow or suitable only for large clients reveal what the system believes about your reputation, and the belief is often traceable to one specific source. Nobody enjoys reading those answers, and they generate more actionable work than the flattering prompts do.
Ask to See a Prompt Set The first question is the most revealing. Ask them to show you the prompt set from a current or recent client, with the client's name removed. A team doing real work has this and will show it, because the prompts are craft rather than secret sauce.
Because there is no independent scoreboard in this channel, an engagement can run for a year on the strength of a number the supplier produces. That is an unusual amount of trust to extend, and it makes knowing what to check more important here than in any other marketing channel.
The Signals That Mean Something Four things are hard to fake and worth watching closely. Your own pages beginning to appear in cited sources, which is directly observable in any assistant that shows citations.
The same caution applies to referral growth figures, which circulate widely without their context. One widely shared statistic showing several hundred percent growth in assistant referrals came from a sample of nineteen analytics properties. That is a real observation and a genuinely small sample, and the difference matters when you are deciding where to move budget.
Weight toward the commercial tiers. Roughly a third on buying intent, a quarter on evaluation, a quarter on problem framing and the remainder split between definitional and branded is a reasonable starting distribution.
One check is worth running independently once a quarter, without telling anyone. Take ten prompts from the agreed set, run them yourself in a signed out session, and compare what you find against the most recent report. Broad agreement is reassuring. A consistent gap in the agency's favour is the single most informative finding available to you, and it is not something a report will ever surface.
Where the Work Is Genuinely the Same The foundations do not change. Crawlable pages, sane site structure, fast rendering, accurate structured data, internal links that reflect how topics relate, and content that answers a real question all serve both channels.
The Honest Uncertainty Anyone claiming precision about this channel is overselling. Retrieval behaviour changes without notice, published studies use small samples, and vendor research tends to flatter the vendor. Opollo's finding that AI referral traffic converted at 14.2 percent against 2.8 percent from search came from 312 business to business brands, and Opollo sells this service.
If you must change the prompt set, add new prompts as a separate cohort and keep the original series running unchanged. Editing the instrument retrospectively destroys the comparison you have been building.
Why One Snapshot Proves Almost Nothing Generation involves randomness, and retrieval can return different pages between runs. The same prompt asked twice in a row can produce different companies in different orders.
Testing too rarely means you find out about a problem a quarter after it started. Testing too often means drowning in variance that looks like signal and reacting to noise. Both failures are common and the second is more expensive, because it produces work.