What To Ask Before Hiring An AI SEO Company
The distinction to draw is between flat results with the inputs completed, and flat results with the inputs missing. The first is a category or timing problem and may be worth persisting with. The second is a delivery problem.
Second, prompts that presuppose a weakness: is this company expensive, are they slow, are they suitable for small clients. The answers reveal what the system believes about your reputation, and where the belief is wrong it points at a specific source you can correct.
Pricing in this field is unusually opaque, partly because the work is new and partly because the absence of an independent scoreboard makes it hard for a buyer to tell whether they are getting value. That combination invites vague scoping.
When to Change Supplier Three conditions justify it individually. Raw answers cannot be produced on request. The prompt set has been changed without disclosure, which invalidates every comparison in every report you have received. Or two quarters have passed with the agreed inputs completed and no movement on citation presence, accuracy or source coverage.
A frequently quoted comparison showing assistant referrals converting several times better than search came from a vendor selling the service, across 312 business to business brands. A widely shared claim about explosive referral growth rested on nineteen analytics properties. Both are legitimate observations and neither supports the confident generalisation usually attached to them.
Verify the Fix Without Fooling Yourself Re-ask the same four questions quarterly rather than weekly, from a fresh signed out session. Identity work has slow feedback because scattered sources have to be re-crawled before the picture updates, and checking too often produces noise that looks like failure.
Ask Who Writes and Who Reviews Find out whether the writing is done by somebody with subject knowledge or generated and lightly edited. Both happen, and the second is not automatically disqualifying, but you need to know because you are the one who carries the liability for inaccurate claims about your own products.
What a Defensible Business Case Looks Like It states what cannot be measured. It reports inputs completed, with counts. It reports prompt set movement as fractions with visible run counts, split by intent. It includes the soft signals as anecdote clearly labelled as anecdote. It attributes every external statistic.
There is a defensible way to measure this. It produces less certainty than a paid media report and considerably more than a visibility score, and it has the advantage of surviving scrutiny. get your brand recommended by ChatGPT
Ask What They Will Not Do Good practitioners have a list. They will not guarantee a position in an answer, because nobody controls that. They will not fabricate reviews or seed forum threads under false identities, because it is detectable, damaging and increasingly enforced against.
Before leaving, make sure you take the prompt set, the baseline archive and everything published. If those were not yours under the contract, that is a lesson for the next agreement rather than something to negotiate at the exit.
Be wary of proposals where the largest line is content production. It is the easiest work to scale, the easiest to bill and the least likely to be the constraint, particularly before a baseline exists. A proposal weighted toward diagnosis, technical fixes and third party corrections is usually cheaper and almost always sequenced better.
Distinguish between a supplier who is failing and one who is reporting badly, because the remedies differ entirely. Ask for the raw answers and read them yourself before deciding. It is not unusual to find that sound work has been buried under a dashboard nobody understands, and fixing the reporting is far cheaper and less disruptive than replacing a team that is actually doing the job.
The risk is scope drift into activity that is easy to report and hard to value. The protection is to have the retainer specify countable units: prompt set runs per month, listings audited, corrections submitted, pages published or rewritten, outreach attempts made.
The answer you want describes a baseline: a prompt set built from how your customers actually speak, run across the assistants that matter, with raw answers and cited sources recorded. Everything after that should be justified by reference to what the baseline showed.
Pair your name with your sector and location consistently, rather than letting it appear alone. Correct third party listings that conflate you with the other business. Where the confusion is entrenched, consider whether a consistent descriptive phrase used alongside the name in all coverage is worth adopting.
The prompt set is the instrument, and almost every weak measurement programme in this field has a weak prompt set at the bottom of it. Get this wrong and everything downstream measures the wrong thing with great precision.
Screenshots of favourable answers with no run count, which say nothing about how many attempts produced them. Impressions or traffic from unrelated channels included to fill a report. And activity described in the language of effort, such as ongoing optimisation, with no countable output attached.