Building Content That Language Models Quote
One Claim Per Sentence Compound sentences that bundle three ideas cannot be lifted without dragging in material that may not apply. A model faced with a passage where only part is relevant will often skip it in favour of a cleaner source.
Why the Direction Is Plausible Anyway Set the numbers aside and the mechanism is straightforward. Somebody arriving from an assistant has already had their question answered, has already seen a comparison, and has been given your name as a recommendation.
A useful way to think about the sequence is that each stage moved a task from the user to the interface. First the fact, then the summary, and now the comparison. Each move removed a reason to visit a website, and each was followed by an industry insisting the change had been overstated. It is reasonable to expect the pattern to continue rather than to stop at a convenient point.
Two caveats belong next to that number every time it is used. Opollo sells services in this space, so it is vendor research and interested. And business to business brands are not representative of retail, local services or consumer products.
What the Evidence Actually Is The figure quoted most often comes from Opollo, which reported assistant referred traffic converting at 14.2 percent against 2.8 percent from conventional search. The sample was 312 business to business brands, attributed through UTM parameters, covering the third quarter of 2024 through the first quarter of 2025.
Write Passages That Can Be Lifted Citation happens at passage level, not page level. A model attaches a source to a specific claim, which means the unit of work is a self contained paragraph that remains true and useful when removed from its surroundings.
How to Test Rather Than Trust Everything above is a starting hypothesis. Run twenty prompts in your own category across all three, from signed out sessions, recording the mode and the date, and count the cited domains for each.
The pattern is consistent across most categories. Review platforms, industry publications, documentation, forum threads and comparison articles appear far more often than brand websites. When a brand site is cited it is usually a specification page, a pricing page or a technical document rather than a homepage or a landing page.
Also watch what happens to your citations over time rather than checking once. A page that earns a citation and then loses it usually has a fresher competitor rather than a technical problem, and the fix is updating your figures rather than rewriting the page. Because retrieval runs live, that maintenance is cheap and it is the difference between a page that keeps earning and one that quietly stops.
Stage Two: The Comparison Moves Inside the Machine The current stage is more consequential. A generated answer does not just supply a fact, it performs the comparison the user would previously have done themselves by reading three results and forming a view.
Citation happens at the level of a passage, not a page. A model attaches a source to a specific claim it lifted, which means the real unit of work is a paragraph that stays true and useful once it has been removed from everything around it.
Ahrefs measured this in July 2025 across 15,000 long-tail prompts and four assistants, finding roughly 80 percent of cited pages did not rank for the original query, with about 12 percent in the top ten. The overlap is real but partial, which is the worst case for planning: you cannot ignore your rankings and you cannot rely on them either.
Two consequences follow immediately. Your page has to be findable by the underlying search step, and once fetched it has to contain a passage worth lifting. Failing either one keeps you out, and most brands fail the second.
The test that keeps this honest is simple. Show the rewritten page to somebody who buys from you and ask whether it is clearer. If the answer is no, no amount of extraction friendliness makes it a good page. get recommended by ai
Include Something Worth Attributing A citation needs something to point at. Passages that contain only sentiment give a model nothing, which is why brand pages full of adjectives are passed over in favour of a competitor's specification table.
One practical consequence of the variation between systems is worth planning for. If your customers are split across two assistants that behave differently, resist building separate programmes for each. The shared requirements account for most of the achievable outcome, and the effort spent on system specific tactics is usually better spent widening the number of third party sources that describe you correctly.
Perplexity is unusually useful to study because it shows its working. Every answer arrives with numbered citations you can click, which means you can reverse engineer what it rewards without guessing. Most assistants hide this. Perplexity puts it on the page.
Equally, do not publish a stripped alternate version of your site for crawlers. Serving different content to machines than to people is cloaking, it has been penalised for two decades, and there is no reason to expect a more forgiving treatment here.