Here is the thing most people optimizing for AI search still get wrong, and it costs them: an answer engine almost never retrieves your page. It retrieves a passage. A few hundred words. Sometimes a single paragraph. Then it decides, from that fragment alone, whether your business is worth naming in the answer. The page you spent a week on is not the unit of competition. The paragraph is.
I've spent the last couple of years watching what actually gets pulled into AI answers versus what gets ignored, both in our own work at AIrecommend.ai and in the patterns behind our State of AI Search research. The businesses that win aren't the ones with the longest, most comprehensive pages. They're the ones whose individual paragraphs can stand on their own — get lifted out of context, dropped into a model's reasoning, and still say something true and specific. I call that property passage authority, and it's the most underpriced skill in AEO right now.
What actually gets retrieved
Modern answer engines — ChatGPT with search, Perplexity, Google's AI features, Claude — don't read your site the way a person does. When a query comes in, a retrieval layer breaks the web (and its index) into small pieces, usually called chunks or passages, and pulls the handful that look most relevant to the specific question. That short list of passages, not the full pages they came from, is what the model reasons over when it composes an answer.
This is a consequence of how retrieval-augmented generation works. The model has a limited context window and a strong incentive to fill it with the densest, most on-point material it can find. So the retriever's job is to find passages that match the query, and the generator's job is to synthesize an answer from those passages. Your 3,000-word cornerstone article doesn't enter the answer as 3,000 words. It enters as whichever 200 words the retriever decided were the best match — if any of them were.
That single fact reorganizes everything. If retrieval happens at the passage level, then visibility is won or lost at the passage level. A page can be authoritative overall and still lose, because the specific chunk that got pulled was vague, hedged, or buried three scrolls down under a wind-up nobody retrieved.
Why the page is the wrong unit
SEO trained a whole generation of marketers to think in pages. One page, one keyword, one ranking position. That model made sense when a human clicked a blue link and landed on the whole page, in order, top to bottom. The page was the delivery unit because the page was what got delivered.
In AI search, the page is not what gets delivered. A synthesized answer gets delivered, and your contribution to it is a fragment. So optimizing the page as a whole — its total word count, its overall topical coverage, its aggregate keyword density — optimizes a unit the machine never hands to the reader intact. You can have a beautifully complete page and contribute nothing to the answer, because none of your paragraphs were individually good enough to be pulled.
The inverse is also true, and it's the more useful insight: a modest page with three genuinely liftable paragraphs can out-perform a comprehensive one with zero. I've seen thin-looking pages get cited constantly because they happened to contain the one clean, self-contained sentence that answered the sub-question directly. The retriever doesn't grade on effort. It grades on fit.
What makes a passage "liftable"
After looking at a lot of retrieved-versus-ignored passages, the ones that get pulled tend to share a short list of properties. None of them are exotic. They're just rarely done on purpose.
Self-containment. A liftable passage makes sense with nothing above it. It doesn't open with "as we saw in the last section" or "this is why." It restates its own subject. If a retriever grabbed only that paragraph and showed it to you cold, you'd still know what it's about and what it claims. This is the single highest-leverage habit, and almost nobody does it, because good human writing flows and leans on what came before. Good machine-retrievable writing is more modular than that.
A claim, stated plainly. Answer engines are trying to answer a question. A passage that contains a direct, declarative claim — "The typical franchise breaks even in X, not Y" — is more useful to that job than a paragraph that circles a topic without committing to anything. Hedged, throat-clearing prose retrieves poorly because it doesn't give the model an answer to lift.
Specificity that can be corroborated. Named things — specific mechanisms, specific numbers, specific categories — give the model something concrete to attribute to you and to cross-check against other sources. Vague passages ("it depends on many factors") are safe to write and useless to retrieve.
Entity clarity. The passage should name who it's about. If the retrievable chunk refers to "we" and "our approach" with no anchor to an actual entity, the model has a harder time attributing the claim — or recommending the business behind it. Say the name.
The self-containment test
Here's a test you can run on any page in ten minutes. Copy each paragraph out, one at a time, into a blank document — no heading, no surrounding text. Read it cold. Ask: if this were the only thing an AI saw from my site, would it know what I'm talking about, and would it have a reason to name me? Most paragraphs fail this badly. They were written to be read in sequence, and stripped of their neighbors they collapse into pronouns and references to things that aren't there anymore. Every paragraph that fails the test is a paragraph that can't win a retrieval slot.
How to structure for passage authority
You don't fix this by writing shorter. You fix it by writing in retrievable units — treating each section as if it might be the only thing anyone ever sees from you, because for an answer engine, it might be.
Lead sections with the answer, then support it. The old "build up to the point" structure hides your best material below the fold of the retriever's attention. Put the claim in the first sentence of the section and use the rest of the paragraph to earn it. This is the opposite of narrative tension, and it feels wrong to writers, but it's how you get the sentence that gets pulled.
Give every meaningful sub-question its own passage. Answer engines fan a single user prompt out into many sub-questions — I've written about query fan-out before — and each sub-question is a separate retrieval. If your page bundles five distinct ideas into one dense block, you've entered one lottery. If it breaks them into five clean, self-contained passages under honest headings, you've entered five. Structure is how you multiply your chances.
Use headings that match how people ask, not how you'd file it internally. A heading is a strong retrieval signal because it tells the retriever what the passage beneath it is about. "How long until a franchise is profitable" retrieves better than "Timeline considerations," because the first matches a real question and the second matches nothing anyone types.
Repeat the subject, not the pronoun. Across a page, re-name the entity and the topic instead of leaning on "it" and "this." It reads slightly more repetitive to a human and dramatically clearer to a machine that just yanked one paragraph out of context. This is a genuine trade-off between human flow and machine legibility, and my read is that for content whose main job is to get you recommended, you lean toward the machine — but that's a judgment call, not a guarantee.
Measuring passage authority
You can't manage what you can't see, and passage-level performance is invisible in a normal analytics dashboard. What I actually look at is which specific claims and phrasings show up when an AI answers questions in your category. When a model paraphrases something that is clearly your framing — your term, your specific number, your way of cutting the problem — that's evidence that one of your passages made it into the retrieval set and survived into the answer. When the answer is generic and names competitors, your passages lost the slot.
The practical loop: identify the questions that matter in your category, ask them across the major engines on a schedule, and read the answers for your fingerprints. Not just "did they mention me" — which of my ideas came through, in whose words. That tells you which passages are earning their keep and which sections are dead weight that no retriever ever touches. Then you rewrite the dead sections to pass the self-containment test and watch whether they start showing up.
A worked example
Take a franchise consultant with a page titled "Our Process." Somewhere in the middle is a paragraph that reads: "Because of the factors described above, timelines vary considerably, and it's important to set realistic expectations before moving forward." That paragraph is a retrieval dead zone. It refers to "the factors described above" (gone the moment it's retrieved), it commits to nothing, and it names no entity. Pulled out of context, it says approximately nothing. No retriever will ever choose it, and if one did, the model would have no reason to attribute anything to this consultant.
Now rewrite it for passage authority: "Most franchise owners we work with reach operational break-even within the first year to eighteen months, depending mainly on the concept and the local market — not the two-to-three months some brokers imply." Same underlying point. But now it's self-contained, it makes a plain claim, it's specific enough to corroborate, and it draws a contrast that answers a real question people ask. That is a paragraph a retriever can pull and a model can use. Nothing about the page's length changed. One paragraph went from invisible to liftable, and that's the entire game.
Where this fits with the rest of your AEO
Passage authority isn't a replacement for the slower work of building entity authority and closing the corroboration gap — it's the layer that makes that work legible to a retriever. You can have a strong reputation and still lose retrieval slots because your pages bury their best claims. And you can write perfectly liftable passages and still struggle if no one else corroborates them. The two operate together: corroboration earns you the trust to be considered, and passage authority makes sure that when a retriever reaches for your material, there's a clean, specific paragraph waiting to be pulled. Fix the passages first, though, because it's the part that's entirely within your control and shows results fastest.
The mistake underneath all the other mistakes
Almost every AEO problem I get called in on traces back to the same root: the business is still optimizing the page when the machine is retrieving the paragraph. They add more words, more sections, more comprehensiveness — and none of it helps, because the problem was never coverage. The problem was that no single passage was clean enough, specific enough, or self-contained enough to get lifted.
Passage authority is unglamorous. It's not a growth hack. It's the discipline of writing paragraphs that can survive being torn out of their context and still say something true, specific, and attributable to you. Do that consistently and you stop competing for page rankings you can't see and start winning retrieval slots you can actually influence. That's the shift. The page was never the unit. The paragraph always was — we just couldn't see it until the machines started reading one paragraph at a time.
Key takeaways
- Answer engines retrieve passages — a few hundred words at a time — not whole pages. Visibility is won or lost at the paragraph level.
- A liftable passage is self-contained: it makes sense with nothing above it, states a plain claim, and names the entity it's about.
- Run the self-containment test — read each paragraph cold, with no context. If it collapses into pronouns and references, it can't win a retrieval slot.
- Break distinct ideas into distinct passages under question-shaped headings. Each clean passage is a separate entry in the retrieval lottery.
- Lead sections with the answer, then support it. Burying your best sentence below a wind-up hides it from the retriever.
- Measure which of your specific ideas and phrasings show up in AI answers — that's the fingerprint of a passage that survived into the response.
Frequently asked questions
Want to be the business AI recommends?
See how AIrecommend.ai builds the entity authority answer engines reward.
Explore AIrecommend.ai