What Is Source Selection?

In one sentence

Source selection is how an AI system decides which of the documents it retrieved actually get cited in the answer.

What Source Selection means

Retrieval produces candidates; selection narrows them to the handful that appear. Being retrieved is necessary but not sufficient.

Selection weighs relevance to the specific sub-question, extractability of the passage, credibility of the source, and how well the claim is corroborated elsewhere.

Why it matters for AI search

This is where most AI SEO effort should concentrate after retrieval is fixed. Plenty of sites are retrieved constantly and cited rarely.

It also explains why a lower-ranked page can be cited over a higher-ranked one — selection isn't ranking.

See it in action

Forty candidates, three citations. What separated them.

What decides selection

Candidate evaluation
user asks → which AI SEO agency should a B2B SaaS company hire?

High authority, vague page

Retrieved. Nothing specific enough to quote — no numbers, no method.

not cited

Mid authority, specific page

Retrieved. States pricing ranges and methodology explicitly. Corroborated on two platforms.

not cited

High authority, gated content

Retrieved the landing page. The substance is behind a form.

not cited

Authority lost to specificity twice.

Selection rewards checkable, extractable substance. A page that states real numbers beats a more authoritative page that states nothing a model can quote.

How to get it right

Improving source selection odds

  • Make sure passages are self-contained and quotable, not just present
  • State specifics — real numbers, documented methodology, named outcomes
  • Build corroboration so claims aren't single-sourced to your own domain
  • Ungate information you want cited; a model can't quote what's behind a form
  • Keep facts current, since stale sources get deprioritised in selection

Common questions

Why are we retrieved but not cited?

Usually because no passage is cleanly extractable, or because your claims aren't corroborated elsewhere so the model treats them cautiously.

Does domain authority decide selection?

It contributes but doesn't decide. Specific, extractable, corroborated content from a mid-authority site regularly beats vague content from a stronger one.

Can we see why we weren't selected?

Not directly. Perplexity shows its source list, which lets you see who won and inspect what they did differently — that's the closest available diagnostic.

These come up alongside Source Selection constantly.

Free 12-point check

Send Ron your site

Two fields. We’ll crawl your site as GPTBot, ClaudeBot and PerplexityBot, benchmark a sample of your category’s prompts, and send you what we find.

  • What each AI crawler actually receives from your site
  • Whether models resolve your brand as a real entity
  • A sample of category prompts and who gets cited
  • The three fixes we’d prioritise first

No commitment, no sales sequence. If we’re not a fit, we’ll say so.

We’ll pull your favicon so you know we found the right site.

We reply from a real address. No sequence, no newsletter.