▲ Now tracking visibility across Google, ChatGPT, Perplexity, Claude, Gemini & AI Overviews

How to get cited by ChatGPT (and why your SEO is half the battle)

A practical guide to entity optimisation, structured sources and the LLM retrieval patterns that decide who gets quoted.
· 13 min read

Last week a client asked the question we now hear most. “Why did Perplexity quote our competitor by name when we answer the same query better?” Fair question. The honest answer is that getting cited by an answer engine is not the same job as ranking on Google, and most teams are still trying to win the new game with the old playbook.

Here is what we have learned from the last 18 months of pulling apart citations across ChatGPT, Perplexity, Claude and Google AI Overviews. Some of it overlaps with classic SEO. A lot of it does not.

How an answer engine actually picks a source

Most LLM products do not browse the live web for every query. They use a retrieval layer (sometimes a search index they built, sometimes Bing, sometimes Brave, sometimes their own crawler) to fetch a small handful of candidate documents, then they ground the generated answer in those documents. The citation you see is one of the documents the retrieval layer surfaced.

That means two things matter independently. The retrieval layer has to pick your page. The generator has to find your page useful enough to quote.

Classic SEO mostly handles the first half. The second half is new work.

What the retrieval layer is looking for

The retrieval layer cares about the same things classic search has always cared about. Topical authority, freshness, structured data, credible inbound links, and a clean technical profile. If you have done good SEO over the last five years, you are in the candidate pool.

What it cares about that you might be missing:

  • Entity coverage. Does your page mention the entities (people, products, companies, standards) the query is about, with the right names and the right relationships? Wikipedia and Wikidata are still the implicit dictionary the retrievers use.
  • Crawl access. Have you accidentally blocked GPTBot, Google-Extended, ClaudeBot, or PerplexityBot? Check your robots.txt today. Half the sites we audit are blocking at least one.
  • Source diversity. Retrievers pull a mix. If your page is the third article on a topic written from the same angle as 200 others, it does not get picked.
  • Schema markup. FAQ, HowTo, Article, Product. Schema is no longer just a SERP feature. It is a parser shortcut for the LLM.

What the generator is looking for

Once your page is in the candidate pool, the generator decides whether to actually quote you. This is where most pages fail.

LLMs preferentially quote text that is short, factual, self contained, and clearly attributable. A page that buries the answer 600 words in, behind three rhetorical questions and a story about the founder, will not get quoted even if it ranks. A page that opens with a clean two sentence definition will.

If a smart intern can copy three sentences from your page and paste them into a client email as the answer, the LLM can quote them too.

Practical changes that have moved the needle for our clients:

  • Lead with the answer. The first 200 words of every page should contain a clean, quotable version of the core claim.
  • Use short, declarative sentences for the parts you want quoted. Reserve flourish for everywhere else.
  • Cite your own sources. LLMs treat pages that cite primary research as more trustworthy than pages that do not.
  • Use stable definitions. If you call a thing “answer engine optimisation” on one page and “generative engine optimisation” on another, you split your own entity.

Why your SEO is still half the battle

Some retrievers (Perplexity, Bing, Google AI Overviews) lean heavily on existing search indices. If you do not rank, you do not get retrieved. If you do not get retrieved, you do not get quoted. Classic SEO is the floor.

But it is not the ceiling. We have client pages that rank on page two of Google and get quoted in ChatGPT every week, because the page is structured for the generator even though it has not won the SERP yet. We have other pages that rank in the top three on Google and never get quoted, because the content was written for human persuasion, not for machine extraction.

The takeaway is unglamorous. Do the SEO. Then do the second pass for the generator. Two jobs. One page.

What we are doing for clients now

Across our book of work, the retrofit pattern is consistent. We pick the 20 pages that drive the most pipeline. We rewrite the leads to be quotable. We add or fix entity coverage. We make sure the bots are unblocked. We add schema where it is missing. We track citations weekly across six answer engines using a tool we built internally.

Most clients see citation share lift inside 90 days. Some sooner. The ones who do not are usually the ones with serious indexation problems we have to fix first, which is itself the answer to the original question. Your SEO foundation has to hold up. Then the LLM work compounds on top of it.

Download Premium WordPress Themes Free
Download Nulled WordPress Themes
Download Premium WordPress Themes Free
Premium WordPress Themes Download
download udemy paid course for free
download coolpad firmware
Download Premium WordPress Themes Free
udemy free download
Written by

Sean Clancy

Founder. Has spent the last 18 months pulling apart how LLMs decide what to quote.

Uncategorized

ChatGPT Ads in Australia: what the first three months actually tell us

Uncategorized

How much do ChatGPT Ads cost in Australia, and are they worth it?

Insights

What share of voice means in an answer engine world

Want this in your inbox?

One email per month. Real writing on what is changing in search, who it affects, and what we are doing about it for clients.

Talk to a senior strategist.

Take the playbook with you.

The local SEO checklist we run on every retainer. 52 checks, GBP walkthrough, NAP audit template. Free, no sales call.

No spam. Unsubscribe in one click. We’ll never share your email.