We pulled 5,676 real AI citations. Here's who AI actually cites

An honest look at our own scan data: across 2,209 answers, only 39% cited a source - and which engine you ask changes everything. Plus the recognisable domains that keep showing up, and how fragmented the rest is.

Most "what AI cites" posts quote someone else's study. This one is our own data, with the caveats in plain sight. We pulled every citation our scanners recorded over a recent nine-day window - June 16–24, 2026 - across the brands and questions our users actually track. That's 2,209 answers, 5,676 citations, 1,616 distinct domains. Here's what it actually looks like under the hood.

Read this as patterns, not a universal ranking. It's a recent snapshot, skewed toward the categories our users care about, and the long tail is full of niche, industry-specific sites we're deliberately not naming (some are customers' own domains). The shape of the data is the takeaway; the exact league table is ours, not the internet's.

Finding 1: most AI answers cite nothing

Only 39% of answers carried a single citation. The other ~60% named brands, made comparisons and gave recommendations with no visible source at all - answering from memory, not the live web. If your mental model is "AI shows its work," most of the time it doesn't. That alone reframes the channel: you can't earn your way into a majority of answers by being citeable, because those answers were never going to cite anyone. It's also why mentions and citations have to be tracked as two different things.

Finding 2: whether you get cited is the engine's decision, not yours

The single clearest pattern is how violently the citation rate splits by engine:

  • Perplexity - 100% of its answers cited at least one source (and ~9 citations per answer). It is a pure search-and-cite engine; it essentially never answers without showing receipts.
  • ChatGPT - ~90% cited sources, averaging ~6 per sourced answer. When it searches, it cites generously.
  • Claude - ~67% cited sources. More selective about when it reaches for the web.
  • Gemini and Grok - 0% in this dataset. Important honesty note: that's a configuration artifact - those engines ran without live web search in our default scans, so they answered purely from memory. It is not evidence that those engines "never" cite; it's evidence that an engine with web search off cites nothing, which is exactly the point.
Share of answers carrying at least one citation, by engine (2,209 answers, June 16–24 2026). Gemini and Grok ran with web search off, so they cite nothing - a configuration artifact, not a verdict on the engines.

So "can I get cited?" has no single answer. On Perplexity, almost always; on a memory-only engine, never. The same brand, the same question, a completely different game depending on who's answering.

Finding 3: the recognisable head is exactly what you'd guess

Filter to the cross-category domains anyone would recognise, and the leaderboard is unsurprising in a reassuring way - these are the surfaces worth being present on:

  • Tech-review publishers - TechRadar (the single most-cited domain at ~2.4%), Android Central, Tom's Guide, MakeUseOf, TechTarget.
  • Community & video - YouTube (~2.1%), Reddit (~1.3%), Instagram, Facebook.
  • Reference - Wikipedia (~2.0%).
  • Stores & aggregators - Google Play, the App Store, AlternativeTo.

Nothing exotic: established review sites, the big community platforms, the encyclopedia, the app stores. If you want to influence the citeable answers, those are the doors - and almost none of them are your own website.

Finding 4: the long tail is enormous

Here's the stat that surprised us most. 5,676 citations spread across 1,616 domains - and the most-cited domain in the entire set accounts for just 2.4% of citations. There is no monopoly. After the recognisable head, citations fan out into hundreds of niche, industry-specific and regional sites, each appearing a handful of times. For a specialised category, the page that wins the citation is often a small site that happens to answer the exact question well - which is genuinely good news if you're not a big brand.

The combined lesson: a majority of answers cite nothing, the ones that do are dominated by a predictable handful of authority sites, and beyond those it's a wide-open long tail. So the GEO playbook is two-track - be present on the authority surfaces engines trust, and publish the clean, specific, quotable page that wins the niche question outright.

What we'd do with this

Stop obsessing over getting your homepage cited; engines cite third parties far more than first parties. Instead: make sure your presence on review, reference and community sites is accurate and current, then own the long-tail questions in your category with pages built to be quoted. And track it per engine - because as finding 2 shows, an average hides the only number that matters.


Want this view for your own brand instead of our aggregate? A free Zene audit shows exactly which domains the engines cite when they answer questions about you - and which of them belong to your competitors.

Frequently asked questions

Do AI engines always cite their sources?

No - and it depends almost entirely on the engine. In our scan data, Perplexity cited a source on 100% of answers and ChatGPT on about 90%, while engines running without live web search cited nothing at all. Across everything, only 39% of answers carried any citation, so the majority of AI answers name brands with no visible source behind them.

Which websites do AI engines cite most?

In our sample the recognisable head was the usual mix: tech-review publishers (TechRadar, Android Central, Tom's Guide), community and video (Reddit, YouTube), reference (Wikipedia), and app stores and aggregators (Google Play, the App Store, AlternativeTo). But no single domain accounted for more than ~2.4% of citations - the long tail is enormous and industry-specific.

Can you optimise for AI citations?

Yes, indirectly. Because engines pull from high-authority reference, review and community sites far more than from any single brand's own pages, the lever is to be well-represented on those surfaces - accurate Wikipedia presence, real reviews, helpful community threads - and to publish clean, quotable pages of your own for the engines that search live.

Muhammet İLBAŞ
Written by
Muhammet İLBAŞ
Founder & engineer, building Zene in public
Share

Put it into practice

Free AI visibility checker →Schema markup generator →llms.txt generator →All free toolsCompare Zene vs alternatives

Keep reading

All articles →

Find out if AI recommends you.

Put this guide into practice - get your free visibility score in minutes.

Zene dashboard: a brand's AI visibility score, the next automatic scan, and per-engine cards for ChatGPT, Claude, Gemini and Perplexity