Ask five research firms how much of AI search runs on Reddit and you will get five answers between roughly 2 percent and 47 percent. All five are defensible. None of them are lying. They are simply counting different things, and almost nobody quoting them says which.
This matters more than a footnote, because marketing budgets get approved on the strength of a single number in a slide. If that number is 40 percent when the honest figure for your situation is closer to 2, the plan built on top of it will not survive contact with reality. So this article does something slightly unusual: instead of picking the most impressive statistic and running with it, we lay out what each study measured, why the results diverge by more than an order of magnitude, and which figure you should actually use depending on what you are trying to decide.
The short version is that Reddit is genuinely one of the two or three most influential sources in AI answers, that its influence is far less stable than most agencies admit, and that the gap between "retrieved" and "cited" is where most of the confusion lives.
The same reality, three different denominators
Every credible figure on Reddit's role in AI search is a fraction. The confusion comes from the bottom half of that fraction, not the top. There are three denominators in common circulation, and they produce wildly different headline numbers from the same underlying behaviour.
The first denominator is all citations across all answers. Measured this way, Reddit lands at around 3.11 percent of every citation issued, according to Profound's analysis of 4 billion AI citations and 300 million answer-engine responses published in November 2025. That sounds small. It was still enough to make Reddit the single most cited domain in that dataset, ahead of YouTube at 2.13 percent and Wikipedia at 1.35 percent. When the winner holds 3 percent, you are looking at an extremely fragmented field, which is the real insight buried in that number.
The second denominator is a platform's top cited sources. Narrow the field to the domains an engine leans on most and Reddit's share jumps. Arfadia's own AI Citation Rate Report 2026 found Reddit supplying 46.7 percent of top citations in Perplexity. Earlier Profound cuts covering August 2024 to June 2025 put Reddit at 6.6 percent of Perplexity's top sources and 2.2 percent of Google AI Overviews.
The third denominator is answers that contain at least one reference to the domain. Semrush analysed more than 150,000 AI citations across 5,000 keywords in June 2025 and found Reddit referenced in 40.1 percent of AI answers, ahead of Wikipedia at 26.3 percent and YouTube at 23.5 percent. These shares overlap, because one answer can cite several domains at once. This is the denominator that produces the biggest, most quotable, most misleading number.
Why Reddit is simultaneously a 3 percent source and a 40 percent source
Every citation issued across every tracked engine forms the denominator. Reddit still ranked first overall, which tells you the field is unusually fragmented.
Only the domains one engine leans on most form the denominator. Useful for channel selection, useless as a general claim about AI search.
Any answer citing Reddit at least once counts. Shares overlap, so several domains can each score high on the same set of answers.
Sources: Profound analysis of 4 billion citations and 300 million answer-engine responses, November 2025, produced in collaboration with Reddit • Arfadia AI Citation Rate Report 2026 • Semrush analysis of 150,000+ citations across 5,000 keywords, June 2025
Created by Arfadia • arfadia.com/blog
Read those three cards again and notice that none of them contradicts the others. A domain can hold a small slice of total citation volume, a large slice of one engine's preferred sources, and appear somewhere in a large share of answers, all at once. The numbers only look incompatible when the denominator is stripped off, which is exactly what happens when a statistic travels through three vendor blog posts before reaching a pitch deck.
Retrieved is not the same as cited, and the gap is enormous
Here is the finding that reframes everything above. Ahrefs analysed 1.4 million ChatGPT prompts and found that while the model reaches for Reddit constantly, it names Reddit as a distinct source only about 1.93 percent of the time. Their conclusion was blunt: ChatGPT uses Reddit extensively to understand a topic and gauge consensus, but it almost never gives Reddit the credit.
Dejan.ai went further in 2026 and quantified the funnel. In their measurement, ChatGPT's live search retrieved Reddit in 76 percent of searches but selected it only 0.61 percent of the time, a rejection rate of 99.39 percent. The same research found zero Reddit citations in Claude across 139,601 grounding sources sampled between May and July 2026.
This distinction has a direct practical consequence. If your success metric is a visible Reddit link inside a ChatGPT answer, you are chasing a 0.61 percent selection rate and you will report failure for months. If your success metric is whether the model's picture of your category includes your product as a credible option, retrieval matters and the citation is incidental. Those are two different programmes with two different reporting frameworks, and conflating them is how Reddit work gets cancelled at month four for hitting the wrong target.
The volatility nobody puts in the pitch deck
Reddit's position is not a fixed asset. It moves, sometimes violently, and often for commercial rather than technical reasons.
Semrush tracked citation behaviour over three months in 2025 and documented ChatGPT citing Reddit in roughly 60 percent of prompt responses in early August, collapsing to about 10 percent by mid September. Wikipedia fell on a similar curve over the same window, from around 55 percent to under 20 percent. The change correlated with Google removing the num=100 search parameter, a piece of infrastructure plumbing with no obvious connection to Reddit at all.
Then there is litigation. Reddit sued Perplexity in October 2025 over scraping, and third-party trackers subsequently reported Perplexity's Reddit citation share dropping sharply. Meanwhile SE Ranking's June 2026 update found Reddit's share moving the other way, roughly doubling from about 2.3 percent to 4.5 percent since November 2025, making it the second most cited domain in their data.
Underneath all of it sit contracts. Reuters reported in February 2024 that Reddit's data licensing deal with Google is worth about 60 million dollars per year. The Columbia Journalism Review put the OpenAI partnership struck in May 2024 at an estimated 70 million a year. When more than 130 million dollars of annual licensing shapes which domain a model reaches for, "Reddit is the most cited source" is partly a statement about procurement, not just about content quality.
How each engine actually behaves
Averaging across engines destroys the only information that would help you plan. The platforms diverge sharply, and the divergence is stable enough to act on.
| Engine | Reddit's role | Figure and source |
|---|---|---|
| Perplexity | Largest single source domain, ranked first | 46.7% of top citations, Arfadia AI Citation Rate Report 2026. Around 24% of Perplexity's January citations, Tinuiti Q1 2026 |
| ChatGPT | Heavily retrieved, rarely credited | Ranked second overall behind Wikipedia, Profound 2025. Cited as a distinct source ~1.93%, Ahrefs. Retrieved 76%, selected 0.61%, Dejan.ai 2026 |
| Google AI Overviews | Second ranked cited domain | Profound 2025. Cites Reddit in far fewer queries than ChatGPT in absolute volume, BrightEdge |
| Google AI Mode | Third ranked cited domain | Profound 2025. Wikipedia, Reddit and YouTube lead within AI Mode responses, SE Ranking |
| Grok | Second ranked cited domain | Profound 2025 |
| Gemini | Almost absent as a visible citation | Cited Reddit in roughly 0.1% of responses, Tinuiti Q1 2026 |
| Microsoft Copilot | Ranked 31st, effectively marginal | Profound 2025 |
| Claude | No Reddit citations found in the sample | Zero across 139,601 grounding sources, May to July 2026, Dejan.ai |
The spread here is the actionable part. Reddit presence is close to essential if your buyers use Perplexity, valuable but invisible if they use ChatGPT, and close to irrelevant as a citation play if they use Gemini or Copilot. For Indonesian businesses this skews the calculation in a specific direction, because Perplexity's share of AI assistant use in Indonesia runs at 21.18 percent against 7.88 percent globally, per Statcounter data cited in the Arfadia report. The engine that likes Reddit most is disproportionately popular in the market.
Why AI engines want community text in the first place
There is a tempting but wrong explanation for all this, which is that models cite Reddit because Reddit ranks well. BrightEdge data suggests something more interesting. Nearly 20 percent of ChatGPT's Reddit citations co-appear with institutional sources such as Healthline, Mayo Clinic, Cleveland Clinic, WebMD, Forbes or NerdWallet in the same answer.
That co-appearance pattern says the model is not treating Reddit as a substitute for authority. It is treating community discussion as a complement to it: the clinic explains the mechanism, the forum explains what it was actually like. BrightEdge also found ChatGPT surfacing Reddit in roughly 55 percent more queries than Google AI Overviews in absolute volume, the exact inverse of the YouTube pattern where Google cites the video platform in roughly 30 times more queries than ChatGPT does.
Which brings us to the ranking assumption. Nearly half of AI Overview citations, 47 percent, come from pages outside the organic top five, according to BrightEdge figures in the Arfadia report. A page you cannot rank can still be quoted, and a community thread can outrank your own site inside an AI answer. The strongest single predictor of citation that has been measured is not position but brand mention volume across the open web, at a correlation of r = 0.664 in Ahrefs data. Reddit is one of the larger contributors to that mention volume, which is a different and more durable reason to be present than chasing a link.
Four questions that expose whether a Reddit statistic means anything
What is the denominator?
All citations, one engine's top sources, or answers containing a mention. Without this, the number is decoration. A figure quoted without its denominator should be treated as unusable.
Which engine, and over what window?
Reddit's share ranges from first place on Perplexity to zero observed citations on Claude. A blended cross-engine average hides the only detail that would change your plan.
Retrieved or cited?
These differ by two orders of magnitude on ChatGPT. Pick which one your reporting measures before the engagement starts, not after the first quarterly review goes badly.
Studies commissioned by or produced with a platform are still useful, but the incentive belongs in the footnote. Profound's headline figure was produced in collaboration with Reddit, and we cite it while saying so.
Who funded it?
Sources: Profound, November 2025 • Ahrefs 1.4 million prompt analysis • Dejan.ai, 2026 • Semrush three month tracking study, 2025 • Arfadia AI Citation Rate Report 2026
Created by Arfadia • arfadia.com/blog
What this means for how you actually work
Three conclusions follow, and none of them is "post more on Reddit".
The first is that cadence beats volume, because of rotation. Between 40 and 60 percent of cited sources rotate out month to month, per AirOps figures in the Arfadia report. Presence has to be maintained rather than achieved once, which argues for a modest programme running for years over a burst that looks impressive for one quarter and then decays.
The second is that Reddit is a layer, not a channel. Your own pages give an engine something authoritative to cite. Community discussion gives it the peer consensus it looks for alongside that. Neither substitutes for the other, which is why we treat Reddit work as an execution layer inside a generative engine optimization programme rather than as a standalone subscription.
The third is the uncomfortable one. Reddit is actively removing content it judges to have been manufactured for AI visibility. The platform reported blocking around 23 million spam views per day and revoking close to 2 million inauthentic votes per day in July 2026, and said it now uses language models to catch coordinated patterns older systems missed. The version of this work with a future is the version that would survive a moderator reading it closely, which is a constraint on method rather than a reason to stay away.
That constraint is why our Reddit marketing automation service, built with technology partner AI Rush, automates discovery, scoring and drafting but stops short of publishing. Software finds the conversations and writes the options. A person decides what goes out. Volume figures in any proposal are a ceiling on what the system may surface, never a quota it has to fill, which means a quiet month correctly produces nothing at all.
None of this requires believing a single tidy percentage. It requires knowing which percentage you are looking at.
Frequently Asked Questions
Is Reddit really the most cited domain in AI search?
By several credible measurements, yes. Profound's November 2025 analysis of 4 billion citations found Reddit the most cited domain aggregated across tracked engines at 3.11 percent of all citations, and a Peec AI analysis of 30 million sources reported via Search Engine Land in March 2026 also put Reddit first across ChatGPT, Google AI Mode, Gemini, Perplexity and AI Overviews. Two caveats belong beside that. The Profound study was produced in collaboration with Reddit, and the ranking is not stable across engines, since Reddit ranks first on Perplexity, 31st on Microsoft Copilot, and had zero observed citations in a Claude sample.
Why do different studies report such different numbers for Reddit?
Because they use different denominators. Reddit is about 3.11 percent of all citations issued, roughly 46.7 percent of Perplexity's top cited sources, and appears in 40.1 percent of AI answers that contain at least one reference to it. These describe the same underlying behaviour and look wildly different. A statistic quoted without stating its denominator cannot be compared to any other statistic, which is why blended figures should be treated with suspicion.
What is the difference between a source being retrieved and being cited?
Retrieval means the engine read the page while forming an answer. Citation means it named the page in the output. On ChatGPT these differ by roughly two orders of magnitude. Ahrefs found Reddit cited as a distinct source about 1.93 percent of the time despite heavy use, and Dejan.ai measured ChatGPT retrieving Reddit in 76 percent of searches while selecting it only 0.61 percent of the time. If your reporting counts visible links only, you will understate the effect of community presence considerably.
How stable is Reddit's position in AI answers?
Less stable than most vendors admit. Semrush tracked ChatGPT's Reddit citations falling from roughly 60 percent of prompt responses in early August 2025 to about 10 percent by mid September 2025. Separately, AirOps figures indicate 40 to 60 percent of cited sources rotate out month to month. Commercial factors move the numbers too, since Reddit sued Perplexity over scraping in October 2025 and trackers reported Perplexity's Reddit citation share dropping sharply afterwards.
Does Reddit matter for Indonesian businesses specifically?
More than the global averages suggest, because of which engine Indonesians use. Perplexity holds 21.18 percent of AI assistant use in Indonesia against 7.88 percent globally, per Statcounter data cited in the Arfadia AI Citation Rate Report 2026, and Perplexity is the engine that leans on Reddit most heavily. The engine with the strongest Reddit appetite is disproportionately popular in this market, which shifts the calculation for Indonesian brands relative to a purely Western reading of the same data.
Do I need to rank in Google to be cited in AI answers?
No, and assuming otherwise leads to the wrong plan. BrightEdge figures cited in the Arfadia report show 47 percent of AI Overview citations coming from pages outside the organic top five. The strongest single measured predictor of citation is brand mention volume across the open web at a correlation of r = 0.664 in Ahrefs data, not ranking position. A community thread can outrank your own site inside an AI answer even when your site ranks better in classic search.
How much do data licensing deals affect which sources AI engines cite?
Enough that they belong in the analysis. Reuters reported in February 2024 that Reddit's licensing arrangement with Google is worth about 60 million dollars per year, and the Columbia Journalism Review put the OpenAI partnership from May 2024 at an estimated 70 million per year. That is over 130 million dollars annually shaping which domain models reach for. Treat "most cited source" as partly a statement about commercial agreements rather than purely a verdict on content quality.
What is the right way to measure a Reddit programme then?
Decide before you start whether you are measuring visible citations or category consensus, because the two produce very different numbers. Track cadence rather than volume, given that cited sources rotate 40 to 60 percent monthly. Report by engine instead of averaging across them, since Reddit's influence ranges from first place to effectively absent depending on the assistant. And treat an honest zero in a quiet period as a valid result rather than a reporting failure.
Sources & References:
- Profound, "The Data on Reddit and AI Search", 10 November 2025, analysing 4 billion AI citations and 300 million answer-engine responses from August 2024 to late October 2025. Reddit most cited domain overall at 3.11 percent of all citations, YouTube 2.13 percent, Wikipedia 1.35 percent. By platform: first on Perplexity, second on ChatGPT behind Wikipedia, second on Google AI Overviews, second on Grok, third on Google AI Mode, 31st on Microsoft Copilot. This study was produced in collaboration with Reddit and is therefore partially vendor-aligned, noted here rather than omitted.
- Peec AI analysis of 30 million sources, reported by Search Engine Land, 31 March 2026, finding Reddit the most cited domain across ChatGPT, Google AI Mode, Gemini, Perplexity and AI Overviews, followed by YouTube and LinkedIn.
- Semrush, June 2025, analysis of more than 150,000 AI citations across 5,000 keywords: Reddit referenced in 40.1 percent of AI answers, Wikipedia 26.3 percent, YouTube 23.5 percent. Shares overlap because a single answer may cite several domains.
- Semrush, "The Most-Cited Domains in AI: A 3-Month Study", tracking August to November 2025: ChatGPT cited Reddit in approximately 60 percent of prompt responses in early August 2025, falling to approximately 10 percent by mid September 2025, correlating with Google's removal of the num=100 search parameter. Wikipedia declined on a comparable curve.
- Ahrefs analysis of 1.4 million ChatGPT prompts, finding heavy Reddit retrieval alongside distinct-source citation of approximately 1.93 percent. Ahrefs Brand Radar separately ranked Reddit the third most cited domain in LLMs.
- Dejan.ai, 2026: ChatGPT live search retrieved Reddit in 76 percent of searches and selected it in 0.61 percent, a 99.39 percent rejection rate; zero Reddit citations observed in Claude across 139,601 grounding sources sampled May to July 2026.
- Arfadia AI Citation Rate Report 2026, DOI 10.5281/zenodo.21100366, indexed on Google Scholar: Reddit supplying 46.7 percent of top citations in Perplexity; Perplexity share of AI assistant use in Indonesia 21.18 percent against 7.88 percent globally per Statcounter; brand mention to citation correlation r = 0.664 per Ahrefs; 47 percent of AI Overview citations from pages outside the organic top five per BrightEdge; 40 to 60 percent monthly rotation of cited sources per AirOps.
- Tinuiti Q1 2026 AI Citations Trends Report, produced with Profound, released March 2026, covering seven platforms across nine commercial categories: social media share of AI citations rising to approximately 9 percent by January 2026 with Reddit the dominant social source; Perplexity drawing approximately 31 percent of January citations from social media with Reddit around 24 percent of the Perplexity total; Gemini citing Reddit in approximately 0.1 percent of responses.
- BrightEdge, "How Google AI Overviews and ChatGPT Use Reddit Differently": ChatGPT surfacing Reddit in roughly 55 percent more queries than Google AI Overviews in absolute volume, the inverse of the YouTube pattern where Google cites YouTube in roughly 30 times more queries than ChatGPT; nearly 20 percent of ChatGPT Reddit citations co-appearing with Healthline, Mayo Clinic, Cleveland Clinic, WebMD, Forbes or NerdWallet.
- SE Ranking, June 2026 update: Reddit share roughly doubling from about 2.3 percent to 4.5 percent since November 2025, becoming the second most cited domain in that dataset. SE Ranking 2025 also identified Wikipedia, Reddit and YouTube as the top cited domains within Google AI Mode responses.
- Data licensing context: Reuters, 22 February 2024, reporting the Reddit and Google contract at approximately 60 million dollars per year. Columbia Journalism Review, estimating the Reddit and OpenAI partnership announced May 2024 at around 70 million dollars per year. Reddit filed suit against Perplexity over scraping in October 2025, after which third-party trackers reported a sharp decline in Perplexity's Reddit citation share.
- Reddit enforcement figures: Reddit official statement, "How We're Keeping Reddit Real and Safe in the AI Era", redditinc.com, 6 July 2026, reporting approximately 23 million spam views blocked per day, approximately 25,000 net new spammy posts and comments caught per day, and close to 2 million inauthentic votes revoked per day over the preceding three months.
- Reddit is a trademark of Reddit, Inc. PT Arfadia Digital Indonesia is not affiliated with, endorsed by, or sponsored by Reddit, Inc. Figures in this article are reproduced with their original denominators and study windows stated; readers comparing them should confirm that any two figures share a denominator before treating them as comparable.