# Does AI prefer Reddit? Why the answer depends on which AI you ask

> OpenAI rejects 99.4% of the Reddit pages it retrieves, Anthropic cites zero, and only Google keeps Reddit near its baseline citation rate.

Canonical: https://brandonlazovic.dev/articles/does-ai-prefer-reddit/  
Author: Brandon Lazovic  
Published: 2026-07-27

## The short version

- Reddit's presence in AI answers tracks each engine's own citation behavior rather than a shared AI preference: OpenAI rejects 99.39% of the Reddit pages it retrieves, Anthropic cites zero, and Google keeps Reddit near its own baseline rate.
- OpenAI's citation-mining data makes Reddit the single most-rejected domain it measures: 3,012 citations out of 491,024 candidate offers across six months, per Dejan.
- Anthropic's Claude never cited Reddit once across 139,601 grounding sources sampled from May through July 2026, per the same analysis.
- Google cites Reddit in about 2% of its answer sources, second only to YouTube. Dejan's extrapolation puts Google's Reddit selection rate at roughly 9% to 60%, about 14 times higher than OpenAI's measured 0.61% even at the low end.

AI answer engines do not share a preference for Reddit. In citation-mining data from Dejan covering the last six months, OpenAI's models reject 99.39% of the Reddit pages they retrieve, making Reddit the single most rejected domain in Dejan's entire dataset. [1] Anthropic's Claude has cited Reddit zero times across 139,601 grounding sources sampled from May through July 2026. [1] Only Google keeps Reddit at anything close to a normal rate, citing it in roughly 2% of its answer sources, second only to YouTube. [1] Because the popular GEO advice to win on Reddit describes only one engine's behavior, treating it as a general law wastes effort on ChatGPT and Claude and misses what actually drives the result: Google's own search ranking of Reddit.

> OpenAI rejects 99.39% of the Reddit pages it retrieves. Google keeps Reddit at close to its own baseline citation rate for every domain.

## Does ChatGPT (OpenAI) actually prefer Reddit?

No. Reddit is the single most-rejected major domain in OpenAI's citation data. Across Dejan's six-month sample, Reddit was logged as a candidate source 491,024 times, and OpenAI's models cited only 3,012 of those pages, a 99.39% rejection rate, worse than Wikipedia's 94.36% or arXiv's 99.23%. Reddit shows up in ChatGPT conversations mainly because it is offered as a candidate constantly, in 76% of sampled searches. The model then throws almost all of it away.

In Dejan's methodology, three ideas that casual GEO talk collapses into one stay separate: a grounding source (a page the model retrieves and could cite), a citation (a page the model actually uses in its answer), and a mention (the word "Reddit" appearing somewhere in a prompt or response). The 491,024 figure is grounding-source offers, not citations. In that same sample Reddit was a candidate in 76% of OpenAI's sampled probes (59,570 of 78,331), and each of those probes surfaced roughly eight separate Reddit pages, which is how one domain generates nearly half a million candidate offers in six months. [1]

The rejection rate is not close to typical. Dejan's own comparison table sets Reddit against two other high-volume domains from the same window:

| Domain | Retrieved | Cited | Rejected | Rejection rate |
|---|---|---|---|---|
| Reddit | 491,024 | 3,012 | 488,012 | 99.39% |
| Wikipedia | 229,879 | 12,968 | 216,911 | 94.36% |
| arXiv | 46,700 | 359 | 46,341 | 99.23% |

Dejan states plainly that Reddit is "the most rejected website" in its entire dataset. [1] Across every domain OpenAI considers, its average selection rate runs at 12.85%. Reddit's sits at 0.61%, about one-twenty-first as often as the average domain gets selected. [1]

## But isn't Reddit already one of the most-cited domains in ChatGPT?

By share of citations, yes, and that is compatible with a 99% rejection rate rather than a contradiction of it. In mid-2025 the AI-visibility vendor Profound reported Reddit as the second most-cited domain in ChatGPT behind Wikipedia, with its citation volume up roughly 400% and Reddit accounting for about 6% of all the sources ChatGPT cited. [3]

That share and Dejan's 99.39% rejection rate answer two different questions about the same funnel. Profound's figure is a share of final citations: out of everything ChatGPT cited, what fraction was Reddit. For Dejan the measure is a selection rate: out of the Reddit pages ChatGPT retrieved as candidates, what fraction it actually used. [1] A domain offered as a candidate as relentlessly as Reddit, present in 76% of OpenAI's sampled searches at roughly eight pages each, can convert well under 1% of those candidates and still finish among the largest shares of the citation pie, because the pool feeding it is so large. Because offered-volume runs high while conversion stays low, that profile produces a big aggregate share and a brutal per-candidate rejection rate at the same time. Since both platforms have kept changing how they weight Reddit in the months since, that share is a moving mid-2025 snapshot rather than a fixed fact.

## What about Anthropic's Claude and Google's AI answers?

Anthropic's Claude has zero Reddit citations in Dejan's dataset: across 139,601 grounding sources sampled from May through July 2026, Reddit never appears, meaning it is rarely even offered as a candidate. Google shows the opposite pattern. In that same window, Reddit accounts for roughly 2% of Google's 697,768 cited sources, ranking second only to YouTube. [1]

The Anthropic and OpenAI numbers look similar at a glance, both near-zero Reddit presence, but they describe different mechanisms. OpenAI's models see Reddit constantly and reject almost all of it, an active filter. Because Reddit is absent from Claude's candidate pool rather than filtered out of it, Claude mostly never sees Reddit as an option to begin with. [1] Dejan's data cannot explain why Claude's retrieval pipeline surfaces Reddit so rarely, only that it does.

Google's roughly 2% citation share understates how differently Google treats Reddit as a candidate. In a separate probe of 3,450 sampled searches, Reddit.com appeared as a result 933 times, across 916 searches, a 26.6% search-level presence rate, against Reddit's presence in 76% of OpenAI's sampled probes. [1] Google surfaces Reddit as a candidate far less often than OpenAI does. When Google does surface it, Google keeps far more of what it sees.

## Why does Google keep Reddit when OpenAI throws it away?

Google keeps Reddit near its own baseline selection rate for two structural reasons: Reddit already ranks well in Google's organic search results, and Google holds a direct commercial pipeline to Reddit's content. Since Google exposes only the sources it cites and never the discarded pool, Dejan cannot measure its rejection rate directly, but extrapolating from candidate-pool sampling puts Reddit's Google selection rate at roughly 9% to 60%, about 14 times more favorable than OpenAI's 0.61%. [1]

The range exists because Google's citation reporting is one-for-one: it shows what got cited, never what got discarded. Dejan's workaround uses OpenAI as a control, since OpenAI's platform exposes both sides of the ratio, candidate pool and citations together, then applies the same candidate-to-citation logic to a separate 3,450-search sample, where Reddit appeared as a raw result 3% of the time against 2% of Google's actual cited sources. [1] At any given level of Google filtering, a variable Google does not publish, that gap between 3% and 2% implies a Reddit selection rate between roughly 9% (if Google filters as aggressively as OpenAI's 12.85% overall rate) and 60% (if Google keeps 80% to 90% of everything it grounds). On that point Dejan is explicit that this is a proxy: the candidate-side number comes from sampled DataForSEO searches, not from Google's own grounding pool, which Google does not publish. [1]

Dejan's own summary states the contrast plainly: "OpenAI is handed Reddit constantly and rejects almost all of it. Google is handed Reddit less often and keeps it at close to its normal rate." [1] "Handed less often" describes the candidate pool: Google's sampled searches carried Reddit 26.6% of the time, against 76% for OpenAI's probes. In Dejan's phrasing, "keeps it at close to its normal rate" describes the selection ratio once Reddit does appear as a candidate, a different measurement than the raw citation share.

Google's retention advantage plausibly runs deeper than organic ranking alone: Reuters reported in February 2024, citing people familiar with the matter, that Google and Reddit had signed a content-licensing agreement worth about $60 million a year. [2] That arrangement would give Google's systems direct commercial access to Reddit's content beyond what a public crawl provides. Since the figure is sources-based rather than on the record, with Reddit's own S-1 disclosing only an aggregate $203 million across the data-licensing deals it signed in January 2024, without naming Google or breaking out a per-partner number, treat the licensing mechanism as a well-reported but unconfirmed contributing explanation.

## So is "optimize for Reddit" bad advice?

"Optimize for Reddit" is engine-specific advice dressed up as a general rule. It is a reasonable bet if your buyers or research process routes through Google's AI Overviews or AI Mode, where Reddit holds a real, measurable share of citations. It is close to wasted effort if the target is ChatGPT, where Dejan's data shows Reddit gets rejected 99.4% of the time, or Claude, which does not retrieve Reddit as a source at all. [1]

The tactic became popular because Reddit's AI presence is visible and easy to point at: threads showing up in Google's AI Overviews, screenshots circulating on social media, a widely repeated claim that Google has favored Reddit since 2024. All of that is consistent with Dejan's Google numbers. For OpenAI or Anthropic none of it transfers, since the underlying data there runs in the opposite direction. A tactic validated on one engine's outputs and marketed as an AI-wide law is exactly the failure mode Dejan's cross-engine comparison exposes.

My read on Dejan's numbers, and the reusable idea underneath the whole piece, is what I would call citation-as-ranking-proxy: a domain's AI-citation presence in a given engine is largely a proxy for how that engine's own search already ranks it, not evidence of an AI-specific preference. Inside Google's AI answers, Reddit's success traces back to its organic ranking, the same backlinks, engagement signals, freshness, and topical relevance that earn any domain a place in Google's results, plus now a commercial licensing relationship. The AI layer sits downstream of that ranking. Optimizing for "AI citation" while ignoring organic ranking treats a symptom as the cause.

## What should you actually do about UGC and AI citations?

Treat AI-citation strategy as an extension of search strategy: audit which AI engine your audience actually uses, then check whether user-generated content already ranks for your terms in that engine's underlying search index. Google Search rewards structural signals like domain authority and freshness regardless of platform, and Reddit's AI presence rides on its organic ranking in Google, a proxy relationship rather than an AI-specific preference. [1]

Three practical moves follow from the data rather than from the general instinct to "get on Reddit." First, check which engines actually drive your traffic and citations before allocating effort. If ChatGPT and Claude dominate your referral logs, Dejan's numbers say Reddit-focused UGC work will mostly get filtered out or never retrieved at all. Second, if Google's AI Overviews or AI Mode matter to you, the fix is still classic SEO: whatever makes a page rank organically on Google is the same thing that gets it into Google's AI answers, UGC or not. Third, treat any single vendor's citation-mining numbers, including the ones in this piece, as directionally useful and methodologically bounded. In Dejan's case the sample is drawn from its own clients' prompts rather than a random cross-section of all AI queries, and Google's rejection rate here is inferred, not measured. [1]

The honest summary: Reddit's AI presence is a downstream artifact of its Google organic ranking, visible inside Google's AI answers because those answers are built on Google's organic index. Chase the organic ranking. The AI citations that are reachable will follow, on whichever engines make them reachable.

## Sources

1. Dejan: No, AI doesn't prefer Reddit. Search does. — https://dejan.ai/blog/reddit-ai/
2. Reuters: Reddit AI content licensing deal with Google, sources say (Feb 22, 2024) — https://www.reuters.com/technology/reddit-ai-content-licensing-deal-with-google-sources-say-2024-02-22/
3. Josh Blyskal, Profound: Reddit citations up 400% in ChatGPT (June 2025) — https://www.linkedin.com/posts/joshua-blyskal_reddit-citations-are-up-400-in-chatgpt-activity-7336057282968920068-m8cA
