← All guides

Guides

Reddit vanished from ChatGPT's citations. Our per-engine data shows a messier picture.

By Anil Sayar, founder of Promvia

Published · Updated

In short

In mid-August 2026, independent trackers measured Reddit's share of ChatGPT Search citations collapsing from roughly 4% of answers to under 1% within days. Reddit's robots.txt blocks every crawler by default — access runs through paid agreements, and when a pipeline is cut off, citations built on it can vanish overnight. Our own per-engine checks show the drop is uneven: Google's engines kept citing Reddit steadily, Perplexity went from citing it to not, and answers vary engine by engine. If your AI-visibility plan leans on one platform, this is what that risk looks like when it lands.

What actually happened

Reddit's robots.txt is two lines: User-agent: * / Disallow: /. Every crawler is refused by default, and has been since 2024 — real access runs through commercial agreements, most prominently the licensing deal with Google. So Reddit's presence in any AI engine's answers is not a fact of the open web. It is a contract, and contracts change.

What our own per-engine checks show

We track a handful of prompts about our own category across six engines and store every source each answer cites. Four prompts in one niche is a small corpus — directional evidence, not a study — but it is first-party, per-engine, and dated, which the headline numbers are not. Between August 19 and 24, three different things happened to Reddit in our data at the same time.

  • Google's engines kept citing Reddit as if nothing happened: AI Overviews and AI Mode returned Reddit threads in roughly a third to a half of our checks on every full sweep, before and after the reported drop. Google's licensing deal is separate, and it shows.
  • Perplexity flipped: nine Reddit citations across our August 19 checks, zero in every check after. In a corpus this small that could be answer variance — but the direction matches the reported cut, and it is the kind of step-change per-engine tracking exists to catch.
  • ChatGPT still cited Reddit for several of our tracked prompts after August 14. A collapse in aggregate share across millions of answers does not mean zero in any given niche — which cuts both ways: your niche may be hit harder or barely at all. The aggregate number cannot tell you; only checking your own prompts can.

Update, August 26: a week on, the split has hardened

We kept the per-day counter running. A week after the first cliff, the pattern is no longer ambiguous in our data: Perplexity has not cited Reddit once since August 19 — zero across 64 consecutive checks — while ChatGPT kept citing it on most days and both Google surfaces never moved. Same Reddit policy, three different outcomes.

Scrapers bypass technological protections to steal data, then sell it to clients hungry for training material.
Source: Ben Lee, Reddit’s chief legal officer, quoted by the Associated Press (PBS News, 22 October 2025)

There is a structural lesson in which engine fell hardest, though not the familiar one. The common explanation is that Perplexity fetches sources live at answer time, so a block hits it first. Our own server logs do not support that for its API answers: none of 104 Perplexity answers that cited our site had a verified fetch of the cited URL in the 10 minutes before, while PerplexityBot had crawled the cited URL beforehand in 174 of 174 citations (the full join is here). On that evidence, a crawler block reaches Perplexity's answers through its index rather than through a live fetch, and when it shows depends on when that index next drops the page.

Update, September 9: the zero did not hold — and our own instrument moved

Two corrections, both against us. First, the Perplexity zero was a gap, not a regime change. Reddit came back into Perplexity's citations from August 26 — the day the update above was published — at a lower rate than before: 5/28 on August 27, 7/59 on September 3, 7/97 on September 7. Six days of zero followed by a return is not the pattern the August 26 section described, and we should have kept counting before writing 'hardened'.

Second, ChatGPT in our data went from 63 Reddit-citing checks out of 184 (August 19–30) to 0 out of 364 since August 31 — the cleanest before/after in the whole corpus, and we cannot use it. On the evening of August 30 we switched that lane's model from gpt-4o to gpt-4o-mini, effective from the 31st. The zero begins on the day our instrument changed, so it says nothing about Reddit or about OpenAI's retrieval; it may say something about which model cites forums. Google's surfaces kept citing Reddit throughout (September 7: AI Overviews 27/78, AI Mode 12/78). The AI Source Index now lists our own instrument changes next to the changelog so this cannot be read as a market move again.

What this means if Reddit was part of your plan

For two years the standard off-site advice — ours included, in the off-site authority guide — has been that engines cited community sources heavily, so genuine, useful Reddit participation was worth real effort. That advice is not wrong, but August 2026 exposed its dependency: the value of a Reddit thread in ChatGPT's answers depends on a commercial agreement you are not a party to and cannot see change.

  • Don't abandon Reddit — Google's engines still cite it, humans still read it, and the last collapse reversed. A thread that helps people keeps its Google-side and human value regardless of what OpenAI and Reddit negotiate.
  • Do stop treating any single platform as your AI-visibility channel. The same lesson applies to YouTube, to review directories, to news sites — any source an engine reads under an agreement can go dark for that engine.
  • Measure per engine, not in aggregate. One blended visibility score would have averaged this event into noise. Engine-level rows show you exactly who dropped a source and when.
  • Watch where the citations went. In our checks the space Reddit left was filled by vendor comparison pages, review directories and YouTube — surfaces you can actually publish to or maintain a profile on.

The general version of this lesson is in how AI engines choose sources: engines cite what they can read, and what they can read is increasingly a matter of deals rather than crawling. The practical response is a portfolio — your own quotable pages plus presence on several platforms — so no single contract change can zero you out.

Frequently asked questions

Did Reddit block AI crawlers completely?

Reddit's robots.txt has disallowed all crawlers by default since 2024 — access runs through commercial agreements. What changed in August 2026, per multiple independent trackers, is that Reddit content largely stopped appearing in ChatGPT Search citations, consistent with a cut pipeline. Google's engines, which license Reddit data separately, kept citing it.

Should I stop posting on Reddit for AI visibility?

No. Google's AI engines still cite Reddit, human readers still see it, and the previous citation collapse later reversed. The change is in how much weight to put on it: treat Reddit as one platform in a portfolio, not as the strategy.

How do I know if my niche was affected?

Check the prompts you care about against real engines and look at the cited sources per engine. Aggregate statistics can't tell you — our own small corpus had ChatGPT still citing Reddit after the reported drop, while Perplexity went to zero for six days and then came back at a lower rate.

See where you stand today

Run your questions across up to 7 AI visibility surfaces and get your baseline — small runs often finish in about a minute. Free plan, no credit card.

Start free →Try the free AI Traffic Checker

Keep reading

Off-site authority: the GEO lever most guides skip →How AI assistants choose which sources to cite →AI share of voice: the metric that shows who owns your category's answers →