Original research · Pre-registered

The Perplexity Reddit Dependency

Perplexity's citations reportedly skew toward Reddit at nearly 47%. For topics with thin community discussion — much of B2B and enterprise content — what fills that gap? A pre-registered study.

Ritik Namdev Ritik Namdev ·Published September 2026 ·v0 — design stage ·10 min read
The short version

Reported data puts Reddit at nearly 47% of Perplexity's top citations. But that describes topics with substantial community discussion. For B2B, enterprise, and specialized topics where Reddit discussion is thin, what wins the citation instead? This page registers a study to find out. It mirrors the design used for ChatGPT's Wikipedia dependency.

The same question, a different engine

Just as ChatGPT's Wikipedia dominance describes a subset of well-covered topics, Perplexity's reported Reddit skew describes topics with active community discussion. A large share of B2B, enterprise, and specialized professional content simply doesn't generate the Reddit discussion volume consumer topics do. That makes the "what fills the gap" question at least as practically relevant here as for ChatGPT.

What is known

Evidence

Perplexity's citations reportedly skew toward Reddit at roughly 46.7% of top citations — Partial , vendor-reported from a non-public corpus.

Open question

What wins the citation for topics with thin Reddit discussion — no published study addresses this.

Why Reddit specifically became the default

It's worth naming what makes Reddit distinctive as a citation source, since the gap-fill prediction below depends on it. Reddit threads offer first-person, lived-experience answers. That's exactly the kind of answer a citation-forward engine needs for comparative questions, like "which is actually better" or "is this worth it."

That property, genuine, unpolished, first-hand testimony at scale, is largely absent for niche B2B and enterprise topics. Public discussion there is thinner. Much of what does exist reads as vendor marketing, not independent testimony.

Defining "thin discussion" precisely

To keep the eventual finding a genuine test, not a selectively illustrated one, "thin Reddit discussion" gets defined by a concrete threshold, set before data collection starts. A topic qualifies if a direct search for it across relevant subreddits returns no thread with a meaningful number of substantive comments. It won't be judged qualitatively after the fact, by whichever topics happen to produce interesting results.

The research question

Restrict to queries about topics with minimal Reddit discussion volume. Which domain types capture the citation instead? Official documentation, industry publications, first-party vendor content, or something else?

Study design

Collection pipeline
  1. 01 Identify topics Where Reddit has thin or no discussion
  2. 02 Query Perplexity Across the fixed query set for those topics
  3. 03 Log citations What replaces the community-discussion default
  4. 04 Classify sources By domain type and content structure
  5. 05 Compare Against topics where Reddit discussion is dense

A worked hypothetical example

Here's the design made concrete, with an invented case. Consider a query about a specialized category of industrial B2B software, with essentially no Reddit discussion beyond a single old thread and a handful of comments. That topic would qualify as thin-discussion under the defined threshold.

The study's prediction, RD1, is that the citation for such a query gets captured instead. Maybe by the vendor's own comprehensive documentation. Maybe by an industry analyst report, or a trade publication's coverage. No Reddit source exists at sufficient depth to compete. Whether this pattern holds consistently across many such B2B gap-topics, not just this one invented example, is exactly what the study is designed to determine.

Pre-registered hypotheses

#HypothesisPrediction
RD1In the absence of Reddit discussion, first-party vendor and documentation content captures a disproportionate share of citations on PerplexitySupported
RD2B2B/enterprise queries show meaningfully lower Reddit citation share than consumer queries in the baseline (Reddit-present) comparison groupSupported

Perplexity's ~47% Reddit citation share describes topics with active community discussion. Most B2B and enterprise content doesn't have that — and nobody has published what wins the citation there instead.

Share on X

What this means for B2B and niche brands

Suppose this gets confirmed. B2B and enterprise brands can't realistically generate the volume of organic community discussion consumer brands attract. But this would give them a more tractable target for Perplexity visibility, instead of trying to compete with Reddit's dominance directly.

How this compares to the Wikipedia dependency study

This study and the Wikipedia dependency study share an identical structure. Identify a gap where an engine's dominant citation source has nothing, then measure what fills it. Each applies that structure to a different engine and source pair.

Running both under the same design also creates a natural comparison, once both have real data. Does the same kind of source, official documentation, established publications, fill the gap for both engines? Or do ChatGPT and Perplexity diverge even in how they handle an absent default? That would be a genuinely interesting finding neither individual study is built to surface alone.

What to do now, before results exist

If your topic sits in a thin-Reddit category, the safest bet right now is simple. Become the documentation Perplexity would otherwise have nothing better to cite. Publish the specific, comprehensive reference page a Reddit thread would normally provide: honest tradeoffs, real numbers, edge cases.

A page that reads like vendor marketing is unlikely to win this gap even once data confirms the pattern. A page that reads like a knowledgeable practitioner's own writeup has a much better shot, since that's the closest match to what Reddit itself would have supplied.

Limitations

  • "Thin Reddit discussion" requires a defined threshold, to be specified before data collection.
  • Results are specific to Perplexity and shouldn't be assumed to generalize to other engines' behavior in analogous gaps.
How to cite this
Namdev, R. (2026). The Perplexity Reddit Dependency (v1). Retrieved from https://ritiknamdev.com/blog/reddit-dependency-perplexity

Published under CC BY 4.0 — reuse freely with attribution.

Related work on this site

The Perplexity counterpart to the Wikipedia dependency study. See Perplexity citation statistics for the underlying skew figure.

§ References

Sources

Figures attributed to third parties above have not been independently verified unless stated otherwise.

FAQ

Frequently asked questions

Why does this matter more for B2B and enterprise topics?
Because those topics are exactly where Reddit discussion tends to be thinnest. A niche enterprise software category, or a specialized B2B service, is far less likely to have substantial Reddit threads than a consumer product. That makes the "what fills the gap" question especially relevant there.
Is this just the Wikipedia-dependency study with a different engine?
Same structural question, but a genuinely different engine and source type. Perplexity's citation-forward design and Reddit-heavy skew make its behavior, in the absence of that default, a distinct empirical question. It isn't assumed to mirror ChatGPT's.
Why would first-party vendor content specifically win the gap, rather than, say, review sites?
That's exactly the open part the study is designed to test, not assume. RD1's prediction, first-party and documentation content, is a reasoned bet based on what's typically available for thin-discussion B2B topics. It's not a foregone conclusion. Review-site or analyst-report dominance are both plausible alternative outcomes the data could support instead.
Could a brand artificially seed Reddit discussion to close this gap instead?
That approach carries real risk. Reddit's community norms and moderation are generally hostile to detectable brand-seeded discussion. And if Perplexity's citation preference is specifically for organic, community-driven threads, manufactured discussion may not reproduce the same citation behavior even if it isn't removed. This study doesn't test that directly, but it's a relevant caution for anyone considering the idea.
Ritik Namdev
Written by

Ritik Namdev

Growth · SEO · GEO

Growth marketer documenting a brand-new site's climb into Google and the AI engines - in public, with real numbers. Every tactic here is tested on real sites before it's published.

The Lab · Weekly

One experiment. Every week.

The field notes in your inbox - one thing I tested, the raw numbers behind it, and what it means for getting cited by AI.

Free forever. Unsubscribe anytime.