// AI NEWS · PLATFORM

Reddit's AI Search Picks the Already-Popular Comment, a Large-Scale Audit Finds

A University of Illinois preprint that traced 30,000 Reddit-generated answers back through 14.68 million comments found that the feature overwhelmingly surfaces top-level, highly upvoted, formally worded replies — and strips the first-person voice out of the ones it does quote.

2026-09-18 · by Moe Ameen

What happened

Researchers at the University of Illinois Urbana-Champaign — Agam Goyal, Claire Wang, and Eshwar Chandrasekharan — published a large-scale audit of Reddit's generative search feature (the answers surfaced by what began as "Reddit Answers," launched December 2024 and merged into Reddit search on May 26, 2026). The preprint, titled "The Wisdom of the Loudest: A Large-Scale Audit of Generative Search on Reddit," went up on arXiv on September 13, 2026. It has not yet been peer-reviewed, so treat the figures as a first pass rather than settled science.

To run it, the team put 10,000 questions drawn from 20 advice- and support-seeking subreddits through the feature three times, producing 30,000 answers, and traced which comments the system selected across 14.68 million candidate comments. The central finding is that generative search on Reddit is not neutral summarization: it strongly favors comments that were already prominent. A comment's vote ranking inside its thread was the single strongest predictor of being chosen, and the median selected comment sat at the 91st percentile for score in its thread versus the 45th percentile for comments that were passed over. Ninety-two percent of selected comments were direct top-level replies to the original post, and selected comments had appeared a median of 1.2 hours after the post, against 5.9 hours for the ones left out — early and visible wins.

Style mattered too. Formal, directive language survived selection at higher rates: a comment one standard deviation more formal had roughly 49% higher odds of being picked (odds ratio about 1.49), and directive phrasing using words like "should" and "must" also helped. Comments rich in personal-experience markers or supportive, empathetic language were less likely to be selected. And when a first-person comment did make it in, the synthesis stripped its voice — first-person singular words like "I" and "my" fell from about 3.3% of the quoted source comments to roughly 0.06% of the final written answers. The lived-experience texture that makes Reddit useful largely does not survive the summarization step.

Why it matters for creators

  • Reddit is one of the most-cited sources across AI search, so how ITS own AI picks comments is a preview of how the popularity-follows-popularity loop works everywhere: the visible get quoted, and being quoted makes them more visible.
  • Being early is a ranking factor. Selected comments landed a median of 1.2 hours after the post — for anyone participating in communities, showing up fast and near the top of a thread beats showing up thorough and late.
  • Format beats feeling. Clear, structured, directive answers ("here is what to do") survive selection; long personal anecdotes and supportive replies get passed over or have their first-person voice removed.
  • It is a caution, not a playbook to game. You cannot buy upvotes credibly, and Reddit users punish inauthentic marketing — the durable move is being genuinely early, clear, and useful, not manufacturing prominence.
  • A citation is still not a visit. As Reddit's own leadership has warned about AI Overviews, being the comment an AI quotes builds recognition but does not reliably send a click — so the value is in being the recognizable source, everywhere an AI looks.

How to act on this with Kompozy

The honest read of this audit is a discoverability lesson, not a hack: the content an AI surfaces from a community is the content that was already prominent, early, clearly structured, and directive. You cannot fake upvotes on Reddit without getting caught, and you shouldn't try. What you CAN do is show up consistently and fast, with clear, on-brand, useful posts, across Reddit and every other place an AI reads — and that consistency-at-speed is exactly the gap a content engine closes. Reddit is a native Direct Connect destination in [Kompozy](/) via Bundle.social, alongside the eight primary social platforms plus blog and email, so the same brand voice ships everywhere in one pass.

Concretely: hand Kompozy one source — a post, a transcript, a note — and it generates roughly 25–35 finished assets across 18 formats, each governed by a [Persona Brief](/glossary/persona-brief) that keeps the voice tight, clear, and directive rather than rambling. [Autopilot](/glossary/autopilot) then schedules and publishes them so you are early and present in the threads and feeds that matter instead of late and sporadic. Pair the community presence with the surfaces AI actually cites most — a full blog article and structured social posts — and you build the recognizable, first-mover footprint the audit shows gets selected. Kompozy does not astroturf Reddit; it removes the reason most creators go quiet — the production and scheduling grind — so genuine, consistent participation is finally sustainable.

Quick takeaways

  • A University of Illinois audit (arXiv preprint, September 13, 2026) traced 30,000 Reddit-generated answers through 14.68M comments across 20 advice/support subreddits.
  • Vote ranking was the strongest predictor of selection: the median chosen comment sat at the 91st percentile for score vs the 45th for passed-over comments.
  • 92% of selected comments were top-level replies to the post, appearing a median of 1.2 hours after it (vs 5.9 hours for non-selected).
  • Formal, directive language survived at higher rates (~49% higher odds per SD of formality); experiential and supportive language did worse.
  • First-person voice was stripped in synthesis — "I"/"my" fell from ~3.3% of quoted comments to ~0.06% of final answers.

Frequently asked questions

What did the Reddit AI search audit find?

Researchers at the University of Illinois Urbana-Champaign audited Reddit's generative search by running 10,000 questions from 20 advice/support subreddits three times (30,000 answers) and tracing selection across 14.68 million comments. They found the feature strongly favors comments that are already prominent — highly upvoted, top-level replies posted early — and prefers formal, directive language over personal, experiential, or supportive language. The written answers also stripped first-person voice from the comments they quoted.

How much does comment popularity affect selection?

Vote ranking within a thread was the single strongest predictor of being selected. The median selected comment sat at the 91st percentile for score in its thread, compared with the 45th percentile for comments that were not selected, and 92% of selected comments were direct top-level replies to the original post.

Is this an official Reddit finding?

No. It is an independent academic audit, "The Wisdom of the Loudest: A Large-Scale Audit of Generative Search on Reddit," posted to arXiv on September 13, 2026 by Agam Goyal, Claire Wang, and Eshwar Chandrasekharan. It is a preprint that had not yet been peer-reviewed, so treat the specific figures as early results.

What should creators do about it?

Do not try to fake upvotes — Reddit communities catch and punish that. Focus on being genuinely early, clear, and useful: post promptly, near the top of relevant threads, in structured and directive language rather than long anecdotes. A content engine like Kompozy makes that consistency sustainable by generating on-brand posts from one source and scheduling them across Reddit (a Direct Connect destination) and every other platform.

Related news

← All AI news · Get started →