A University of Illinois preprint that traced 30,000 Reddit-generated answers back through 14.68 million comments found that the feature overwhelmingly surfaces top-level, highly upvoted, formally worded replies — and strips the first-person voice out of the ones it does quote.
2026-09-18 · by Moe Ameen
Researchers at the University of Illinois Urbana-Champaign — Agam Goyal, Claire Wang, and Eshwar Chandrasekharan — published a large-scale audit of Reddit's generative search feature (the answers surfaced by what began as "Reddit Answers," launched December 2024 and merged into Reddit search on May 26, 2026). The preprint, titled "The Wisdom of the Loudest: A Large-Scale Audit of Generative Search on Reddit," went up on arXiv on September 13, 2026. It has not yet been peer-reviewed, so treat the figures as a first pass rather than settled science.
To run it, the team put 10,000 questions drawn from 20 advice- and support-seeking subreddits through the feature three times, producing 30,000 answers, and traced which comments the system selected across 14.68 million candidate comments. The central finding is that generative search on Reddit is not neutral summarization: it strongly favors comments that were already prominent. A comment's vote ranking inside its thread was the single strongest predictor of being chosen, and the median selected comment sat at the 91st percentile for score in its thread versus the 45th percentile for comments that were passed over. Ninety-two percent of selected comments were direct top-level replies to the original post, and selected comments had appeared a median of 1.2 hours after the post, against 5.9 hours for the ones left out — early and visible wins.
Style mattered too. Formal, directive language survived selection at higher rates: a comment one standard deviation more formal had roughly 49% higher odds of being picked (odds ratio about 1.49), and directive phrasing using words like "should" and "must" also helped. Comments rich in personal-experience markers or supportive, empathetic language were less likely to be selected. And when a first-person comment did make it in, the synthesis stripped its voice — first-person singular words like "I" and "my" fell from about 3.3% of the quoted source comments to roughly 0.06% of the final written answers. The lived-experience texture that makes Reddit useful largely does not survive the summarization step.
The honest read of this audit is a discoverability lesson, not a hack: the content an AI surfaces from a community is the content that was already prominent, early, clearly structured, and directive. You cannot fake upvotes on Reddit without getting caught, and you shouldn't try. What you CAN do is show up consistently and fast, with clear, on-brand, useful posts, across Reddit and every other place an AI reads — and that consistency-at-speed is exactly the gap a content engine closes. Reddit is a native Direct Connect destination in [Kompozy](/) via Bundle.social, alongside the eight primary social platforms plus blog and email, so the same brand voice ships everywhere in one pass.
Concretely: hand Kompozy one source — a post, a transcript, a note — and it generates roughly 25–35 finished assets across 18 formats, each governed by a [Persona Brief](/glossary/persona-brief) that keeps the voice tight, clear, and directive rather than rambling. [Autopilot](/glossary/autopilot) then schedules and publishes them so you are early and present in the threads and feeds that matter instead of late and sporadic. Pair the community presence with the surfaces AI actually cites most — a full blog article and structured social posts — and you build the recognizable, first-mover footprint the audit shows gets selected. Kompozy does not astroturf Reddit; it removes the reason most creators go quiet — the production and scheduling grind — so genuine, consistent participation is finally sustainable.
Researchers at the University of Illinois Urbana-Champaign audited Reddit's generative search by running 10,000 questions from 20 advice/support subreddits three times (30,000 answers) and tracing selection across 14.68 million comments. They found the feature strongly favors comments that are already prominent — highly upvoted, top-level replies posted early — and prefers formal, directive language over personal, experiential, or supportive language. The written answers also stripped first-person voice from the comments they quoted.
Vote ranking within a thread was the single strongest predictor of being selected. The median selected comment sat at the 91st percentile for score in its thread, compared with the 45th percentile for comments that were not selected, and 92% of selected comments were direct top-level replies to the original post.
No. It is an independent academic audit, "The Wisdom of the Loudest: A Large-Scale Audit of Generative Search on Reddit," posted to arXiv on September 13, 2026 by Agam Goyal, Claire Wang, and Eshwar Chandrasekharan. It is a preprint that had not yet been peer-reviewed, so treat the specific figures as early results.
Do not try to fake upvotes — Reddit communities catch and punish that. Focus on being genuinely early, clear, and useful: post promptly, near the top of relevant threads, in structured and directive language rather than long anecdotes. A content engine like Kompozy makes that consistency sustainable by generating on-brand posts from one source and scheduling them across Reddit (a Direct Connect destination) and every other platform.