Reddit's AI Search Buried the Human Voice. The Voice Was Never the Variable.

Reddit's AI search prefers the comment that reads like a memo.
A University of Illinois Urbana-Champaign team ran 10,000 questions through the feature three times, analyzed the 30,000 answers, and traced them back through 14.68 million comments across 20 advice and support subreddits. A comment one standard deviation more formal had 49 percent higher odds of being selected. Comments carrying markers of personal experience, first-person pronouns and past tense, the exact thing people open Reddit for, came in at 0.789. Then the written answer scrubbed off what survived. Words like "I" and "my" ran 3.3 percent in the quoted comments and 0.06 percent in the answers built on top of them.
There's an obvious conclusion sitting on that pile, and I'll watch people reach for it all month. Sound less like a person. Write the buttoned-up version. Feed the machine what it eats.
That's the wrong lesson, and the paper's own title says so. It's called "The Wisdom of the Loudest".
Formality got the headline because it was the strongest language predictor. Vote rank inside the thread was the one that moved selection.
One standard deviation there multiplied the odds by 2.88. That's more than double what formality managed, and formality's 1.488 was the number before the researchers controlled for anything. Once they accounted for a comment's score, position and age, formality dropped to 1.213 and the experiential penalty softened to 0.860.
The median selected comment sat at the 91st percentile for score in its thread. The median comment that got passed over sat at the 45th. Comments scoring zero or below were 5.1 percent of everything collected and 0.53 percent of what got picked.
The system is reading the room's verdict and repeating it back. Formal writing correlates with upvotes, upvotes drive selection, and somewhere along that chain a correlation gets repackaged as a writing tip.
The authors are blunt about the limits. Their sentence: "Our study is observational and should not be interpreted causally." It's a preprint. It hasn't been peer reviewed. And the conclusion being drawn from it is the one it supports least.
I ran the paper's own measure on my last three posts.
The researchers scored experiential voice partly on first-person pronoun density. That's cheap to replicate, so I ran it against the last three things published on this site, dated September 5, 6 and 7.
The September 7 post, the one about OpenAI charging me $120 after I told them I was leaving, runs 4.74 percent "I" and "my." That's higher than the 3.3 percent the researchers measured in the Reddit comments that actually made it into answers. The September 6 post, a read on how Anthropic is framing its EU AI Act compliance, runs 0.25 percent. The September 5 post about shoppers standing confused in the aisle runs 0.35 percent.
Same writer. Same site. Three consecutive days. Nineteen times the first-person density on one of them.
I wasn't turning a dial. The OpenAI post was first person because it happened to me and there was no honest way to write it otherwise. The Anthropic post was a read on a company's public posture and my pronouns had no business anywhere near it.
That's where the "write formally for AI search" advice breaks at the root. Register follows from whether you were actually there. Anybody selling you a voice adjustment as an ai content strategy is selling you a knob that doesn't exist.
Strip the language findings out and four variables are left. Every one of them is mechanical.
Timing. Selected comments showed up a median of 1.2 hours after the post. The ones that got passed over showed up at 5.9 hours.
Position. Direct replies to the original post were 53 percent of all comments collected and 92 percent of the ones selected.
Length. One standard deviation more of it, and the odds went up 1.79.
Links. A comment carrying an external link ran 2.25.
Early, direct, substantial, sourced. Not one of those asks you to sand down how you sound. They're all about where you show up and when.
Here's the uncomfortable half. A system that selects the 91st-percentile comment is finding the answer the crowd already found, and then handing it a second distribution channel. The good reply posted six hours late by somebody who actually lived the problem stays buried, and now it stays buried in two places instead of one.
That's what retrieval does. Anything that ranks by existing signal will hand its advantage to whoever already had one.
So an ai content strategy built on rewriting your voice is optimizing a 1.213 while ignoring a 2.88. It's the marketing equivalent of repainting the barn while the fence is down.
What actually moves is boring and it's the same as it was. Be early to the thing in your category people are asking about. Answer the question that was asked instead of the adjacent one you'd rather talk about. Say enough to be usable. Point at the source. And publish where you own the URL, because every platform eventually meters the thing it gave you for free.
None of that requires you to stop sounding like yourself. The voice was never the variable. It was just the easiest thing to measure, so it got measured, and then it got merchandised.
If your brand is currently paying somebody to make your copy sound more institutional so the machines will like it better, stop. Take the same money and go answer the twelve questions your buyers are actually typing, within a day of them typing it, with your name on the answer. That's the work. If you want someone to run it with you, that's what a food and beverage marketing agency is supposed to be for.
adage, emmy, telly & webby award-winning digital marketing consultant for purpose-driven food & beverage brands.




