AI isn’t enough to protect social media communities from AI
Ars Technica Scharon Harding ● Covered by 2 sources
Reddit's r/AskHistorians got hit by mass, automated removals of years-old posts, seemingly flagged as AI-made junk. Turns out leaning on AI to police AI content just creates new messes for human moderators to clean up.
There's a certain irony baked into the modern moderation playbook: platforms are throwing AI at the AI slop problem, and the fires are spreading instead of dying out. Ars Technica's look at r/AskHistorians makes the case plainly. In April, moderators of that Reddit community watched their Slack channel light up with modmail alerts as dozens of posts and comments, some a decade old, got automatically yanked from the subreddit. Nothing had changed about those old contributions. What changed was the detection system sitting on top of them, apparently deciding that longstanding, carefully sourced historical answers looked enough like machine-generated filler to get pulled.
That's the uncomfortable part. AskHistorians built its reputation on rigorous, human-written answers, the kind that take hours of research and citation-checking to produce. Those are exactly the qualities an automated slop-detector is supposed to protect. Instead, the tooling apparently swept up genuine work alongside whatever it was actually hunting, and moderators were left doing archaeology on their own subreddit to figure out what got erased and why.
The deeper issue isn't a buggy classifier or a bad training run. It's the assumption that authenticity is something software can reliably certify at scale. Social media's value has always come from people willing to write a genuinely useful PC-building guide or a blog post nobody else could have written. Automated filters trained to spot AI-generated junk end up making probabilistic guesses about tone, phrasing, and pattern, and those guesses inevitably catch real, careful human writing in the net, especially older content written in styles that don't match whatever the model considers
My take
Handing moderation to more AI is a lazy shortcut dressed up as a scaling solution, and communities that took years to build trust are the ones paying for it. Platforms love automation because it's cheap, but cheap detection means false positives on exactly the thoughtful, human contributions that made the platform worth using, which is a bad trade any way you slice it.
Read more about this at: Ars Technica