As large language models (LLMs) have become easier to access, bad actors can more readily generate spam across the internet. Reddit says it has developed tools that use LLMs to detect and reduce spam and coordinated artificial hype — in many cases targeting content originally generated with LLMs themselves.
Scale and effectiveness
According to Reddit, its systems block about 23 million spam views per day and identify roughly 25,000 new spam posts and comments every day. In a company blog post, Reddit says it employs LLMs to catch the “highly subtle, coordinated patterns of fake behavior and artificial hype” that older systems sometimes missed.
Reddit also reports that these measures reduced users’ exposure to spam by 20% from January through March compared with the prior three months.
A shared challenge for platforms
Social platforms have used automated spam-reduction tools for years, but the spread of LLMs has introduced large volumes of content that can be harder to detect. Platforms such as YouTube, Meta, and Instagram allow AI-generated content if creators disclose it, while TikTok offers users a toggle to control how much AI-generated content they see.
Faster detection of AI-generated content could also enable platforms to flag and remove violative material such as hate speech more quickly. However, experts continue to emphasize that AI-driven content moderation is most effective when combined with human moderation: automation plus human review produces the most reliable results.
Why this matters
Reddit’s approach illustrates that platforms are increasingly using AI to fight AI-driven abuse. LLMs provide both a vector for abuse and a tool for defense, and combating spam and manipulation will require ongoing technological updates and human oversight to remain effective.



