Industry

Reddit Deploys Large Language Models to Curb AI-Generated Spam

Reddit has begun relying on large language models to detect and block spam—much of which is itself generated by similar models.

Reddit Deploys Large Language Models to Curb AI-Generated Spam

Reddit has begun using large language models (LLMs) to detect and block spam on its platform, targeting a substantial portion of content that is itself generated by similar automated systems.

Scope and reported results

  • The company says its systems block about 23 million spam views per day.
  • Reddit reports identifying roughly 25,000 new spam posts and comments each day.
  • According to Reddit, deploying LLM-based tools reduced users' exposure to spam by 20 percent from January to March compared with the prior quarter.

Why this matters

Reddit has long pitched itself as a place for human-driven discussion—“the last human room” online—used by people seeking content that feels less produced by algorithms. As generative language models proliferated, however, large volumes of automated and coordinated content began appearing on the site, in many cases evading older, rule-based filters.

By using LLMs to police spam, Reddit is effectively fighting machine-generated content with other machines. These models are able to operate at the speed and scale necessary to detect patterns and coordinated manipulations that legacy filters missed.

Implications and limits

  • The reported 20 percent reduction in spam exposure suggests the approach has improved the user experience in the short term.
  • The move also highlights a tension: platforms that defined themselves against algorithmic content now depend on advanced automation to remain usable.
  • Reddit has not published detailed technical information about the models or their false-positive rates, which makes it difficult to fully assess long-term effectiveness and potential collateral moderation issues.

Conclusion

Reddit's shift underscores a broader pattern across social platforms: as AI-generated content scales, so does the need for automated defenses. Maintaining a readable and trustworthy feed increasingly requires machines to manage the output of other machines.