Mark RadarMARK RADAR
About
EN
Sign in

Reddit Brings LLMs to Community Moderation

2 reports · First detected 2026-08-06 · Last active 2026-08-06

Reddit’s communities largely rely on volunteer moderators to write and enforce their own rules. Its longstanding AutoModerator system uses keywords, regular expressions and account attributes, making it powerful but often difficult to configure and less adept at interpreting context. Rules Hub instead uses large language models to assess whether posts and comments violate the intent of a rule, aiming to reduce routine work involving spam, abuse and generic violations while keeping human moderators in control.

Reddit began testing Rules Hub with a small group of communities on April 16, 2026, and announced a broader rollout on August 5. A day later, access expanded to every community created since July 5 and all newly created communities going forward. Moderators can test rules against historical content, then choose whether matching posts or comments should be removed, reported or only logged. Existing communities are unchanged for now, with wider availability planned for later in 2026.

All Coverage

2 original reports

The Backstory

The history behind this event
Reddit Deploys Large Language Models to Fight AI-Generated Fake Comments2026-07-11 · 1 reports · similarity 0.82

Reddit signed a data-licensing agreement with Google in 2024 worth $60 million annually, making the authenticity of its content commercially important. Marketers, however, have recently used AI for “generative engine optimization,” or GEO, planting fake comments to manipulate recommendations from chatbots such as ChatGPT. The practice has forced Reddit into an escalating battle against AI-powered troll networks.

Reddit announced an upgraded defense system in July 2026 that uses large language models to analyze accounts. The system currently blocks an average of about 25,000 suspicious posts and comments a day and has reversed nearly 2 million fraudulent votes. It also helped reduce users’ exposure to spam in early 2026 by about 20% from the same period a year earlier. The move signals that the technological arms race between social platforms and AI troll networks has formally intensified.

Mark Radar|MARK RADAR
All times are in Taipei time (GMT+8)