Checked for new stories 19m ago

Updates on Content Moderation

Every AI story we track on Content Moderation — 43 stories so far, each summarized in our own words and linked back to the publisher that reported it.

Pulled from 124 sources

This week

Social Media8 min read

Facebook Is Hosting Huge Numbers of Horrifying AI-Generated Videos of Violent Child Abuse, and Meta Is Barely Even Pretending to Care About Taking Them Down

Futurism
AI Research5 min read

No monitoring system caught the German wiki. Two outside researchers found it by searching the internet.

The Next Web
Social Media6 min read

Instagram’s AI detection is a mess (again)

The Verge
Social Media2 min read

MauroPello/stop-the-slop: Spot AI-generated YouTube scripts

Hacker News

This month

Social Media3 min read

LinkedIn's 'seems like AI slop' button is a hit

Covered by 2 sources
AI Research4 min read

Amazon will train on Twitch streamers’ content by default, unless they opt out

Covered by 4 sources
Social Media2 min read

Over 1 million people have clicked LinkedIn’s AI slop button

The Verge
Cybersecurity2 min read

Another woman joins lawsuit accusing Grok of generating CSAM

Engadget
Social Media2 min read

Twitch streamers can now opt out from training Amazon’s AI

Covered by 4 sources
Social Media3 min read

Twitch CPO squirms, admits everyone hates its new AI training 'feature'

Hacker News
Music & Audio4 min read

Spotify will label ‘AI Persona’ profiles and exclude their music from recommendations

Covered by 2 sources
AI Research6 min read

Mistral AI Releases Shieldstral 1.0 3B: An Open-Weights Policy-Adaptive Multimodal Safety Classifier Matching Models 7× Its Size

MarkTechPost
Social Media3 min read

AI isn’t enough to protect social media communities from AI

Ars Technica
Social Media3 min read

Meta apps displayed ads that contained AI-generated CSAM

Covered by 2 sources
Social Media4 min read

Old Reddit could be the next casualty of Reddit's war on AI scraping

Covered by 2 sources
Social Media2 min read

Telegram CEO says 'takedown extortionist' was responsible for the app being briefly delisted by Apple

Engadget
Social Media3 min read

Telegram’s Brief App Store Removal Renews Questions About Apple’s CSAM Enforcement

CNET
AI Research5 min read

Mistral's Shieldstral: 3B open-weights model for multimodal moderation

Covered by 2 sources
Social Media8 min read

From Substack to YouTube, here are the social platforms cracking down on AI slop

Business Insider
Social Media2 min read

Snapchat no longer rewards fully AI-generated Spotlight content

Covered by 3 sources
Social Media3 min read

LinkedIn to add AI slop report feature that could train better AI slop

Mashable
Cybersecurity4 min read

China drafts cyberbullying rules that reach AI-generated abuse

The Next Web
Dev20 min read

Best practices for applying Amazon Bedrock Guardrails to code generation workflows

AWS Blog
Agents31 min read

‘I Want Everything Completely Uncensored’: Here’s What Grok Users Are Complaining About to the FTC

Gizmodo
Social Media3 min read

Discord admits AI moderation bug wrongfully banned users over harmless images

TechCrunch

Big AI Had a Point When It Said It Needed to Be Told What Is Not Okay

Gizmodo

Spotify removed 57,000 fake podcast episodes promoting illegal drugs, but only after a senator forced its hand

The Next Web

Chinese activist in UK told by X that abusive deepfakes do not breach rules

The Guardian

Synthesia partners with Cinder to scale moderation before a frame renders

The Next Web

Prompting Amazon Nova 2 for content moderation

AWS Blog

‘Never Talk About Goblins’: OpenAI’s Instructions to Codex Have a Weirdly Emphatic No-Creatures Policy

Gizmodo

Claude Leak Shows That Anthropic Is Tracking Users’ Vulgar Language and Deems Them “Negative”

Futurism

The Facebook insider building content moderation for the AI era

TechCrunch

Teenager died after asking ChatGPT for ‘most successful’ way to take his life, inquest told

The Guardian

Wikipedia Editors Tried and Tried to Work With AI Content, Eventually Realized It Was Total Trash and Banned It Entirely

Futurism

OpenAI Cancels Spicy “Adult Mode” Chatbot as Crisis Deepens

Covered by 3 sources

Meta rolls out new AI content enforcement systems while reducing reliance on third-party vendors

TechCrunch

“Educational” YouTube AI Slop Encourages Kids to Play in Traffic

Futurism

Character.AI Is Hosting Epstein Island Roleplays Scenarios and Ghislaine Maxwell Bots

Futurism

Google faces lawsuit after Gemini chatbot instructed man to kill himself

The Guardian

OpenAI policy exec who opposed chatbot’s “adult mode” reportedly fired on discrimination claim

TechCrunch

India orders social media platforms to take down deepfakes faster

TechCrunch
That's everything we have on Content Moderation right now