Can AI moderate social media content? Here’s why it falls short
This article examines the limitations of AI in moderating harmful content on large social media platforms like Meta's apps. It discusses the company's internal shifts in AI strategy, including the development of new models and the acquisition of Scale AI.
Why it matters
The effectiveness of AI moderation is a central concern for the safety of billions of users on social media platforms.
It’s easy, with technology, to miss the forest for the trees. Every new AI advance grabs the limelight, and while everyone argues about the latest model, the older, uglier problem keeps festering underneath: large social platforms, used daily by hundreds of millions of teenagers, are still flooded with harmful, violent and risqué content. This is happening while adolescent brains are going through a critical stretch of development.
Take Meta. Facebook and Instagram sit inside what the company calls its “Family” of apps — those two, plus WhatsApp and Messenger — which together reach an estimated 3.5 billion people a day. Mark Zuckerberg has spent the past couple of years rewiring all of it with home-grown AI and new recommendation systems. The results haven’t matched what he expected, but the underlying way content gets made and served has genuinely shifted.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in