Skip to main content
  1. Home
  2. Social Media
  3. Computing
  4. Web
  5. News

Facebook increasingly using AI to scan for offensive content

Add as a preferred source on Google

“Content moderation” might sound like a commonplace academic or editorial task, but the reality is far darker. Internationally, there are more than 100,000 people, many in the Philippines but also in the U.S., whose job each day is to scan online content to screen for obscene, hateful, threatening, abusive, or otherwise disgusting content, according to Wired. The toll this work takes on moderators can be likened to post-traumatic stress disorder, leaving emotional and psychological scars that can remain for life if untreated. That’s a grim reality seldom publicized in the rush to have everyone engaged online at all times.

Twitter, Facebook, Google, Netflix, YouTube, and other companies may not publicize content moderation issues, but that doesn’t mean the companies aren’t working hard to stop the further harm caused to those who screen visual and written content. The objective isn’t only to control the information that shows up on their sites, but also to build AI systems capable of taking over the most tawdry, harmful aspects of the work. Engineers at Facebook recently stated that AI systems are now flagging more offensive content than humans, according to TechCrunch.

Recommended Videos

Joaquin Candela, Facebook’s director of engineering for applied machine learning, spoke about the growing application of AI to many aspects of the social media giant’s business, including content moderation. “One thing that is interesting is that today we have more offensive photos being reported by AI algorithms than by people,” Candela said. “The higher we push that to 100 percent, the fewer offensive photos have actually been seen by a human.”

The systems aren’t perfect and mistakes happen. In a 2015 report in Wired, Twitter said when its AI system was tuned to recognize porn 99 percent of the time, 7 percent of those blocked images were innocent and filtered incorrectly, such as when photos of half-naked babies or nursing mothers are screened as offensive. The systems will have to learn, and the answer will be deep-learning, using computers to analyze how they, in turn, analyze massive amounts of data.

Another Facebook engineer spoke of how Facebook and other companies are sharing what they’re learning with AI to cut offensive content. “We share our research openly,” Hussein Mehanna, Facebook’s director of core machine learning, told TechCrunch. “We don’t see AI as our secret weapon just to compete with other companies.”

Bruce Brown
Bruce Brown Contributing Editor   As a Contributing Editor to the Auto teams at Digital Trends and TheManual.com, Bruce…
Meta just pulled its most controversial AI image generation feature days after launch
Meta is framing this as "hearing feedback," not as fixing a consent problem.
Instagram Muse Image

A couple of days ago, I covered Meta’s announcement of the Muse Image, an AI tool that lets users generate images based on someone’s Instagram profile without asking the account owner. 

I also highlighted the risks associated with it in another piece, along with steps for opting out. Three days later, the feature is no longer available. 

Read more
Your YouTube playlists can now become actual TV shows, but there’s a catch you need to know
YouTube just gave Partner Program creators the episodic infrastructure that Netflix has been using to keep audiences hooked for years.
Electronics, Mobile Phone, Phone

YouTube just gave its creators a tool that streaming platforms take for granted. I’m talking about the ability to structure content as proper episodic TV. 

If you're in the YouTube Partner Program and you’ve been organizing your videos into playlists while praying that the algorithm and your audience notice, then Shows is the upgrade you've been waiting for.

Read more
I knew there was plenty of AI slop on LinkedIn. Shocking report says the problem is far worse than suspected
LinkedIn app on App Store iPhone

I already knew LinkedIn was overflowing with posts written by AI, recycled leadership advice, and those god-awful lessons about entrepreneurship. A new report suggests the situation is considerably worse than even the platform’s feed makes it appear.

AI-detection company Pangram analyzed more than one million posts scanned through its Chrome extension across LinkedIn, X, Reddit, Medium, and Substack. LinkedIn represented approximately one-third of everything scanned, yet produced 62% of all content Pangram flagged as AI-generated.

Read more