Skip to main content
  1. Home
  2. Computing
  3. Social Media
  4. News

Facebook opens up its image-recognition AI software to everyone

Add as a preferred source on Google

The AI research division at Facebook is open sourcing its image recognition software with the aim of advancing the tech so it can one day be applied to live video. Facebook’s DeepMask, SharpMask, and MultiPathNet software is now available to everyone on GitHub.

Facebook previously laid out its image-recognition systems in a number of research papers, which are also being made available to the public along with its demos. At present, the company’s algorithms work in conjunction with its MultiPathNet convolutional neural networks — an AI that is fed huge amounts of data until it can autonomously recognize other data — allowing Facebook to understand an image based on each pixel it contains.

Recommended Videos

In order to classify and label the objects in an image, Facebook couples its DeepMask segmentation framework with its SharpMask segment refinement module. The final stage in Facebook’s machine vision system utilizes its MultiPathNet deep learning AI to label each object in the photo.

According to Facebook, AI machine vision software has progressed in leaps and bounds over the past few years, allowing the type of image classification that didn’t even exist a short while ago. Facebook claims that open sourcing the software is critical to its advancement.

Example images scanned by Facebook's complete image-recognition system
Example images scanned by Facebook’s complete image-recognition system Image used with permission by copyright holder

Deep learning techniques are springing up all over the big blue behemoth. The AI powers Facebook’s (controversial) facial-recognition feature, manages curation on its News Feed, and is even utilized within its digital assistant for Messenger.

This isn’t the first time Facebook has open sourced its AI. In fact, the company is somewhat of a trailblazer when it comes to sharing its tech. In December, Facebook submitted its state-of-the-art computer server dedicated to AI to the Open Compute Project — a group consisting of tech giants, such as Apple and Microsoft, that share the designs of their respective computer infrastructures.

Facebook is already predicting the future use cases for the image-recognition tech. The company reveals that it could potentially help it to build upon its existing AI generated image descriptions for the visually impaired.

“Currently, visually impaired users browsing photos on Facebook only hear the name of the person who shared the photo, followed by the term “photo,” when they come upon an image in their News Feed,” writes Piotr Dollar, research scientist at Facebook AI Research (FAIR), in a blog post. “Instead we aim to offer richer descriptions, such as ‘Photo contains beach, trees, and three smiling people.’”

Additionally, Facebook claims that its next challenge is to apply its image-recognition techniques to video, “where objects are moving, interacting, and changing over time,” and even Facebook Live broadcasts. “Real-time classification could help surface relevant and important Live videos on Facebook, while applying more refined techniques to detect scenes, objects, and actions over space and time could one day allow for real-time narration,” Dollar adds.

Saqib Shah
Saqib Shah is a Twitter addict and film fan with an obsessive interest in pop culture trends. In his spare time he can be…
ChatGPT’s Chrome extension can now read your tabs, YouTube videos, and highlighted text
ChatGPT now understands what's on your screen, letting you ask about tabs, highlighted text, and YouTube videos directly.
chatgpt-chrome-extension-upgrade-youtube-videos

ChatGPT just got a lot better at understanding what you're actually looking at. OpenAI has rolled out updates to its Chrome extension and desktop app that let ChatGPT read your open tabs, react to highlighted text, and answer questions about YouTube videos. The update arrived shortly after OpenAI shelved its standalone Atlas browser to shift its focus toward tools people already use.

https://twitter.com/ChatGPT/status/2082970812584432115

Read more
Lenovo’s Googlebook lineup could include two laptops, a 2-in-1 tablet, and an AI mouse
Images have emerged online giving us our first look at Lenovo’s upcoming Googlebooks
Googlebook

Google announced Googlebook at I/O in May, but we are still waiting to see what the finished hardware will actually look like. The first devices are expected this fall, and new images published by Android Headlines may have given us an early look at what Lenovo has planned.

Lenovo appears to be preparing at least three Googlebooks, including two traditional laptops and a detachable 2-in-1 tablet. There is also a new wireless AI mouse that seems to have been designed specifically for the platform.

Read more
Shopping for back-to-school? I’d recommend these GaN chargers for saving space and packing more power
From a $9 spare to a $56 desk hub, these chargers cut down on cable clutter so that you can focus on your studies, not how cluttered your desk looks.
Best GaN chargers in 2026 for back to school shopping.

Between a phone, a laptop, a tablet, and perhaps a pair of wireless earbuds that require charging every day or every few days, most students are packing multiple chargers for move-in day. Those big, heavy bricks take up valuable backpack space that could otherwise be free for carrying more items. 

However, GaN chargers fix that problem in two ways. First, they deliver the same amount of power from a much smaller body, thanks to gallium nitride transistors. Furthermore, if you're willing to spend a little more upfront, you can get a high-capacity, multi-device charger capable of powering your entire digital ecosystem from a single wall outlet, something that could be very useful for a dorm room.

Read more