Don't miss the latest stories
Advertise Newsletter
Network
  • The Creative Finder
  • The Bazaar
  • Deals
  • Status Is Down
Community
  • Sign up / Log in
  • Discussion Forums
  • Calendar of Events
NEW

Follow

Share this

Meta
Artificial Intelligence
Imaging
Productivity
Research
Science
Technology
More
  • Video
  • Productivity
  • Research
  • Science
  • Technology
  • Video
MENU
  • Advertise with us
  • Submit tip/feedback
  • Work with us
  • Subscribe to newsletter
  • Subscribe to RSS
Advertise here
Advertisement

Meta Debuts Open-Source ‘ImageBind’ AI That Learns To Mimic Human Perception

By Alexa Heah, 10 May 2023

Subscribe to newsletter
Like us on Facebook
Image ID 251724386 © via Daniel Constante | Dreamstime.com

 

Last month, Meta released an open-source artificial intelligence (AI) tool that helped creatives turn doodles into animation at a click of a button. Now, the technology giant is making yet another system open to the public, dubbed ‘ImageBind’.


Unlike common image generators that create images from a series of words, the software allows users to link text, images, videos, audio clips, 3D measurements, temperature data, and motion data. More impressively, it offers all of these options without having to train on every possibility.


ImageBind works by predicting the connections between different sets of data groups, similar to how humans perceive the environment around them. Instead of just conjuring up an image, the model can attach sounds, temperatures, or even precise locations relevant to the scene.

 

Image via Meta AI

 

Through this, AI moves closer to human comprehension, mimicking the way the brain works to absorb various stimulants in the surroundings, such as the sights, sounds, and other sensory experiences that come as second nature to us.

 

“For example, a creator could couple an image with an alarm clock and a rooster crowing, and use a crowing audio prompt to segment the rooster or the sound of an alarm to segment the block and animate both into a video sequence,” Meta explained.

Advertisement
Advertisement


According to graphs provided by the company, it seems ImageBind far outperforms other algorithms in both audio accuracy and depth accuracy, showing that it’s possible to create a “joint embedding space” across multiple modalities without having to train the AI on every single one.

 

Image via Meta AI

 

“This is important because it’s not feasible for researchers to create datasets with samples that contain, for example, audio data and thermal data from a busy city street, or depth data and a text description of a seaside cliff,” the blog post continued.


Soon, Meta envisions the technology will expand beyond the current “six senses” to include the likes of touch, speech, smell, and even brain fMRI signals that will give a “holistic” approach to machine learning and future mixed-reality gadgets.


Check out the code on GitHub.

 

Image via Meta AI

 

 

 

[via Engadget and Screen Rant, images via various sources]

Receive interesting stories like this one in your inbox
Advertise here

More related news

Advertise here
Also check out these recent news
Web Design
Link to news page

When Your Website Goes Down, This Is the Page Customers Meet Instead

2027
Link to news page

2027 Already Has A Color Of The Year And It’s Beginning On ‘Grounded’ Territory

IKEA
Link to news page

IKEA & Xbox Press Play On Furniture & Storage Inspired By The Iconic Controller

Fashion
Link to news page

Vogue Presents ‘United Flags of Fashion’ With Top Designers For All 50 States

Coca-Cola
Link to news page

Coca-Cola Pours Fresh Life Into Its Iconic Branding With Worldwide Redesign