Acoustic synthesis for AR/VR experiences @AIatMeta
Acoustic synthesis for AR/VR experiences  @AIatMeta
Uploaded June 2022 | Updated September 2026, 4 hours ago
Existing AI models do a good job of understanding images but require more work to understand the acoustics of environments in the related image.

This is why researchers from Meta AI and the University of Texas are open-sourcing three new models for audio-visual understanding of human speech and sounds in video, helping us achieve immersive AR and VR experiences much more quickly.

Using multimodal AI models that can take audio, video, and text signals at one time, AI will be able to deliver sound quality that realistically matches the settings people are immersed in.

Learn more about this state-of-the-art work here (insert link).
Acoustic synthesis for AR/VR experiencesIntroducing SAM Audio: The First Unified Multimodal Model for Audio Separation | AI at MetaLIGHT, a multi-player crowdsourced text adventure game for dialogue researchExpire-Span: Teaching AI How to Forget at ScaleArchitecture of Metas First-Generation AI Inference AcceleratorSAM 3: Building a unified model architecture for detection and trackingMSVP - Metas First In-House Silicon for Video Processing | AI at MetaBuilding diverse MiniHack environments with just a few lines of codeHarmful content can evolve rapidly and it’s crucial for AI systems to evolve alongside it.Using AI to deliver more inclusive biographical content on WikipediaDiplomacy Gameplay with CICERO | AI at MetaMeet Llama 3.1
AI at Meta |

Acoustic synthesis for AR/VR experiences

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER