The Future is Hear: Advances in Computational Audition and Sound Manipulation @allenai
The Future is Hear: Advances in Computational Audition and Sound Manipulation  @allenai
Uploaded June 2022 | Updated September 2026, 3 days ago
Abstract:
Computer Audition is the sonic analog to Computer Vision and encompasses much more than just speech to text. Northwestern University’s Interactive Audio Lab (IAL), headed by Prof. Bryan Pardo, is a world leader in Computer Audition. IAL develops new techniques and technologies for identifying sound sources (telling dog barks from human speech), labeling content (naming the song played live in a live concert), parsing audio scenes into their constituent parts (e.g. splitting the live concert recording into individual instrumental recordings), searching for sounds in large datasets (e.g. matching the sound of an automobile with engine troubles to prior known cases in a database), and manipulating the selected audio (changing the prosody of speech, making adversarial speech examples to spoof voice ID). In this talk, Prof. Pardo will give an overview of recent work in the lab on these topics.

Bio:
Bryan Pardo is head of Northwestern University’s Interactive Audio Lab and co-director of the Northwestern University HCI+Design institute. Prof. Pardo has appointments in in the Department of Computer Science and Department of Radio, Television and Film. He received a M. Mus. in Jazz Studies in 2001 and a Ph.D. in Computer Science in 2005, both from the University of Michigan. He has authored over 100 peer-reviewed publications. He has developed speech analysis software for the Speech and Hearing department of the Ohio State University, statistical software for SPSS and worked as a machine learning researcher for General Dynamics. He has collaborated on and developed technologies acquired and patented by companies like Bose, Adobe and Ear Machine. While finishing his doctorate, he taught in the Music Department of Madonna University. When he is not teaching or researching, he performs on saxophone and clarinet with the bands Son Monarcas and The East Loop.
The Future is Hear: Advances in Computational Audition and Sound ManipulationSkill it! A Data-Driven Skills Framework for Understanding and Training Language ModelsAsta | an agentic ecosystem that advances scientific discoveryOpenBot: Turning Smartphones into Robots | Embodied AI Lecture Series at AI2Rethinking LLM efficiencyWere bringing training text out in the open, Introducing OLMoTraceSide Effects May Include Homogenization and Overusing ClichesGooAQ: Open Question Answering with Diverse Answer Types | AI2Language AI for RNA Virus and RNA VaccineDoing for our robots what nature did for us | Embodied AI Lecture series at AI2Optimal Transport Posterior Alignment for Cross-lingual Semantic ParsingFiguring out how the world works: causality in a world full of real people
Ai2 |

The Future is Hear: Advances in Computational Audition and Sound Manipulation

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER