New AI Can Separate Voices in a Crowd Like Magic
Imagine hearing one person clearly in a noisy room—this AI could make it happen.
arXivLabs is a framework that lets collaborators develop and share new arXiv features directly on the site. Individuals and organizations working with arXivLabs have embraced and accepted arXiv's values of openness, community, excellence, and user data privacy — and arXiv only works with partners who adhere to them.
Have an idea for a project that would add value for arXiv's community? You can learn more about arXivLabs.
- SepRQ can separate multiple voices from a single recording without needing clean examples to learn from.
- It uses a multi-scale approach, analyzing sound at both fine and coarse levels to isolate individual speakers.
- This could lead to better hearing aids, clearer phone calls, and improved voice assistants in noisy settings.
Why It Matters
This could make hearing aids and phone calls work better in noisy places, improving daily communication for millions.