Creating speech zones with self-distributing acoustic swarms-Reference-Cited by-同舟云学术

Creating speech zones with self-distributing acoustic swarms

Published:2023-09-21 Issue:1 Volume:14 Page:
ISSN:2041-1723
Container-title:Nature Communications
language:en
Short-container-title:Nat Commun

Author:

Itani Malek^ORCID,Chen Tuochao^ORCID,Yoshioka Takuya^ORCID,Gollakota Shyamnath^ORCID

Abstract

AbstractImagine being in a crowded room with a cacophony of speakers and having the ability to focus on or remove speech from a specific 2D region. This would require understanding and manipulating an acoustic scene, isolating each speaker, and associating a 2D spatial context with each constituent speech. However, separating speech from a large number of concurrent speakers in a room into individual streams and identifying their precise 2D locations is challenging, even for the human brain. Here, we present the first acoustic swarm that demonstrates cooperative navigation with centimeter-resolution using sound, eliminating the need for cameras or external infrastructure. Our acoustic swarm forms a self-distributing wireless microphone array, which, along with our attention-based neural network framework, lets us separate and localize concurrent human speakers in the 2D space, enabling speech zones. Our evaluations showed that the acoustic swarm could localize and separate 3-5 concurrent speech sources in real-world unseen reverberant environments with median and 90-percentile 2D errors of 15 cm and 50 cm, respectively. Our system enables applications like mute zones (parts of the room where sounds are muted), active zones (regions where sounds are captured), multi-conversation separation and location-aware interaction.

Funder

Gordon and Betty Moore Foundation

Publisher

Springer Science and Business Media LLC

Subject

General Physics and Astronomy,General Biochemistry, Genetics and Molecular Biology,General Chemistry,Multidisciplinary

Link

https://www.nature.com/articles/s41467-023-40869-8.pdf

Reference56 articles.

1. Grumiaux, P.-A., Kitić, S., Girin, L. & Guérin, A. A survey of sound source localization with deep learning methods. J. Acoust. Soc. Am. 152, 107–151 (2022).

2. Yu, J., Han, S. D., Tang, W. N. & Rus, D. A portable, 3d-printing enabled multi-vehicle platform for robotics research and education. 2017 IEEE International Conference on Robotics and Automation (ICRA), pages 1475–1480 (2017).

3. Le Goc, M. et al. Zooids: Building blocks for swarm user interfaces. In Proc of the 29th Annual Symposium on User Interface Software and Technology, page 97–109 (2016).

4. Özgür, A. et al. Cellulo: Versatile handheld robots for education. In Proc of the 2017 ACM/IEEE International Conference on Human-Robot Interaction, page 119–127 (2017).

5. Basiri, M., Schill, F., Floreano, D. & Lima, P. U. Audio-based localization for swarms of micro air vehicles. 2014 IEEE international conference on robotics and automation (ICRA), pages 4729–4734 (2014).

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Meta-barriers for ventilated sound reduction via transformation acoustics;International Journal of Mechanical Sciences;2024-07

2. Scalable Acoustic IoT through Composable Distributed Beamforming Tags;2024 23rd ACM/IEEE International Conference on Information Processing in Sensor Networks (IPSN);2024-05-13

3. A Unified Geometry-Aware Source Localization and Separation Framework for AD-HOC Microphone Array;2024 IEEE International Conference on Acoustics, Speech, and Signal Processing Workshops (ICASSPW);2024-04-14

4. Real-time control of a hearing instrument with EEG-based attention decoding;2024-03-04