Optimizing the spatial configuration of a seven-talker speech display-Reference-Cited by-同舟云学术

Optimizing the spatial configuration of a seven-talker speech display

Published:2005-10 Issue:4 Volume:2 Page:430-436
ISSN:1544-3558
Container-title:ACM Transactions on Applied Perception
language:en
Short-container-title:ACM Trans. Appl. Percept.

Author:

Brungart Douglas S.¹,Simpson Brian D.¹

Affiliation:

1. Human Effectiveness Directorate, Air Force Research Laboratory, WPAFB, OH

Abstract

Although there is substantial evidence that performance in multitalker listening tasks can be improved by spatially separating the apparent locations of the competing talkers, very little effort has been made to determine the best locations and presentation levels for the talkers in a multichannel speech display. In this experiment, a call sign based color and number identification task was used to evaluate the effectiveness of three different spatial configurations and two different level normalization schemes in a seven-channel binaural speech display. When only two spatially adjacent channels of the seven-channel system were active, overall performance was substantially better with a geometrically spaced spatial configuration (with far-field talkers at −90°, −30°, −10°, 0°, +10°, +30°, and +90° azimuth) or a hybrid near-far configuration (with far-field talkers at −90°, −30°, 0°, +30°, and +90° azimuth and near-field talkers at ±90°) than with a more conventional linearly spaced configuration (with far-field talkers at −90°, −60°, −30°, 0°, +30°, +60°, and +90° azimuth). When all seven channels were active, performance was generally better with a “better-ear” normalization scheme that equalized the levels of the talkers in the more intense ear than with a default normalization scheme that equalized the levels of the talkers at the center of the head. The best overall performance in the seven-talker task occurred when the hybrid near-far spatial configuration was combined with the better-ear normalization scheme. This combination resulted in a 20% increase in the number of correct identifications relative to the baseline condition with linearly spaced talker locations and no level normalization. Although this is a relatively modest improvement, it should be noted that it could be achieved at little or no cost simply by reconfiguring the HRTFs used in a multitalker speech display.

Publisher

Association for Computing Machinery (ACM)

Subject

Experimental and Cognitive Psychology,General Computer Science,Theoretical Computer Science

Link

https://dl.acm.org/doi/pdf/10.1145/1101530.1101538

Reference8 articles.

1. A speech corpus for multitalker communications research;Bolia R.;Journal of the Acoustical Society of America,2000

2. Auditory localization of nearby sources. i: Head-related transfer functions;Brungart D.;Journal of the Acoustical Society of America,1999

3. The effects of spatial separation in distance on the informational and energetic masking of a nearby speech signal;Brungart D.;Journal of the Acoustical Society of America,2002

4. Evaluation of the “cocktail party effect” for multiple speech stimuli within a spatial audio display;Crispien K.;Journal of the Audio Engineering Society,1995

5. Speech intelligibility and localization in a multi-source environment;Hawley M.;Journal of the Acoustical Society of America,1999

Cited by 17 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Vigilance to Spatialized Auditory Displays: Initial Assessment of Performance and Workload;Human Factors: The Journal of the Human Factors and Ergonomics Society;2022-12-01

2. Investigating cognitive workload in concurrent speech-based information communication;International Journal of Human-Computer Studies;2022-01

3. Evaluation of Information Comprehension in Concurrent Speech-based Designs;ACM Transactions on Multimedia Computing, Communications, and Applications;2020-11-30

4. Investigating efficient speech-based information communication: a comparison between the high-rate and the concurrent playback designs;Multimedia Systems;2020-07-02

5. Spatial Release From Masking in Adults With Bilateral Cochlear Implants: Effects of Distracter Azimuth and Microphone Location;Journal of Speech, Language, and Hearing Research;2018-03-15