Familiar and unfamiliar speaker recognition assessment and system emulation for cochlear implant users-Reference-Cited by-同舟云学术

Familiar and unfamiliar speaker recognition assessment and system emulation for cochlear implant users

Published:2023-02-01 Issue:2 Volume:153 Page:1293-1306
ISSN:0001-4966
Container-title:The Journal of the Acoustical Society of America
language:en
Short-container-title:

Author:

Mamun Nursadul¹,Ghosh Ria¹,Hansen John H. L.¹^ORCID

Affiliation:

1. Cochlear Implant Processing Laboratory—Center for Robust Speech Systems (CRSS-CILab), The University of Texas at Dallas , 800 West Campbell Road, Richardson, Texas 75080, USA

Abstract

In the area of speech processing, human speaker identification under naturalistic environments is a challenging task, especially for hearing-impaired individuals with cochlear implants (CIs) or hearing aids (HAs). Motivated by the fact that electrodograms reflect direct CI stimulation of input audio, this study proposes a speaker identification (ID) investigation using two-dimensional electrodograms constructed from the responses of a CI auditory system to emulate CI speaker ID capabilities. Features are extracted from electrodograms through an identity vector (i-vector) framework to train and generate identity models for each speaker using a Gaussian mixture model-universal background model followed by probabilistic linear discriminant analysis. To validate the proposed system, perceptual speaker ID for 20 normal hearing (NH) and seven CI listeners was evaluated with a total of 41 different speakers and compared with the scores from the proposed system. A one-way analysis of variance showed that the proposed system can reliably predict the speaker ID capability of CI (F[1,10] = 0.18, p = 0.68) and NH (F[1,20] = 0, p = 0.98) listeners in naturalistic environments. The impact of speaker familiarity is also addressed, and the results show a reduced performance for speaker recognition by CI subjects using their CI processor, highlighting limitations of current speech processing strategies used in CIs/HAs.

Funder

National Institute of Health

Univ. of Texas at Dallas

Publisher

Acoustical Society of America (ASA)

Subject

Acoustics and Ultrasonics,Arts and Humanities (miscellaneous)

Link

https://pubs.aip.org/asa/jasa/article-pdf/153/2/1293/16654585/1293_1_online.pdf

Reference42 articles.

1. The CCi-MOBILE vocoder;J. Acoust. Soc. Am.,2018

2. Within subject comparison of advanced coding strategies in the Nucleus 24 cochlear implant,1999

3. Understanding voice perception;Br. J. Psychol.,2011

4. Testing with the YOHO CD-ROM voice verification corpus,1995