The Mandarin Chinese Speech Database: A Large Corpus for Auditory Neutral Nonsense Pseudo-Sentences-Reference-Cited by-同舟云学术

The Mandarin Chinese Speech Database: A Large Corpus for Auditory Neutral Nonsense Pseudo-Sentences

Published:2024-08-22 Issue: Volume: Page:
ISSN:
Container-title:
language:
Short-container-title:

Author:

Zhou Anqi¹,Li Qiuhong¹,Wu Chao¹

Affiliation:

1. Peking University

Abstract

Word frequency, context, and length are three core elements that impact speech perception. Considering the limitations of previous Chinese stimulus databases, such as non-standardized sentence structures, uncontrolled emotional information that may exist in semantics, and a relatively small number of voice items, we developed an abundant and reliable Chinese Mandarin nonsense pseudo-sentences database with fixed syntax (pronoun + subject + adverbial + predicate + pronoun + object), lengths (6 two-character words), and high-frequency words in daily life. The high-frequency keywords (subject, predicate, and object) were extracted from China Daily. Ten native Chinese participants (five women and five men) evaluated the sentences. After removing sentences with potential emotional and semantic content valence, 3,148 meaningless neutral sentence text remained. The sentences were recorded by six native speakers (three males and three females) with broadcasting experience in a neutral tone. After examining and standardizing all the voices, 18,820 audio files were included in the corpus (https://osf.io/ra3gm/?view_only=98c3b6f1ee7747d3b3bcd60313cf395f). For each speaker, 12 acoustic parameters (duration, F0 mean, F0 standard deviation, F0 minimum, F0 maximum, harmonics-to-noise ratio, jitter, shimmer, in-tensity, root-mean-square amplitude, spectral center of gravity, and spectral spread) were retrieved, and there were significant gender differences in the acoustic features (all p < 0.001). This database could be valuable for researchers and clinicians to investigate rich topics, such as children’s reading ability, speech recognition abilities in different populations, and oral cues for orofacial movement training in stutterers.

Publisher

Springer Science and Business Media LLC

Reference56 articles.

1. The cocktail party problem;McDermott JH;Current Biology,2009

2. Some Experiments on the Recognition of Speech, with One and with Two Ears;Cherry EC;Journal of the Acoustical Society of America,1953

3. Effect of priming on energetic and informational masking in a same-different task;Jones JA;Ear And Hearing,2012

4. Altered maturation and atypical cortical processing of spoken sentences in autism spectrum disorder;Alho J;Progress In Neurobiology,2021

5. Measuring Mandarin Speech Recognition Thresholds Using the Method of Adaptive Tracking;Wang Y;Journal Of Speech, Language, And Hearing Research : Jslhr,2019