The Human Takes It All: Humanlike Synthesized Voices Are Perceived as Less Eerie and More Likable. Evidence From a Subjective Ratings Study-Reference-Cited by-同舟云学术

The Human Takes It All: Humanlike Synthesized Voices Are Perceived as Less Eerie and More Likable. Evidence From a Subjective Ratings Study

Published:2020-12-16 Issue: Volume:14 Page:
ISSN:1662-5218
Container-title:Frontiers in Neurorobotics
language:
Short-container-title:Front. Neurorobot.

Author:

Kühne Katharina,Fischer Martin H.,Zhou Yuefang

Abstract

Background: The increasing involvement of social robots in human lives raises the question as to how humans perceive social robots. Little is known about human perception of synthesized voices.Aim: To investigate which synthesized voice parameters predict the speaker's eeriness and voice likability; to determine if individual listener characteristics (e.g., personality, attitude toward robots, age) influence synthesized voice evaluations; and to explore which paralinguistic features subjectively distinguish humans from robots/artificial agents.Methods: 95 adults (62 females) listened to randomly presented audio-clips of three categories: synthesized (Watson, IBM), humanoid (robot Sophia, Hanson Robotics), and human voices (five clips/category). Voices were rated on intelligibility, prosody, trustworthiness, confidence, enthusiasm, pleasantness, human-likeness, likability, and naturalness. Speakers were rated on appeal, credibility, human-likeness, and eeriness. Participants' personality traits, attitudes to robots, and demographics were obtained.Results: The human voice and human speaker characteristics received reliably higher scores on all dimensions except for eeriness. Synthesized voice ratings were positively related to participants' agreeableness and neuroticism. Females rated synthesized voices more positively on most dimensions. Surprisingly, interest in social robots and attitudes toward robots played almost no role in voice evaluation. Contrary to the expectations of an uncanny valley, when the ratings of human-likeness for both the voice and the speaker characteristics were higher, they seemed less eerie to the participants. Moreover, when the speaker's voice was more humanlike, it was more liked by the participants. This latter point was only applicable to one of the synthesized voices. Finally, pleasantness and trustworthiness of the synthesized voice predicted the likability of the speaker's voice. Qualitative content analysis identified intonation, sound, emotion, and imageability/embodiment as diagnostic features.Discussion: Humans clearly prefer human voices, but manipulating diagnostic speech features might increase acceptance of synthesized voices and thereby support human-robot interaction. There is limited evidence that human-likeness of a voice is negatively linked to the perceived eeriness of the speaker.

Publisher

Frontiers Media SA

Subject

Artificial Intelligence,Biomedical Engineering

Reference97 articles.

1. My robot is happy today: how older people with mild cognitive impairments understand assistive robots' affective output,;Antona,2019

2. The right kind of unnatural: designing a robot voice,;Aylett,2019

3. Creating robot personality: effects of mixing speech and semantic free utterances,;Aylett,2020

4. Speech synthesis for the generation of artificial personality;Aylett;IEEE Trans. Affect. Comput,2017

5. The perception and analysis of the likeability and human likeness of synthesized speech,;Baird,2018

Cited by 54 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. The effectiveness of human vs. AI voice-over in short video advertisements: A cognitive load theory perspective;Journal of Retailing and Consumer Services;2024-11

2. Integration of ChatGPT Into a Course for Medical Students: Explorative Study on Teaching Scenarios, Students’ Perception, and Applications;JMIR Medical Education;2024-08-22

3. Toward a Third-Kind Voice for Conversational Agents in an Era of Blurring Boundaries Between Machine and Human Sounds;ACM Conversational User Interfaces 2024;2024-07-08

4. Do Your Expectations Match? A Mixed-Methods Study on the Association Between a Robot's Voice and Appearance;ACM Conversational User Interfaces 2024;2024-07-08

5. Users’ responses to humanoid social robots: A social response view;Telematics and Informatics;2024-07