The Human Takes It All : Humanlike Synthesized Voices Are Perceived as Less Eerie and More Likable. Evidence From a Subjective Ratings Study

Kühne, Katharina; Fischer, Martin H.; Zhou, Yuefang

doi:10.3389/fnbot.2020.593732

The search result changed since you submitted your search request. Documents might be displayed in a different sort order.

search hit 7 of 131

Back to Result List

The Human Takes It All

Katharina Kühne, Martin H. Fischer, Yuefang Zhou

Background: The increasing involvement of social robots in human lives raises the question as to how humans perceive social robots. Little is known about human perception of synthesized voices. Aim: To investigate which synthesized voice parameters predict the speaker's eeriness and voice likability; to determine if individual listener characteristics (e.g., personality, attitude toward robots, age) influence synthesized voice evaluations; and to explore which paralinguistic features subjectively distinguish humans from robots/artificial agents. Methods: 95 adults (62 females) listened to randomly presented audio-clips of three categories: synthesized (Watson, IBM), humanoid (robot Sophia, Hanson Robotics), and human voices (five clips/category). Voices were rated on intelligibility, prosody, trustworthiness, confidence, enthusiasm, pleasantness, human-likeness, likability, and naturalness. Speakers were rated on appeal, credibility, human-likeness, and eeriness. Participants' personality traits, attitudes to robots, andBackground: The increasing involvement of social robots in human lives raises the question as to how humans perceive social robots. Little is known about human perception of synthesized voices. Aim: To investigate which synthesized voice parameters predict the speaker's eeriness and voice likability; to determine if individual listener characteristics (e.g., personality, attitude toward robots, age) influence synthesized voice evaluations; and to explore which paralinguistic features subjectively distinguish humans from robots/artificial agents. Methods: 95 adults (62 females) listened to randomly presented audio-clips of three categories: synthesized (Watson, IBM), humanoid (robot Sophia, Hanson Robotics), and human voices (five clips/category). Voices were rated on intelligibility, prosody, trustworthiness, confidence, enthusiasm, pleasantness, human-likeness, likability, and naturalness. Speakers were rated on appeal, credibility, human-likeness, and eeriness. Participants' personality traits, attitudes to robots, and demographics were obtained. Results: The human voice and human speaker characteristics received reliably higher scores on all dimensions except for eeriness. Synthesized voice ratings were positively related to participants' agreeableness and neuroticism. Females rated synthesized voices more positively on most dimensions. Surprisingly, interest in social robots and attitudes toward robots played almost no role in voice evaluation. Contrary to the expectations of an uncanny valley, when the ratings of human-likeness for both the voice and the speaker characteristics were higher, they seemed less eerie to the participants. Moreover, when the speaker's voice was more humanlike, it was more liked by the participants. This latter point was only applicable to one of the synthesized voices. Finally, pleasantness and trustworthiness of the synthesized voice predicted the likability of the speaker's voice. Qualitative content analysis identified intonation, sound, emotion, and imageability/embodiment as diagnostic features. Discussion: Humans clearly prefer human voices, but manipulating diagnostic speech features might increase acceptance of synthesized voices and thereby support human-robot interaction. There is limited evidence that human-likeness of a voice is negatively linked to the perceived eeriness of the speaker.…

Metadaten
Author details:	Katharina Kühne ORCiD, Martin H. Fischer ORCiD GND, Yuefang Zhou ORCiD GND
DOI:	https://doi.org/10.3389/fnbot.2020.593732
ISSN:	1662-5218
Title of parent work (English):	Frontiers in Neurorobotics
Subtitle (English):	Humanlike Synthesized Voices Are Perceived as Less Eerie and More Likable. Evidence From a Subjective Ratings Study
Publisher:	Frontiers Research Foundation
Place of publishing:	Lausanne
Publication type:	Article
Language:	English
Date of first publication:	2020/12/16
Publication year:	2020
Release date:	2021/01/28
Tag:	human-robot interaction; paralinguistic features; synthesized voice; text-to-speech; uncanny valley
Volume:	14
Number of pages:	15
Funding institution:	Universität Potsdam
Funding number:	PA 2020_135
Organizational units:	Humanwissenschaftliche Fakultät / Strukturbereich Kognitionswissenschaften
DDC classification:	6 Technik, Medizin, angewandte Wissenschaften / 60 Technik / 600 Technik, Technologie
	6 Technik, Medizin, angewandte Wissenschaften / 61 Medizin und Gesundheit / 610 Medizin und Gesundheit
Peer review:	Referiert
Grantor:	Publikationsfonds der Universität Potsdam
Publishing method:	Open Access / Gold Open-Access
License (German):	CC-BY - Namensnennung 4.0 International
External remark:	Zweitveröffentlichung in der Schriftenreihe Postprints der Universität Potsdam : Humanwissenschaftliche Reihe ; 700

The Human Takes It All

Export metadata

Additional Services