科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Computers in Human Behavior Artificial Humans2025-10-04· Psychology

Attractive synthetic voices

Camila Bruder, Pamela Breda, Pauline Larrouy-Maestri

原始摘要(英文原文)· Original abstract
With recent advances in Artificial Intelligence (AI), synthetic voices have become increasingly prevalent in our everyday soundscape. This study examined listeners’ perception of human and neural Text-To-Speech (TTS) voices. In an online experiment, 75 participants listened to different versions of a short utterance spoken by eight different voices (half human, half TTS), each presented in four expressed emotions (neutral, happy, sad, angry). For each stimulus, participants rated voice attractiveness and willingness to interact, and selected the perceived emotion from a forced-choice list. In a second part, participants were asked to classify each voice as human or AI-generated. Results revealed that participants were often “fooled” by the TTS voices, misidentifying them as human. Voice ratings were influenced by the perceived emotion regardless of the voice type, with happy-sounding voices rated more positively than those perceived as sad or angry. However, TTS voices were rated as less attractive and socially appealing overall, though with large individual differences. These findings indicate that TTS voices are approaching human ones in how they are perceived by listeners, highlighting progress in their naturalness. • Human voices are still rated as more attractive than synthetic ones – but perhaps not for long. • Participants are often fooled by TTS voices, misidentifying them as human. • Perceived emotion influences ratings similarly for human and TTS voices, with strong effect of valence. • Older adults are less accurate in distinguishing TTS from human voices.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Attractive synthetic voices — 科研速览 Science Skim