Randomized trial investigates how speaker gender affects emotion perception in Korean listeners, suggesting modulation by gendered voice characteristics.
Speech conveys rich paralinguistic information, notably the speaker’s emotional state. The acoustic expression of emotion, however, is subject to considerable variability shaped by factors such as speaker and listener gender, as well as broader cultural and linguistic contexts. This study investigates how four emotions - Happy, Sad, Angry, and Anxious - are perceived by 33 native Korean listeners. The stimuli consisted of low-pass filtered emotional utterances from ‘The Open AI Dataset Project (AI-Hub)’, which allowed for a focus on prosodic cues while removing semantic content. Results showed that recognition accuracy varied by both emotion and speaker gender: Happy was most consistently identified across all voices, while Angry was more accurately recognized in male speech and Sad in female speech. Perceived emotional intensity also differed by speaker gender: female speakers received higher intensity ratings for Happy than male speakers. In particular, female speakers’ Happy and Sad were perceived as more intense than their own Angry . These results are discussed in light of perceptual weighting of acoustic cues in emotion recognition, suggesting that gendered voice characteristics modulate how listeners extract emotional meaning from prosody alone.
No takes yet. Share an insight, caveat, or question.
Yoon et al. (2026) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: