PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 7, 2004112 citations

Voice signatures

View Full Paper
ISIzhak ShafranMRMichael RileyMMMehryar Mohri

Key Points

Key points are not available for this paper at this time.

Abstract

Most current spoken-dialog systems only extract sequences of words from a speaker's voice. This largely ignores other useful information that can be inferred from speech such as gender, age, dialect, or emotion. These characteristics of a speaker's voice, voice signatures, whether static or dynamic, can be useful for speech mining applications or for the design of a natural spoken-dialog system. This paper explores the problem of extracting automatically and accurately voice signatures from a speaker's voice. We investigate two approaches for extracting speaker traits: the first focuses on general acoustic and prosodic features, the second on the choice of words used by the speaker. In the first approach, we show that standard speech/nonspeech HMM, conditioned on speaker traits and evaluated on cepstral and pitch features, achieve accuracies well above chance for all examined traits. The second approach, using support vector machines with rational kernels applied to speech recognition lattices, attains an accuracy of about 8.1 % in the task of binary classification of emotion. Our results are based on a corpus of speech data collected from a deployed customer-care application (HMIHY 0300). While still preliminary, our results are significant and show that voice signatures are of practical interest in real-world applications.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Shafran et al. (2004) studied this question.

synapsesocial.com/papers/6a1eec15aba71003cac8e07bhttps://doi.org/10.1109/asru.2003.1318399
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Speech recognition using fundamental frequency and voicing in acoustic modeling2002 · 23 citations
  2. 2Statistical dialect classification based on mean phonetic features2002 · 16 citations
  3. 3Accent identification2002 · 68 citations
  4. 4Recognition of negative emotions from the speech signal2005 · 149 citations
  5. 5Automated natural spoken dialog2002 · 61 citations