PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 26, 201034 citations

SCARF: a segmental conditional random field toolkit for speech recognition

View Full Paper
GZGeoffrey ZweigPNPatrick Nguyen

Key Points

Key points are not available for this paper at this time.

Abstract

This paper describes a new toolkit- SCARF- for doing speech recognition with segmental conditional random fields. It is designed to allow for the integration of numerous, possibly redundant segment level acoustic features, along with a complete language model, in a coherent speech recognition framework. SCARF performs a segmental analysis, where each segment corresponds to a word, thus allowing for the incorporation of acoustic features defined at the phoneme, multi-phone, syllable and word level. SCARF is designed to make it especially convenient to use acoustic detection events as input, such as the detection of energy bursts, phonemes, or other events. Language modeling is done by associating each state in the SCRF with a state in an underlying n-gram language model, and SCARF supports the joint and discriminative training of language model and acoustic model parameters. SCARF is available for download from

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zweig et al. (2010) studied this question.

synapsesocial.com/papers/6a15623acb801b7f954e70c0https://doi.org/10.21437/interspeech.2010-307
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1From knowledge-ignorant to knowledge-rich modeling : a new speech research parading for next generation automatic speech recognition2004 · 57 citations
  2. 2Advances in speech transcription at IBM under the DARPA EARS program2006 · 125 citations
  3. 3Conditional Random Fields: Probabilistic Models for Segmenting and Labeling Sequence Data2001 · 12,994 citations
  4. 4From knowledge-ignorant to knowledge-rich modeling : A new speech research paradigm for next generation automatic speech recognition2004 · 70 citations
  5. 5Advances in transcription of broadcast news and conversational telephone speech within the combined EARS BBN/LIMSI system2006 · 49 citations