An isolated word recognizer has been evaluated using a large data base of telephone-band digit utterances recorded by 100 talkers. Three reference template configurations of the recognizer have been studied, one talker independent and two talker dependent. The talker dependent configurations are a 1 template per word system obtained by robust training and a 5 template per word system. Performance has been studied in terms of the distribution of error rates across the three template configurations, the effect on performance of varying a decision rule parameter, applying a rejection threshold, and normalizing test and reference utterance lengths has also been investigated. Overall average error rates obtained are 2.17% for the talker independent system and 2.77% and 0.77% for the 1-template and 5-template talker dependent systems, respectively.
No takes yet. Share an insight, caveat, or question.
Rosenberg et al. (2005) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: