Speech amplitude, low-pass (below 1 kHz) and high-pass (above 1 kHz) zero-crossing rates were represented as weighted sums of orthonormal functions. The weighting coefficients were used to identify unknown speakers. For a ten speaker population recognition rates as high as 96.6 percent were obtained using the same data for training and testing.
No takes yet. Share an insight, caveat, or question.
Wasson et al. (1975) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: