Training data clustering for improved speech recognition

  title={Training data clustering for improved speech recognition},
  author={Ananth Sankar and Françoise Beaufays and Vassilios Digalakis},
We present an approach to cluster the training data for automatic speech recognition (ASR). A relative-entropy based distance metric between training data clusters is deened. This metric is used to hierarchically cluster the training data. The metric can also be used to select the closest training data clusters given a small amount of data from the test speaker. The selected clusters are then used to estimate a set of hidden Markov models (HMMs) for recognizing the speech from the test speaker… CONTINUE READING
Highly Cited
This paper has 40 citations. REVIEW CITATIONS