
In this speech recognition system, a set of templates for each phoneme includes clusters of speech patterns based on two speech features: "physical" features (formant spectra of men versus women) and "utterance" features (unvoiced vowels and nasalization), derived from a plurality of reference speakers.











