Statistical dialect classification based on mean phonetic features
Generate an AI Snapshot to get a quick, structured summary of this paper.
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
Work done on a text-dependent method for automatic utterance classification and dialect model selection using mean cepstral and duration features on a per-phoneme basis and a linear discriminant to separate the dialects in feature space is described.
Abstract
Describes work done on a text-dependent method for automatic utterance classification and dialect model selection using mean cepstral and duration features on a per-phoneme basis. From transcribed dialect data, we build a linear discriminant to separate the dialects in feature space. This method is potentially much faster than our previous selection algorithm. We have been able to achieve error rates of 8% for distinguishing Northern US speakers from Southern US speakers, and average error rates of 13% on a variety of finer pairwise dialect discriminations. We also present a description of the training and test corpora collected for this work.
