Authors IndexSessionsTechnical programAttendees

 

Session: Other Topics in ASR Robustness, Adaptation and Language Modeling

Title: HIGH PERFORMANCE TELEPHONE BANDWIDTH SPEAKER INDEPENDENT CONTINUOUS DIGIT RECOGNITION

Authors: Piero Cosi, John-Paul Hosom, Alberto Valente

Abstract: The development of a high-performance telephone-bandwidth speaker independent connected digit recognizer for Italian is described. The CSLU Speech Toolkit was used to develop and implement the hybrid ANN/HMM system, which is trained on context-dependent categories to account for coarticulatory variation. Various front-end processing and system architecture were compared and, when the best features (MFCC with CMS + Delta) and network (4-layer fully connected feed-forward network) were considered, there was a 98.92% word recognition accuracy and a 92.62% sentence recognition accuracy) on a test set of the FIELD continuous digits recognition task.

a01pc006.ps a01pc006.pdf