login

Language independent and language adaptive large vocabulary speech recognition

Published 30 November 1998Open access
Tanja Schultz, Alex Waibel
Citations63
View PDF

TL;DR

This paper describes the design of a multilingual speech recognizer using an LVCSR dictation database which has been collected under the project GlobalPhone and presents several recognition results in language independent and language adaptive setups.

Abstract

This paper describes the design of a multilingual speech recognizer using an LVCSR dictation database which has been collected under the project GlobalPhone.This project at the University of Karlsruhe investigates LVCSR systems in 15 languages of the world, namely Arabic, Chinese, Croatian, English, French, German, Italian, Japanese, Korean, Portuguese, Russian, Spanish, Swedish, Tamil, and Turkish.Based on a global phoneme set we built different multilingual speech recognition systems for five of the 15 languages.Context dependent phoneme models are created data-driven by introducing questions about language and language groups to our polyphone clustering procedure.We apply the resulting multilingual models to unseen languages and present several recognition results in language independent and language adaptive setups.

Keywords

Computer Science