Logo image
適用於華英雙語語音辨識之聲學單位合併方法
Thesis

適用於華英雙語語音辨識之聲學單位合併方法

卓楷斌
Masters, 國立清華大學, 資訊系統與應用研究所
2011

Abstract

華英雙語辨識系統 聲學模型合併 華英雙語問題集 Mandarin-English bilingual recognition system mergence of bilingual acoustic models Mandarin-English bilingual question sets
The long-term goal of this research is to construct a Mandarin-English bilingual speech recognition system on devices mounted on automobiles with limited storage size. Thus, the purpose of this thesis is to effectively reduce the model size and to maintain considerable performance as a unilingual system without using language identification. In this thesis, similar acoustic models are merged to reduce the number of model parameters. Similar acoustic units between the two languages are found by analyzing different phonetic notations with either knowledge-driven or data-driven techniques. In addition to directly merging the two acoustic models, this thesis also proposes the use of decision trees to merge states of different HMMs (hidden Markov models). Experimental result shows that, merging the models in a finer level via decision trees not only effectively reduces the model size but also enhances robustness of the bilingual models. By comparing to the baseline models, the state mergence using decision trees can reduce model size to one third of the original one and achieve an improvement of 1.2% in correction rate of bilingual recognition.

Metrics

1 Record Views

Details

Logo image