Combining Acoustic Data Driven G2P and Letter-to-Sound Rules for Under Resource Lexicon Generation

Type of publication:	Conference paper
Citation:	Rasipuram_INTERSPEECH_2012
Publication status:	Accepted
Booktitle:	Proceedings of Interspeech
Year:	2012
Location:	Portland, Oregon
Abstract:	In a recent work, we proposed an acoustic data-driven grapheme-to-phoneme (G2P) conversion approach, where the probabilistic relationship between graphemes and phonemes learned through acoustic data is used along with the orthographic transcription of words to infer the phoneme sequence. In this paper, we extend our studies to under-resourced lexicon development problem. More precisely, given a small amount of transcribed speech data consisting of few words along with its pronunciation lexicon, the goal is to build a pronunciation lexicon for unseen words. In this framework, we compare our G2P approach with standard letter-to-sound (L2S) rule based conversion approach. We evaluated the generated lexicons on PhoneBook 600 words task in terms of pronunciation errors and ASR performance. The G2P approach yields a best ASR performance of 14.0% word error rate (WER), while L2S approach yields a best ASR performance of 13.7% WER. A combination of G2P approach and L2S approach yields a best ASR performance of 9.3% WER.
Keywords:	grapheme, grapheme-to-phoneme converter, letter-to-sound rules, Lexicon, multilayer perceptron, phoneme
Projects:	Idiap
Authors:	Rasipuram, Ramya Magimai-Doss, Mathew
Added by:	[UNK]
Total mark:	0
Attachments
Rasipuram_INTERSPEECH_2012.pdf
Notes

processing time: 0.0010 seconds.