Sciweavers

ICASSP
2009
IEEE

Unsupervised speaker adaptation for telephone call transcription

13 years 2 months ago
Unsupervised speaker adaptation for telephone call transcription
The use of the PC and Internet for placing telephone calls will present new opportunities to capture vast amounts of un-transcribed speech for a particular speaker. This paper investigates how to best exploit this data for speaker-dependent speech recognition. Supervised and unsupervised experiments in acoustic model and language model adaptation are presented. Using one hour of automatically transcribed speech per speaker with a word error rate of 36.0%, unsupervised adaptation resulted in an absolute gain of 6.3%, equivalent to 70% of the gain from the supervised case, with additional adaptation data likely to yield further improvements. LM adaptation experiments suggested that although there seems to be a small degree of speaker idiolect, adaptation to the speaker alone, without considering the topic of the conversation, is in itself unlikely to improve transcription accuracy.
R. Wallace, Kishan Thambiratnam, Frank Seide
Added 18 Feb 2011
Updated 18 Feb 2011
Type Journal
Year 2009
Where ICASSP
Authors R. Wallace, Kishan Thambiratnam, Frank Seide
Comments (0)