Sciweavers

TASLP
2010

Speaker Diarization Exploiting the Eigengap Criterion and Cluster Ensembles

12 years 11 months ago
Speaker Diarization Exploiting the Eigengap Criterion and Cluster Ensembles
A novel system for speaker diarization is proposed that combines the eigengap criterion and cluster ensembles. No explicit assumptions on the number of speakers are made. Two variants of the system are developed. The first variant does not cluster the speech segments that are detected as outliers, while the second one does. The aforementioned system variants are assessed with respect to various metrics, such as the overall classification error, the average cluster purity, and the average speaker purity. Experiments are conducted on twoperson dialogue scenes in movies as well as on news broadcasts from MDE RT-03 Training Data Speech Corpus released by the U.S. National Institute of Standards and Technology. In the latter case, the diarization error rate is also reported. It is demonstrated that the clustering performance does not degrade when outliers are present. Moreover, thanks to the eigengap criterion, the evaluation metrics are improved.
Nikoletta Bassiou, Vassiliki Moschou, Constantine
Added 21 May 2011
Updated 21 May 2011
Type Journal
Year 2010
Where TASLP
Authors Nikoletta Bassiou, Vassiliki Moschou, Constantine Kotropoulos
Comments (0)