Title :
Variational Bayesian speaker diarization of meeting recordings
Author :
Valente, Fabio ; Motlicek, Petr ; Vijayasenan, Deepu
Author_Institution :
IDIAP Res. Inst., Martigny, Switzerland
Abstract :
This paper investigates the use of the Variational Bayesian (VB) framework for speaker diarization of meetings data extending previous related works on Broadcast News audio. VB learning aims at maximizing a bound, known as Free Energy, on the model marginal likelihood and allows joint model learning and model selection according to the same objective function. While the BIC is valid only in the asymptotic limit, the Free Energy is always a valid bound. The paper proposes the use of Free Energy as objective function in speaker diarization. It can be used to select dynamically without any supervision or tuning, elements that typically affect the diarization performance i.e. the inferred number of speakers, the size of the GMM and the initialization. The proposed approach is compared with a conventional state-of-the-art system on the RT06 evaluation data for meeting recordings diarization and shows an improvement of 8.4% relative in terms of speaker error.
Keywords :
Bayes methods; Gaussian processes; speaker recognition; Free Energy; GMM; VB framework; broadcast news audio; meeting recordings; model marginal likelihood; variational Bayesian speaker diarization; Audio recording; Bayesian methods; Broadcasting; Density estimation robust algorithm; Error analysis; Microphones; Probability; Meetings Data; Speaker Diarization; Variational Bayesian Methods;
Conference_Titel :
Acoustics Speech and Signal Processing (ICASSP), 2010 IEEE International Conference on
Conference_Location :
Dallas, TX
Print_ISBN :
978-1-4244-4295-9
Electronic_ISBN :
1520-6149
DOI :
10.1109/ICASSP.2010.5495087