مرکز منطقه ای اطلاع رساني علوم و فناوري - Speaker trait characterization in web videos: Uniting speech, language, and facial features

DocumentCode :

1668569

Title :

Speaker trait characterization in web videos: Uniting speech, language, and facial features

Author :

Weninger, Felix ; Wagner, Christoph ; Wollmer, Martin ; Schuller, Bjorn ; Morency, Louis-Philippe

Author_Institution :

Machine Intell. & Signal Process. Group, Tech. Univ. Munchen, München, Germany

fYear :

2013

Firstpage :

3647

Lastpage :

3651

Abstract :

We present a multi-modal approach to speaker characterization using acoustic, visual and linguistic features. Full realism is provided by evaluation on a database of real-life web videos and automatic feature extraction including face and eye detection, and automatic speech recognition. Different segmentations are evaluated for the audio and video streams, and the statistical relevance of Linguistic Inquiry and Word Count (LIWC) features is confirmed. In the result, late multimodal fusion delivers 73, 92 and 73% average recall in binary age, gender and race classification on unseen test subjects, outperforming the best single modalities for age and race.

Keywords :

face recognition; feature extraction; speaker recognition; video signal processing; acoustic feature; audio stream; automatic feature extraction; automatic speech recognition; binary age; eye detection; face detection; facial feature; gender; language feature; linguistic feature; linguistic inquiry-word count feature; multimodal approach; multimodal fusion; race classification; real-life Web video; speaker characterization; speaker trait characterization; speech feature; video stream; visual feature; Acoustics; Face; Feature extraction; Pragmatics; Speech; Videos; Visualization; computational paralinguistics; multi-modal fusion; speaker classification;

fLanguage :

English

Publisher :

ieee

Conference_Titel :

Acoustics, Speech and Signal Processing (ICASSP), 2013 IEEE International Conference on

Conference_Location :

Vancouver, BC

ISSN :

1520-6149

Type :

conf

DOI :

10.1109/ICASSP.2013.6638338

Filename :

6638338

Link To Document :

https://search.ricest.ac.ir/dl/search/defaultta.aspx?DTC=49&DC=1668569