مرکز منطقه ای اطلاع رساني علوم و فناوري - Using a Free-Parts Representation for Visual Speech Recognition

DocumentCode :

3167004

Title :

Using a Free-Parts Representation for Visual Speech Recognition

Author :

Lucey, Patrick ; Lucey, Simon ; Sridharan, Sridha

Author_Institution :

Queensland University of Technology

fYear :

205

fDate :

6-8 Dec. 205

Firstpage :

Lastpage :

Abstract :

Motivated by the success of free-parts based representations in face recognition, we have attempted to address some of the problems associated with applying such a philosophy to the task of speaker-independent visual speech recognition. A major problem with canonical area-based approaches in automatic visual speech recognition is the dependence these approaches have on locating and tracking the speaker’s region of interest (ROI) correctly. By employing a free-parts representation,we assume that the position/structure of patches within the mouth image can be relaxed so they can "freely" move to varying extents, hence reducing the influence of the front-end effect. In this paper, we show that by using a free-parts representation we gain some robustness against the problem of ROI localisation and tracking compared to current area-based feature extraction techniques such as the discrete cosine transform (DCT). Also in this paper, we expose the importance of representation for the task of visual speech recognition highlighted by the poor results current representations yield.

Keywords :

Automatic speech recognition; Discrete cosine transforms; Face recognition; Feature extraction; Hidden Markov models; Laboratories; Mouth; Pixel; Robustness; Speech recognition;

fLanguage :

English

Publisher :

ieee

Conference_Titel :

Digital Image Computing: Techniques and Applications, 2005. DICTA '05. Proceedings 2005

Conference_Location :

Queensland, Australia

Print_ISBN :

0-7695-2467-2

Type :

conf

DOI :

10.1109/DICTA.2005.84

Filename :

1587657

Link To Document :

https://search.ricest.ac.ir/dl/search/defaultta.aspx?DTC=49&DC=3167004