DocumentCode
3775958
Title
Linear multimodal fusion in video concept analysis based on node equilibrium model
Author
Jie Geng;Zhenjiang Miao;Qinghua Liang;Shu Wang
Author_Institution
Institute of Information Science, Beijing Jiaotong University, No.3 Shangyuancun, Haidian District, Beijing, China, 100044
fYear
2015
Firstpage
316
Lastpage
320
Abstract
Multiple modalities such as color, texture, shape and motion need to be analyzed separately and fused together to get the comprehensive result in content-based video concept analysis. We propose a multimodal fusion method based on a mechanical node equilibrium model. It treats the scores ofmultiple modalities and the fused score as physical nodes. Between these nodes, we define correlations which are treated as forces to move the nodes to a new position. Finally, the whole node system will be at an equilibrium status which is regarded as the fusion result. Essentially, the proposed method is a linear fusion model with linear fusion equations. The correlations are optimized by an expectation maximum (EM) algorithm which is quite efficient needing only several iterations.
Keywords
"Correlation","Detectors","Training","Computational modeling","Mathematical model","Force","Linear programming"
Publisher
ieee
Conference_Titel
Pattern Recognition (ACPR), 2015 3rd IAPR Asian Conference on
Electronic_ISBN
2327-0985
Type
conf
DOI
10.1109/ACPR.2015.7486517
Filename
7486517
Link To Document