• DocumentCode
    3775958
  • Title

    Linear multimodal fusion in video concept analysis based on node equilibrium model

  • Author

    Jie Geng;Zhenjiang Miao;Qinghua Liang;Shu Wang

  • Author_Institution
    Institute of Information Science, Beijing Jiaotong University, No.3 Shangyuancun, Haidian District, Beijing, China, 100044
  • fYear
    2015
  • Firstpage
    316
  • Lastpage
    320
  • Abstract
    Multiple modalities such as color, texture, shape and motion need to be analyzed separately and fused together to get the comprehensive result in content-based video concept analysis. We propose a multimodal fusion method based on a mechanical node equilibrium model. It treats the scores ofmultiple modalities and the fused score as physical nodes. Between these nodes, we define correlations which are treated as forces to move the nodes to a new position. Finally, the whole node system will be at an equilibrium status which is regarded as the fusion result. Essentially, the proposed method is a linear fusion model with linear fusion equations. The correlations are optimized by an expectation maximum (EM) algorithm which is quite efficient needing only several iterations.
  • Keywords
    "Correlation","Detectors","Training","Computational modeling","Mathematical model","Force","Linear programming"
  • Publisher
    ieee
  • Conference_Titel
    Pattern Recognition (ACPR), 2015 3rd IAPR Asian Conference on
  • Electronic_ISBN
    2327-0985
  • Type

    conf

  • DOI
    10.1109/ACPR.2015.7486517
  • Filename
    7486517