• DocumentCode
    2598436
  • Title

    Learning video primitives from natural video sequence

  • Author

    Li, Chuanzhen ; Zhang, Qin

  • Author_Institution
    Inf. & Eng. Sch., Commun. Univ. of China, Beijing, China
  • Volume
    1
  • fYear
    2011
  • fDate
    15-17 Oct. 2011
  • Firstpage
    492
  • Lastpage
    496
  • Abstract
    Block based video modeling is a hot issue of video information processing. In past literature, the size of block is set to a fixed value. Different with previous works, we find that the optimal size of primitives depend on the video content rather than fixed value. In this paper, in order to model natural video sequence, we segment video sequence to a number of spatial-temporal neighborhoods, and categorize video neighborhoods into two types: structural video primitives and textural video primitives. Structural video primitives represent structural pixels and their movement and textural video primitives represent the texture neighborhoods and their movement. We learn the size of video primitives based on genetic algorithm and spatial-temporal neighborhoods entropy. Then we map spatial-temporal neighborhoods to primitives using the structural similarity index. The experimental results demonstrate that the size of primitives depends on the content of the video rather than a fixed value. Using our method, the structural video primitives and textural video primitives are separated better than using fixed size, and the computational time for learning primitives has been greatly reduced. The primitives we learned can be used to video reconstruction, video segmentation and other applications.
  • Keywords
    computational complexity; genetic algorithms; image reconstruction; image segmentation; image sequences; image texture; learning (artificial intelligence); video signal processing; block based video modeling; computational time; fixed size; genetic algorithm; learning primitives; natural video sequence; spatial-temporal neighborhoods entropy; structural pixels; structural similarity index; structural video primitives; textural video primitives; video content; video information processing; video neighborhoods; video reconstruction; video segmentation; Computational modeling; Entropy; Genetic algorithms; Humans; Indexes; Manifolds; Video sequences; genetic algorithm; spatial-temporal neighborhoods; structural video primitives; textural video primitives;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Image and Signal Processing (CISP), 2011 4th International Congress on
  • Conference_Location
    Shanghai
  • Print_ISBN
    978-1-4244-9304-3
  • Type

    conf

  • DOI
    10.1109/CISP.2011.6099969
  • Filename
    6099969