• DocumentCode
    1798999
  • Title

    Scene text recognition in multiple frames based on text tracking

  • Author

    Xuejian Rong ; Chucai Yi ; Xiaodong Yang ; YingLi Tian

  • fYear
    2014
  • fDate
    14-18 July 2014
  • Firstpage
    1
  • Lastpage
    6
  • Abstract
    Text signage as visual indicators in natural scene plays an important role in navigation and notification in our daily life. Most previous methods of scene text extraction are developed from a single scene image. In this paper, we propose a multi-frame based scene text recognition method by tracking text regions in a video captured by a moving camera. The main contributions of this paper are as follows. First, we present a framework of scene text recognition in multiple frames based on feature representation of scene text character (STC) for character prediction and conditional random field (CRF) model for word configuration. Second, a feature representation of STC is employed from dense sampled SIFT descriptors and Fisher Vector. Third, we collect a dataset for text information extraction from natural scene videos. Our proposed multi-frame scene text recognition is more compatible with image/video-based mobile applications. The experimental results demonstrate that STC prediction and word configuration in multiple frames based on text tracking significantly improves the performance of scene text recognition.
  • Keywords
    text detection; Fisher vector; SIFT descriptors; conditional random field model; feature representation; image-based mobile applications; multiframe based scene text recognition method; multiframe scene text recognition; multiple frames; natural scene videos; scene text character; scene text recognition; text extraction; text information extraction; text signage; text tracking; tracking text regions; video-based mobile applications; Accuracy; Encoding; Feature extraction; Predictive models; Text recognition; Trajectory; Vectors; Fisher Vector; Scene text recognition; feature representation of scene text character; text tracking; video dataset of scene text;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Multimedia and Expo (ICME), 2014 IEEE International Conference on
  • Conference_Location
    Chengdu
  • Type

    conf

  • DOI
    10.1109/ICME.2014.6890248
  • Filename
    6890248