• DocumentCode
    3695207
  • Title

    Simplifying the reading of historical manuscripts

  • Author

    Abedelkadir Asi;Rafi Cohen;Klara Kedem;Jihad El-Sana

  • Author_Institution
    Department of Computer Science, Ben-Gurion University of the Negev, Beer-Sheva, Israel
  • fYear
    2015
  • Firstpage
    826
  • Lastpage
    830
  • Abstract
    Complex document layouts pose prominent challenges for document image understanding algorithms. These layouts impose irregularities on the location of text paragraphs which consequently induces difficulties in reading the text. In this paper we present a robust framework for analyzing historical manuscripts with complex layouts. This framework aims to provide a convenient reading experience for historians through topnotch algorithms for text localization, classification and dewarping. We segment text into spatially coherent regions and text-lines using texture-based filters and refine this segmentation by exploiting Markov Random Fields (MRFs). A principled technique is presented for dewarping curvy text regions using a non-linear geometric transformation. The framework has been validated using a subset of a publicly available dataset of historical documents and it provided promising results.
  • Keywords
    "Kernel","Transforms","Image segmentation"
  • Publisher
    ieee
  • Conference_Titel
    Document Analysis and Recognition (ICDAR), 2015 13th International Conference on
  • Type

    conf

  • DOI
    10.1109/ICDAR.2015.7333877
  • Filename
    7333877