DocumentCode :
1456343
Title :
A robust algorithm for text string separation from mixed text/graphics images
Author :
Fletcher, Lloyd Alan ; Kasturi, Rangachar
Author_Institution :
Dept. of Electr. Eng., Pennsylvania State Univ., University Park, PA, USA
Volume :
10
Issue :
6
fYear :
1988
fDate :
11/1/1988 12:00:00 AM
Firstpage :
910
Lastpage :
918
Abstract :
The development and implementation of an algorithm for automated text string separation that is relatively independent of changes in text font style and size and of string orientation are described. It is intended for use in an automated system for document analysis. The principal parts of the algorithm are the generation of connected components and the application of the Hough transform in order to group components into logical character strings that can then be separated from the graphics. The algorithm outputs two images, one containing text strings and the other graphics. These images can then be processed by suitable character recognition and graphics recognition systems. The performance of the algorithm, both in terms of its effectiveness and computational efficiency, was evaluated using several test images and showed superior performance compared to other techniques
Keywords :
computer graphics; computerised pattern recognition; computerised picture processing; transforms; Hough transform; character recognition; computer graphics; computerized picture processing; document analysis; graphics recognition; mixed text/graphics images; text string separation; Character generation; Character recognition; Graphics; Image analysis; Image processing; Image recognition; Image segmentation; Robustness; Testing; Text analysis;
fLanguage :
English
Journal_Title :
Pattern Analysis and Machine Intelligence, IEEE Transactions on
Publisher :
ieee
ISSN :
0162-8828
Type :
jour
DOI :
10.1109/34.9112
Filename :
9112
Link To Document :
بازگشت