Title :
A robust algorithm for text string separation from mixed text/graphics images
Author :
Fletcher, Lloyd Alan ; Kasturi, Rangachar
Author_Institution :
Dept. of Electr. Eng., Pennsylvania State Univ., University Park, PA, USA
fDate :
11/1/1988 12:00:00 AM
Abstract :
The development and implementation of an algorithm for automated text string separation that is relatively independent of changes in text font style and size and of string orientation are described. It is intended for use in an automated system for document analysis. The principal parts of the algorithm are the generation of connected components and the application of the Hough transform in order to group components into logical character strings that can then be separated from the graphics. The algorithm outputs two images, one containing text strings and the other graphics. These images can then be processed by suitable character recognition and graphics recognition systems. The performance of the algorithm, both in terms of its effectiveness and computational efficiency, was evaluated using several test images and showed superior performance compared to other techniques
Keywords :
computer graphics; computerised pattern recognition; computerised picture processing; transforms; Hough transform; character recognition; computer graphics; computerized picture processing; document analysis; graphics recognition; mixed text/graphics images; text string separation; Character generation; Character recognition; Graphics; Image analysis; Image processing; Image recognition; Image segmentation; Robustness; Testing; Text analysis;
Journal_Title :
Pattern Analysis and Machine Intelligence, IEEE Transactions on