The Community for Technology Leaders
2013 12th International Conference on Document Analysis and Recognition (2011)
Beijing, China
Sept. 18, 2011 to Sept. 21, 2011
ISSN: 1520-5363
ISBN: 978-0-7695-4520-2
pp: 126-130
The current OCR cannot segment words and characters from video images due to complex background as well as low resolution of video images. To have better accuracy, this paper presents a new gradient based method for words and character segmentation from text line of any orientation in video frames for recognition. We propose a Max-Min clustering concept to obtain text cluster from the normalized absolute gradient feature matrix of the video text line image. Union of the text cluster with the output of Canny operation of the input video text line is proposed to restore missing text candidates. Then a run length algorithm is applied on the text candidate image for identifying word gaps. We propose a new idea for segmenting characters from the restored word image based on the fact that the text height difference at the character boundary column is smaller than that of the other columns of the word image. We have conducted experiments on a large dataset at two levels (word and character level) in terms of recall, precision and f-measure. Our experimental setup involves 3527 characters of English and Chinese, and this dataset is selected from TRECVID database of 2005 and 2006.
Video document analysis, Word segmentation, Video character extraction, Gradient features, Video character recognition
Umapada Pal, Bolan Su, Chew Lim Tan, Palaiahnakote Shivakumara, Souvik Bhowmick, "A New Gradient Based Character Segmentation Method for Video Text Recognition", 2013 12th International Conference on Document Analysis and Recognition, vol. 00, no. , pp. 126-130, 2011, doi:10.1109/ICDAR.2011.34
93 ms
(Ver )