Ninth International Conference on Document Analysis and Recognition (ICDAR 2007) Vol 1
HMM-Based Recognizer with Segmentation-free Strategy for Unconstrained Chinese Handwritten Text
Curitiba, Parana, Brazil
September 23-September 26
ISBN: 0-7695-2822-8
T.-H. Su, Harbin Institute of Technology, Harbin, P. R. CHINA
T.-W. Zhang, Harbin Institute of Technology, Harbin, P. R. CHINA
H.-J. Huang, Harbin Institute of Technology, Harbin, P. R. CHINA
Y. Zhou, Harbin Institute of Technology, Harbin, P. R. CHINA
A segmentation-free strategy based on Hidden Markov Models (HMMs) is presented for offline recognition of unconstrained Chinese handwriting. As the first step, handwritten textlines are converted to observation sequence by sliding windows and character segmentation stage is avoided prior to recognition. Following that, embedded Baum-Welch algorithm is adopted to train character HMMs. Finally, best character string maximizing the a posteriori is located through Viterbi algorithm. Experiments are conducted on the HIT-MW database written by more than 780 writers. The results show: First, our baseline recognizer outperforms one segmentation-based OCR product with 35% relative improvement; second, more discriminative feature and compact representation, and state-tying technique to alleviate the data sparsity can enhance the recognizer with high confidence. The final recognizer has improved the performance by 10.77% than the baseline system.
Citation:
T.-H. Su, T.-W. Zhang, H.-J. Huang, Y. Zhou, "HMM-Based Recognizer with Segmentation-free Strategy for Unconstrained Chinese Handwritten Text," icdar, vol. 1, pp.133-137, Ninth International Conference on Document Analysis and Recognition (ICDAR 2007) Vol 1, 2007