On the segmentation of touching characters
暂无分享,去创建一个
In many OCR systems, character segmentation is a necessary preprocessing step for character recognition. It is a critical step because incorrectly segmented characters are not likely to be correctly recognized. The most difficult cases in character segmentation are broken characters and touching characters. The problem of segmenting touching characters in various fonts and size in machine-printed text is addressed. The author classifies the touching characters into five categories: touching characters in fixed-pitch fonts, proportional and serif fonts, ambiguous touching characters, and strings with broken and touching characters. Different methods for detecting multiple character segments and for segmenting touching characters in these categories are developed. The methods use features of characters and fonts and profile models.<<ETX>>
[1] Haruo Asada,et al. Resolving Ambiguity in Segmenting Touching Characters , 1992 .
[2] Roy L. Hoffman,et al. Segmentation Methods for Recognition of Machine-Printed Characters , 1971, IBM J. Res. Dev..
[3] Theodosios Pavlidis,et al. On the Recognition of Printed Characters of Any Font and Size , 1987, IEEE Transactions on Pattern Analysis and Machine Intelligence.