A new watershed model based system for character segmentation in degraded text lines

详细信息查看全文

作者：Aladhahalli Shivegowda Kavitha^a ; ^{kavitha_sanjay_as@yahoo.co.in} ; Palaiahnakote Shivakumara^b ; ^{shiva@um.edu.my} ; Govindaraj Hemantha Kumar^a ; ^{ghk.2007@yahoo.com} ; Tong Lu^c ; ^{lutong@nju.edu.cn}
关键词：Degraded historical document image ; Filtering ; Watershed model ; Gradient values ; Character segmentation
刊名：AEU - International Journal of Electronics and Communications
出版年：2017
出版时间：January 2017
年：2017
卷：71
期：Complete
页码：45-52
全文大小：2957 K
卷排序：71

文摘

Character segmentation from text lines in degraded historical document images is challenging due to complex background and non-availability of regular structures of text patterns. This paper proposes a new method based on watershed model for segmenting characters from text lines in degraded historical document images. The proposed method filters out noise pixels by exploring Sobel and Laplacian values of pixels, which results in edges that represent text components. We then propose watershed model for studying non-linear spacing between characters based on the fact that watersheds provide information about water flow and volume of collection of water. Experimental results on different datasets, which include degraded historical document images called Indus documents and other Indian scripts, show that the proposed method segments characters better than the existing character segmentation methods in terms of recall and precision.

地址：北京市海淀区学院路29号邮编：100083

电话：办公室：(+86 10)66554848；文献借阅、咨询服务、科技查新：66554700