Offline Urdu Nastaleeq Optical Character Recognition Based on Stacked Denoising Autoencoder

被引:0
作者
Ahmad, Ibrar [1 ,2 ]
Wang, Xiaojie [1 ]
Li, Ruifan [1 ]
Rasheed, Shahid [3 ]
机构
[1] Beijing Univ Posts & Telecommun, Sch Comp Sci, CIST, 10 Xitucheng Rd, Beijing 100876, Peoples R China
[2] Univ Peshawar, Dept Comp Sci, Peshawar 25120, Pakistan
[3] PTCL, Islamabad 44000, Pakistan
基金
中国国家自然科学基金;
关键词
offline printed ligature recognition; urdu nastaleeq; denoising autoencoder; deep learning; classification;
D O I
暂无
中图分类号
TN [电子技术、通信技术];
学科分类号
0809 ;
摘要
Offline Urdu Nastaleeq text recognition has long been a serious problem due to its very cursive nature. In order to get rid of the character segmentation problems, many researchers are shifting focus towards segmentation free ligature based recognition approaches. Majority of the prevalent ligature based recognition systems heavily rely on hand-engineered feature extraction techniques. However, such techniques are more error prone and may often lead to a loss of useful information that might hardly be captured later by any manual features. Most of the prevalent Urdu Nastaleeq test recognition was trained and tested on small sets. This paper proposes the use of stacked denoising autoencoder for automatic feature extraction directly from raw pixel values of ligature images. Such deep learning networks have not been applied for the recognition of Urdu text thus far. Different stacked denoising autoencoders have been trained on 178573 ligatures with 3732 classes from un-degraded (noise free) UPTI (Urdu Printed Text Image) data set. Subsequently, trained networks are validated and tested on degraded versions of UPTI data set. The experimental results demonstrate accuracies in range of 93% to 96% which are better than the existing Urdu OCR systems for such large dataset of ligatures.
引用
收藏
页码:146 / 157
页数:12
相关论文
共 30 条
[1]  
Ahmad Z, 2007, PROC WRLD ACAD SCI E, V26, P249
[2]  
Akram Q. U. A., 2010, P GRAD C COMP SCI GC, V1
[3]  
[Anonymous], 2006, NIPS
[4]  
[Anonymous], 2010, P INT C INF EM TECHN
[5]  
[Anonymous], 2002, MULT TOP C 2002 INMI, DOI DOI 10.1109/INMIC.2002.1310191
[6]  
[Anonymous], 2010, JWAoS Engineering
[7]  
[Anonymous], P C LANG TECHN CLT 1
[8]  
[Anonymous], 2012, INT J COMPUTER APPL
[9]  
[Anonymous], 2013, SEGMENTATION BASED U
[10]  
[Anonymous], 1992, Structured Document Image Analysis