Two streams deep neural network for handwriting word recognition

被引:0
作者
Alaa Sulaiman
Khairuddin Omar
Mohammad F. Nasrudin
机构
[1] Faculty of Information Science and Technology University Kebangsaan Malaysia (UKM),Pattern Recognition Research Group, Center of Artificial Intelligence Technology
来源
Multimedia Tools and Applications | 2021年 / 80卷
关键词
Handwritten word recognition; Deep model; Convolutional layer; ConvLSTM;
D O I
暂无
中图分类号
学科分类号
摘要
Handwritten word recognition is one of the hot topics in automatic handwritten text recognition that received a lot of attention in recent years. Unlike character recognition, word recognition deals with considerable variations in word shape and written style. This paper proposes a novel deep model for language-independent handwritten word recognition. The proposed deep structure has two parallel stages for jointly learning character and word-level information. In the character-level stage, a weakly character segmentation method is performed and then applies a series of Long short-term memory (LSTM) layers for character-level representation. The word-level stage employs a series of convolutional layers for the shape and structure representation of the word. These representations are then concatenated and followed by a series of fully connected layers for jointly learning the words and the character-level information. Since the character segmentation is language independent and error-prone, the proposed deep structure only applies weakly separation scheme and does not rely on any character segmentation algorithm. Thus, it effectively utilizes character level representation without bounding on any language model. In the proposed methodology, we use two new data augmentation strategies based on a psychological assumption to increase the model generalization performance. Experimental results on five public datasets including Arabic, English and German languages demonstrate that the proposed deep model has a superior performance to the state-of-the-art methods.
引用
收藏
页码:5473 / 5494
页数:21
相关论文
共 73 条
[1]  
AlKhateeb JHY(2011)Offline handwritten arabic cursive text recognition using hidden markov models and re-ranking Pattern Recogn Lett 32 1081-1088
[2]  
Ren J(2014)Word spotting and recognition with embedded attributes IEEE Transactions on Pattern Analysis and Machine Intelligence 36 2552-2566
[3]  
Jiang J(2019)Script identification in natural scene image and video frames using an attention based convolutional-lstm network Pattern Recogn 85 172-184
[4]  
Al-Muhtaseb H(2011)Improving offline handwritten text recognition with hybrid hmm/ann models IEEE Transactions on Pattern Analysis and Machine Intelligence 33 767-779
[5]  
Almazán J(2002)Automatic recognition of handwritten numerical strings: a recognition and verification strategy IEEE Trans Pattern Anal Mach Intell 24 1438-1454
[6]  
Gordo A(2019)A novel variational model for noise robust document image binarization Neurocomputing 325 288-302
[7]  
Fornés A(2019)Deep adaptive learning for writer identification based on single handwritten word images Pattern Recogn 88 64-74
[8]  
Valveny E(1997)Long short-term memory Neural Comput 9 1735-1780
[9]  
Bhunia AK(1962)Visual pattern recognition by moment invariants IRE Trans Information Theory 8 179-187
[10]  
Konwer A(2015)Reading text in the wild with convolutional neural networks Int J Comput Vis 116 1-20