离线孟加拉语手写文本识别:各种深度学习方法的综合研究

Farhan Sadaf, S. M. Taslim Uddin Raju, Abdul Muntakim
{"title":"离线孟加拉语手写文本识别:各种深度学习方法的综合研究","authors":"Farhan Sadaf, S. M. Taslim Uddin Raju, Abdul Muntakim","doi":"10.1109/ICEEE54059.2021.9718890","DOIUrl":null,"url":null,"abstract":"Offline Handwritten Text Recognition (HTR) is a technique for translating handwritten images into digitally editable text format. Due to the presence of cursive letters, punctuation marks, and compound characters, it is more complex to recognize Bangla handwritten text. Over the years, several approaches to the optical model of the HTR system have been developed, including Hidden Markov Model (HMM) or deep learning techniques such as Convolutional Recurrent Neural Networks (CRNN), and current state-of-the-art Gated-CNN based architectures. Despite this, there are relatively limited works available for Bangla word recognition. In this paper, we introduce an end-to-end system for Bangla word recognition. We used a variety of popular pre-trained CNN architectures, including Xception, MobileNet, and DenseNet, followed by recurrent units such as LSTM or GRU. Furthermore, we experimented with Puigcerver’s CRNN based and Flor’s Gated-CNN based optical model architectulimited works available in Bangla.res. Flor architecture provided the highest recognition rate in our experiment, with a CER of 12.83% and a WER of 36.01%.","PeriodicalId":188366,"journal":{"name":"2021 3rd International Conference on Electrical & Electronic Engineering (ICEEE)","volume":"77 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2021-12-22","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"3","resultStr":"{\"title\":\"Offline Bangla Handwritten Text Recognition: A Comprehensive Study of Various Deep Learning Approaches\",\"authors\":\"Farhan Sadaf, S. M. Taslim Uddin Raju, Abdul Muntakim\",\"doi\":\"10.1109/ICEEE54059.2021.9718890\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Offline Handwritten Text Recognition (HTR) is a technique for translating handwritten images into digitally editable text format. Due to the presence of cursive letters, punctuation marks, and compound characters, it is more complex to recognize Bangla handwritten text. Over the years, several approaches to the optical model of the HTR system have been developed, including Hidden Markov Model (HMM) or deep learning techniques such as Convolutional Recurrent Neural Networks (CRNN), and current state-of-the-art Gated-CNN based architectures. Despite this, there are relatively limited works available for Bangla word recognition. In this paper, we introduce an end-to-end system for Bangla word recognition. We used a variety of popular pre-trained CNN architectures, including Xception, MobileNet, and DenseNet, followed by recurrent units such as LSTM or GRU. Furthermore, we experimented with Puigcerver’s CRNN based and Flor’s Gated-CNN based optical model architectulimited works available in Bangla.res. Flor architecture provided the highest recognition rate in our experiment, with a CER of 12.83% and a WER of 36.01%.\",\"PeriodicalId\":188366,\"journal\":{\"name\":\"2021 3rd International Conference on Electrical & Electronic Engineering (ICEEE)\",\"volume\":\"77 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2021-12-22\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"3\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2021 3rd International Conference on Electrical & Electronic Engineering (ICEEE)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/ICEEE54059.2021.9718890\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2021 3rd International Conference on Electrical & Electronic Engineering (ICEEE)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/ICEEE54059.2021.9718890","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 3

摘要

脱机手写文本识别(HTR)是一种将手写图像转换为数字可编辑文本格式的技术。由于草书字母、标点符号和复合字符的存在,识别孟加拉语手写文本更加复杂。多年来,HTR系统光学模型的几种方法已经开发出来,包括隐马尔可夫模型(HMM)或深度学习技术,如卷积循环神经网络(CRNN),以及当前最先进的基于门特cnn的架构。尽管如此,可用于孟加拉语单词识别的工作相对有限。本文介绍了一个端到端的孟加拉语词识别系统。我们使用了各种流行的预训练CNN架构,包括excepeption、MobileNet和DenseNet,其次是循环单元,如LSTM或GRU。此外,我们实验了Puigcerver的基于CRNN和Flor的基于gate - cnn的光学模型架构,这些作品在孟加拉可用。floor architecture在我们的实验中提供了最高的识别率,CER为12.83%,WER为36.01%。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
Offline Bangla Handwritten Text Recognition: A Comprehensive Study of Various Deep Learning Approaches
Offline Handwritten Text Recognition (HTR) is a technique for translating handwritten images into digitally editable text format. Due to the presence of cursive letters, punctuation marks, and compound characters, it is more complex to recognize Bangla handwritten text. Over the years, several approaches to the optical model of the HTR system have been developed, including Hidden Markov Model (HMM) or deep learning techniques such as Convolutional Recurrent Neural Networks (CRNN), and current state-of-the-art Gated-CNN based architectures. Despite this, there are relatively limited works available for Bangla word recognition. In this paper, we introduce an end-to-end system for Bangla word recognition. We used a variety of popular pre-trained CNN architectures, including Xception, MobileNet, and DenseNet, followed by recurrent units such as LSTM or GRU. Furthermore, we experimented with Puigcerver’s CRNN based and Flor’s Gated-CNN based optical model architectulimited works available in Bangla.res. Flor architecture provided the highest recognition rate in our experiment, with a CER of 12.83% and a WER of 36.01%.
求助全文
通过发布文献求助,成功后即可免费获取论文全文。 去求助
来源期刊
自引率
0.00%
发文量
0
期刊最新文献
Computer-Aided Polyp Removal Detection in Endoscopic Images FPGA based Histogram Equalization for Image Processing Spreading Loss Model for Channel Characterization of Future 6G Terahertz Communication Networks Impact of Cladding Rectangular Bars on the Antiresonant Hollow Core Fiber Predicting Autism Spectrum Disorder Based On Gender Using Machine Learning Techniques
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1