{"title":"Revolutionizing Digit Image Recognition: Pushing the Limits with Simple CNN and Challenging Image Augmentation Techniques on MNIST","authors":"Khodijah Hulliyah","doi":"10.47738/jads.v4i3.104","DOIUrl":null,"url":null,"abstract":"This study aims to apply Convolutional Neural Networks (CNN) and image augmentation techniques in digit recognition using the MNIST dataset. We built a CNN model and experimented with various image augmentation techniques to improve digit recognition accuracy. The results showed that the use of CNN with image augmentation techniques was effective in improving digit recognition performance. In the data collection stage, we used the MNIST dataset consisting of images of handwritten digits as training and testing data. After building the CNN model, we apply image augmentation techniques such as rotation, shift, and flipping to the training data to enrich the data variety and prevent overfitting. The evaluation results show that the CNN model that has been trained with image augmentation techniques produces significant accuracy, with a maximum accuracy of 99.81%. We also performed an ensemble of several CNN models and found that this approach increased the digit recognition accuracy to 99.79%. This research has the potential for further development. Recommendations for further research include exploring more specific and complex image augmentation techniques, as well as using more challenging datasets. In addition, future research may consider improvements to the CNN architecture used or combining it with other methods such as recurrent neural networks (RNN).","PeriodicalId":479720,"journal":{"name":"Journal of Applied Data Sciences","volume":"356 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2023-09-15","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Journal of Applied Data Sciences","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.47738/jads.v4i3.104","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 0
Abstract
This study aims to apply Convolutional Neural Networks (CNN) and image augmentation techniques in digit recognition using the MNIST dataset. We built a CNN model and experimented with various image augmentation techniques to improve digit recognition accuracy. The results showed that the use of CNN with image augmentation techniques was effective in improving digit recognition performance. In the data collection stage, we used the MNIST dataset consisting of images of handwritten digits as training and testing data. After building the CNN model, we apply image augmentation techniques such as rotation, shift, and flipping to the training data to enrich the data variety and prevent overfitting. The evaluation results show that the CNN model that has been trained with image augmentation techniques produces significant accuracy, with a maximum accuracy of 99.81%. We also performed an ensemble of several CNN models and found that this approach increased the digit recognition accuracy to 99.79%. This research has the potential for further development. Recommendations for further research include exploring more specific and complex image augmentation techniques, as well as using more challenging datasets. In addition, future research may consider improvements to the CNN architecture used or combining it with other methods such as recurrent neural networks (RNN).