2018/06/06 by Alam, Samiul, Reasat, Tahsin, Doha, Rashed Mohammad +1 · 1 citation
#68T10 #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #I.5.1 #I.5.4
paper · doi:10.48550/arxiv.1806.02452
To benchmark Bengali digit recognition algorithms, a large publicly available dataset is required which is free from biases originating from geographical location, gender, and age. With this aim in mind, NumtaDB, a dataset consisting of more than 85,000 images of hand-written Bengali digits, has been assembled. This paper documents the collection and curation process of numerals along with the salient statistics of the dataset.