Skip to main content
zenodoopen

Telegram digits dataset

<p>This dataset is MNIST-like, containing digitized handwritten characters extracted from electoral telegrams during the General Elections of Santa Fe, Argentina, in the year 2021. The dataset offers a valuable resource for researchers and practitioners in the field of character recognition, particularly in the context of electoral data analysis. Each sample in the dataset represents a single digit, ranging from 0 to 9, handwritten by different individuals participating in the electoral process. The dataset aims to facilitate the development and evaluation of machine learning and computer vision algorithms for character recognition tasks.</p> <p>It contains 170718 images, splitted in train (119502), validation (25608) and&nbsp;test (25608).</p> <p>This dataset is part of master&#39;s thesis in data science which aims to build an Optical Character Recognition (OCR) system using domain adaptation techniques titled &quot;Classification of digits written in the telegrams of legislative elections in Santa Fe using Domain Adaptation techniques&quot;.</p>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
16
Reuse readiness
8
Engagement
0

Topics