zenodoopen
Republic Print Dataset
<p>The republic print dataset consists of 107 ground truthed scans </p> <p>Using annotation software provided through the Transkribus Platform we annotated scans, concerning mostly 18th century printed documents from the National Archive of the Netherlands, with their layout consisting of baselines and regions. The resulting ground truth was used to train a machine learning model yielding very accurate results. The ground truth was made available as an open access dataset.</p>
ShareScore
44/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 8
- Engagement
- 8