Skip to main content
zenodoopen

Republic Print Dataset

<p>The republic print dataset consists of 107 ground truthed scans&nbsp;</p> <p>Using annotation software provided through the Transkribus Platform we annotated scans, concerning mostly 18th century printed documents from the National Archive of the Netherlands, with their layout consisting of&nbsp;baselines and regions. The resulting ground truth was used to train a machine learning model yielding very accurate results. The ground truth was made available as an open access dataset.</p>

ShareScore

44/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
16
Reuse readiness
8
Engagement
8

Topics