Skip to main content
zenodoopen

FDup deduplication software data benchmark: 10Mi OpenAIRE Publications Dump

<p>This dataset is a random subset of publications extracted&nbsp;from the OpenAIRE Research Graph (<a href="http://doi.org/10.5281/zenodo.4707307">http://doi.org/10.5281/zenodo.4707307</a>). The dataset contains ~10Mi JSON publications records.&nbsp;</p> <p>The file is a zip archive containing gz files, each with one JSON per line. Each JSON is compliant to the schema available at&nbsp;<a href="http://doi.org/10.5281/zenodo.4723403">http://doi.org/10.5281/zenodo.4723403</a>.<br> <br> Learn more about the OpenAIRE Research Graph at&nbsp;<a href="https://graph.openaire.eu/">https://graph.openaire.eu</a>.</p>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
8
Access
16
Reuse readiness
8
Engagement
0

Topics