zenodoopen
ArXiV-Entity/Relation annotated dataset
<p>This dataset is a collection of abstracts from the CS section of ArXiV, each annotated with <a href="https://github.com/dwadden/dygiepp">DyGIE++</a> (SciERC model)</p> <p>The dataset can be used to train triple extractors or to cluster triples (in the Computer Science and AI domains).</p> <p>Supersedes the ArXiV-AIKG dataset as these triples are unconstrained (so they don't forcibly appear in AIKG)</p>
ShareScore
44/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 20
- Reuse readiness
- 8
- Engagement
- 4