Skip to main content
zenodoopen

ArXiV-Entity/Relation annotated dataset

<p>This dataset is a collection of abstracts from the CS section of ArXiV, each annotated with <a href="https://github.com/dwadden/dygiepp">DyGIE++</a> (SciERC model)</p> <p>The dataset can be used to train triple extractors or to cluster triples (in the Computer Science and AI domains).</p> <p>Supersedes the ArXiV-AIKG dataset as these triples are unconstrained (so they don&#39;t forcibly appear in AIKG)</p>

ShareScore

44/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
20
Reuse readiness
8
Engagement
4

Topics