Skip to main content
zenodoopen

DUPS: Diachronic Usage Pair Similarity

<p>The DUPS (Diachronic Usage Pair Similarity) dataset contains similarity judgements of English word usage pairs from different time periods, as described in the paper&nbsp;below.&nbsp;</p> <p>The WUG version of the DUPS dataset (version 2.0.0) contains diachronic Word Usage Graphs constructed from the similarity judgements of English word usage pairs contained in DUPS. In a word usage graph, the usages of a word are represented as nodes connected by edges weighted according to (human-annotated) semantic proximity. A description of the data format as well as the code used to generate the graphs from DUPS can be found at <a href="https://www.ims.uni-stuttgart.de/data/wugs">https://www.ims.uni-stuttgart.de/data/wugs</a>.</p> <p>Both versions of the DUPS dataset can be downloaded from the Files section of this web page.</p> <p>Please cite this paper if you use any version of the dataset in your work:</p> <blockquote> <p>Mario Giulianelli, Marco Del Tredici, and Raquel Fern&aacute;ndez. 2020. <a href="https://aclanthology.org/2020.acl-main.365/">Analysing Lexical Semantic Change with Contextualised Word Representations</a>. In <em>Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (ACL-2020)</em>. Association for Computational Linguistics.</p> </blockquote> <p>&nbsp;</p>

ShareScore

36/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
16
Reuse readiness
0
Engagement
8

Topics