DUPS: Diachronic Usage Pair Similarity
<p>The DUPS (Diachronic Usage Pair Similarity) dataset contains similarity judgements of English word usage pairs from different time periods, as described in the paper below. </p> <p>The WUG version of the DUPS dataset (version 2.0.0) contains diachronic Word Usage Graphs constructed from the similarity judgements of English word usage pairs contained in DUPS. In a word usage graph, the usages of a word are represented as nodes connected by edges weighted according to (human-annotated) semantic proximity. A description of the data format as well as the code used to generate the graphs from DUPS can be found at <a href="https://www.ims.uni-stuttgart.de/data/wugs">https://www.ims.uni-stuttgart.de/data/wugs</a>.</p> <p>Both versions of the DUPS dataset can be downloaded from the Files section of this web page.</p> <p>Please cite this paper if you use any version of the dataset in your work:</p> <blockquote> <p>Mario Giulianelli, Marco Del Tredici, and Raquel Fernández. 2020. <a href="https://aclanthology.org/2020.acl-main.365/">Analysing Lexical Semantic Change with Contextualised Word Representations</a>. In <em>Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (ACL-2020)</em>. Association for Computational Linguistics.</p> </blockquote> <p> </p>
ShareScore
36/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 16
- Reuse readiness
- 0
- Engagement
- 8