Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
5
datasets available to search
ShareScore release 0.9.0
Dataset results
5 results for “semantic targeting”
Decrypting cryptic crosswords: Semantically complex wordplay puzzles as a target for NLP
<p>Cryptic crosswords, the dominant crossword variety in the UK, are a promising target for advancing NLP systems that seek to process semantically complex, highly compositional language. Cryptic clues read like fluent natural language but are adversarially composed of two parts: a definition and a wordplay cipher requiring character-level manipulations. Expert humans use creative intelligence to solve cryptics, flexibly combining linguistic, world, and domain knowledge. In this paper, we make two main contributions. First, we present a dataset of cryptic clues as a challenging new benchmark for NLP systems that seek to process compositional language in more creative, human-like ways. After showing that three non-neural approaches and T5, a state-of-the-art neural language model, do not achieve good performance, we make our second main contribution: a novel curriculum approach, in which the model is first fine-tuned on related tasks such as unscrambling words. We also introduce a challenging data split, examine the meta-linguistic capabilities of subword-tokenized models, and investigate model systematicity by perturbing the wordplay part of clues, showing that T5 exhibits behavior partially consistent with human solving strategies. Although our curricular approach considerably improves on the T5 baseline, our best-performing model still fails to generalize to the extent that humans can. Thus, cryptic crosswords remain an unsolved challenge for NLP systems and a potential source of future innovation.</p>
Decrypting cryptic crosswords: Semantically complex wordplay puzzles as a target for NLP
Open the record for dataset details and reuse information.
Semantic targeting: past, present and future
<p>Keynote address day one of the ISKO UK <a href="https://www.iskouk.org/2009-Conference">2009 biennale conference </a>.</p> <p>This keynote address looked at the evolution of the linguistic approach to content analysis which Crystal has been developing over the past twenty years. It begins with the knowledge management taxonomy used for the Cambridge family of general encyclopedias, and follows its transformation into an Internet taxonomy, with applications in automatic document classification, search engine assistance, e-commerce, online advertising, and<br> Internet security. Recent developments have brought a focus on advertising, a field which has seen ideas develop from simple keyword analysis to contextual advertising and now to semantic targeting.<br> David Crystal explores the difference between these notions, and describes current issues in the way semantic targeting is evolving, including ways of handling site sensitivity, sentiment, intention, and cultural localization.</p> <p> </p>
Rehabilitation of Post-stroke Aphasia by Targeting Phonological, and Lexico-semantic Deficits With Speech Output Tasks
ClinicalTrials.gov study NCT06451731. IPD Sharing: YES. Countries: 0. Publications: 0.
Rehabilitation of post-stroke aphasia by a single protocol targeting phonological, lexical, and semantic deficits with speech output tasks
<p>This study assessed the effectiveness of a novel rehabilitation protocol (PHOLEXSEM), focused on PHonological, SEmantic, and LExical deficits, aiming at improving lexical retrieval, and, generally, spoken output. The study is published here: https://doi.org/10.23736/S1973-9087.24.08576-9</p> <p><span>These data cannot be made publicly available because they include sensitive information, but can be made available to interested researchers upon reasonable request. Please, forward your request to Prof. Nadia Bolognini (n.bolognini@auxologico.it).</span></p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.