Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
278
datasets available to search
ShareScore release 0.9.0
Dataset results
278 results for “Validated dataset”
Three datasets for the validation of a digital leadership inventory
<p>Three datasets for the validation of a digital leadership inventory (n=482, n=362, n=491)</p>
Dataset for the validation of the Delirium Observation Screening Scale in long-term care facilities in Flanders
Open the record for dataset details and reuse information.
Datasets for "Validating Mechanistic Interpretations: An Axiomatic Approach"
Open the record for dataset details and reuse information.
Four psychometrically validated datasets for benchmarking large language models, based on the TIMSS 2008 and 2011 released items.
<p>Four datasets validated according to psychometric principles that can be used to benchmark large language models in terms of achievements in advanced school math, advanced school physics, 8th grade math and 8th grade science.</p> <p>These four datasets are derived from items released by Trends in International Mathematics and Science Study Advanced 2008 and Trends in International Mathematics and Science Study 2011. See <a href="https://nces.ed.gov/timss/released-questions.asp">link</a>.</p> <p>For more information, see our paper <a href="https://arxiv.org/abs/2404.01799">PATCH! Psychometrics-AssisTed benCHmarking of Large Language Models: A Case Study of Mathematics Proficiency</a>.</p>
Validation and Test Datasets for "High-resolution AI image dataset for diagnosing oral submucous fibrosis and squamous cell carcinoma"
<p>This deposition contains the validation and test dataset for our study "High-resolution AI image dataset for diagnosing oral submucous fibrosis and squamous cell carcinoma".</p> <p>The training dataset for this study can be found at the following DOI: [<strong>10.5281/zenodo.12636426</strong>].<br><br></p>
Setup and Dataset for the Validation of an Eulerian-Lagrangian Coupling Method in the m-AIA framework
<p>This repository holds data files, code information and property files which are used to conduct performance analyses of a <br>Parallel Eulerian-Lagrangian Coupling Method. </p>
FASTQ datasets for the validation of the Neisseria pipeline
<p>todo: update</p>
Adversarial Validation for quantifying dissimilarity experiments datasets and codes (new)
<p>Readme.txt describes how to conduct the code.</p> <p>This rar/zip file includes all materials of Adversarial Validation for quantifying dissimilarity experiments. </p> <p>They are ordered by the first number of folder's name.</p> <p>In each folder, the order of running code scripts are labeled by the first number of code's name.</p>
Identification of crashworthy designs combining active learning and the solution space methodology - Validation and training dataset
<p>Dataset and Python codes for training a crashworthiness classifier.</p>
Dataset for Validity and reliability study of Personal Resource Questionnaire-2000 Indonesia version (PRQ2000-INA) to measure perceived social support among people with dementia in Indonesia
<p>This is a dataset for "Validity and reliability study of Personal Resource Questionnaire-2000 Indonesia version (PRQ2000-INA) to measure perceived social support among people with dementia in Indonesia"</p>
Dataset 2nd validation NbS CoBAs tool (SPSS.SAV)
<p>Database of the co-benefit scale of nature-based solutions for the second validation with a sample of students (N=115). Participants have to evaluate 8 urban public spaces with different degrees of naturalization, openness,... (independent or contextual variables. The unit of analysis is the evaluations (N=437). The evaluated co-benefits are psychosocial and are structured around 5 general categories: environmental comfort, psychological restoration, safety, social flow and emotional change.</p>
Dataset to train, validate and reconstruct POC over the global ocean for 2009-2013 based on PlankTOM12
<p>The distribution of in situ UVP5 measurements over the period 2009-2013 was used to create synthetic data by sampling a global biogeochemical ocean model PlankTOM12 at UVP5 time and location.</p> <p>The synthetic data set is used to train, validate and test Machine Learning methods to reconstruct particulate organic carbon concentration. </p> <p>These data are part of publication at the GMD. </p>
Dataset of the publication "Theory and Experimental Validation of Two Techniques for Compensating VT Nonlinearities"
<p>This is a dataset for paper published:</p> <p>G. D’Avanzo <em>et al</em>., "Theory and Experimental Validation of Two Techniques for Compensating VT Nonlinearities," in <em>IEEE Transactions on Instrumentation and Measurement</em>, vol. 71, pp. 1-12, 2022, Art no. 9001312, doi: 10.1109/TIM.2022.3147883.</p> <p> </p> <p>Excel file provides data for VT_B.</p>
Benchmark dataset for verification and validation of elPaSo Core module
<p>This dataset contain the set of vibroacoustic benchmark problems for verification and validation of the FEM research code "elPaSo Core".</p>
In situ SIF dataset for validating OCO2/3 SIF products
<p>The matched in situ SIF observations for validating OCO2/3 SIF products</p>
TA-EG Questionnaire italian Translation - Validation Dataset
<p>This file contains the raw data of TA-EG questionnaire single item information and the comments provided for each item </p>
Validation of a new global irrigation scheme in the ORCHIDEE land surface model - Dataset
<p>Datasets used in the paper 'Validation of a new global irrigation scheme in the ORCHIDEE land surface model', submitted to GMD</p>
Dataset for validation of app myjump2
<p>The dataset shows the results obtained during the data collection</p>
Analysis datasets for NEMO_validation workflow Byrne et al 2023 GMD. "Using the COAsT Python package to develop a standardised validation workflow for ocean physics models"
<p>Analysis datasets in support of Byrne et al. (2023) "Using the COAsT Python package to develop a standardised validation workflow for ocean physics models", <em>Geoscientific Model Development</em>.</p> <p> </p> <p>The datasets are from a comparative analysis of two versions of the European shelf sea AMM15 (Atlantic Margin Model at 1.5km horizontal resolution) configuration. These are NEMO ocean model configurations with different code base versions. The configurations are CO7, which is based on NEMOv3.6, and CO9p0 (also referred to as P0.0), which is based on NEMOv4.0.4.</p>
Proposing more ecologically-valid experiment protocol using YouTube platform - Dataset
<p>Uploaded data related to publication "Proposing more ecologically-valid experiment protocol using YouTube platform". Included in the files are: data analysis project, database, video recordings of testers' behaviors.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.