Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
110
datasets available to search
ShareScore release 0.9.0
Dataset results
110 results for “schema”
Dataset for manuscript "Sleep does not influence schema-facilitated motor memory consolidation"
<p>Dataset containing the raw as well as the subject-level data for the two experiments reported in the manuscript "Sleep does not influence schema-facilitated motor memory consolidation".</p>
Open Context Database SQL Dump: Legacy Schema Tables and New Schema Tables
<p>Open Context (<a href="https://opencontext.org">https://opencontext.org</a>) publishes free and open access research data for archaeology and related disciplines. An open source (but bespoke) Django (Python) application supports these data publishing services. The software repository is here: <a href="https://github.com/ekansa/open-context-py">https://github.com/ekansa/open-context-py</a></p> <p>The Open Context team runs ETL (extract, transform, load) workflows to import data contributed by researchers from various source relational databases and spreadsheets. Open Context uses PostgreSQL (<a href="https://www.postgresql.org">https://www.postgresql.org</a>) relational database to manage these imported data in a graph style schema. The Open Context Python application interacts with the PostgreSQL database via the Django Object-Relational-Model (ORM).</p> <p>In 2023, the Open Context team finished migration of from a legacy database schema to a revised and refactored database schema with stricter referential integrity and better consistency across tables. During this process, the Open Context team de-duplicated records, cleaned some metadata, and redacted attribute data left over from records that had been incompletely deleted in the legacy schema.</p> <p>This database dump includes all Open Context data organized with the legacy schema (table names that start with the 'oc_' or 'link_' prefixes) along with all Open Context data after cleanup and migration to the new database schema (table names that start with 'oc_all_'). The binary media files referenced by these structured data records are stored elsewhere. Binary media files for some projects, still in preparation, are not yet archived with long term digital repositories.</p> <p>These data comprehensively reflect the structured data currently published and publicly available on Open Context. Other data (such as user and group information) used to run the Website are not included. </p> <p> </p> <p><strong>IMPORTANT</strong></p> <p>This database dump contains data from roughly 180 different projects. Each project dataset has its own metadata and citation expectations. If you use these data, you must cite each data contributor appropriately, not just this Zenodo archived database dump.</p> <p> </p> <p> </p>
ENPKG graph schema
<p>Schema of the Experimental Natural Products Knowledge Graph</p>
Neisseria gonorrhoeae clustering to reveal major European WGS-based genogroups in association with antimicrobial resistance (cgMLST and MScgMLST schemas, allelic profile matrices and GrapeTree input file)
<p>This dataset refers to the gene-by-gene analysis of 3791 <em>Neisseria gonorrhoeae</em> genomes from 21 European countries and includes the used cgMLST and MScgMLST loci schemas prepared for the chewBBACA core suite, as well as the associated allelic profile matrices for all genomes. Additionally a <em>.json</em> file is made available for direct input in the GrapTree vizualization software for data/metadata exploration. </p> <p>All novel raw sequence reads used in this study were deposited in the European Nucleotide Archive (ENA) (BioProject PRJEB36482). Additional raw sequence read data used were retrieved from the following ENA BioProjects: PRJEB14933; PRJEB2124; PRJEB23008; PRJEB26560; PRJEB9227; PRJNA275092; PRJNA348107; PRJNA473385; PRJNA315363. </p>
Replicability and Reproducibility of a Schema Evolution Study in Embedded Databases
<p>Archives containing datasets, scripts, and instructions for reproducing a study.</p>
Russian prefixed verbs as constructional schemas: evidence of audience design
<p>Files necessary to replicate the study findings.</p>
Processed data for the "Deriving Semantics-Aware Fuzzers from Web API Schemas" paper
<p>Processed data for the "Deriving Semantics-Aware Fuzzers from Web API Schemas" paper. Each directory in the archive consists of:</p> <p>- metadata.json. Metadata about a test run - tested fuzzer name, run duration, etc</p> <p>- fuzzer.json - Structured fuzzer output</p> <p>- deduplicated_cases.json - Deduplicated reported failures, when fuzzers provide it</p> <p>- sentry.json - Cleaned Sentry events for this run</p> <p>- target.json - Parsed stdout for Gitlab & Disease.sh targets that were tested without Sentry integration</p>
Valentwin: Using Self-Supervised Contrastive Learning on Language Model for Schema Matching Datasets
<div>ValenTwin is a schema matching framework that uses self-supervised contrastive learning to train the model, uses the model to generate embeddings of table columns, then uses different similarity measures to match the column embeddings.</div> <div> </div> <div> <div>We provide two types of zip files for the datasets:<br>1. `data.zip` contains the raw data files, the ground truth files, the sampled data (n=[100, 200, 300, 400, 500] used in the experiments, as well as the contrastive data used to train the model.<br>2. `data-raw.zip` contains only the raw data files and the ground truth files. You can sample the data and generate the contrastive dataset yourself by following step 1 and 2 in the `How to Run` section. <br>Download and unzip one of the zip files to the `data` folder.</div> </div>
Database schema
<p>database schema</p>
Database schema quality analysis dataset
<p>Database schema quality analysis dataset</p>
RaDISAN - Linee guida 2023, Business Rules, schema XSD e tool Excel.
<ol> <li>Linee guida per la compilazione dei file in formato XML per la trasmissione dei dati relativi a campioni prelevati nel 2022.</li> <li>File contenente i controlli (Business rules e Domain rules), campi coinvolti e messaggi d'errore del sistema. Il file contiene tutte le tabelle di controllo utilizzate per i controlli BR ad eccezione delle anagrafiche che vengono pubblicate nella sezione dedicata.</li> <li>Schema XSD da utilizzare pe la costruzione / validazione del file XML.</li> <li>Tool Excel per popolare la tabella dati e creare un corretto file XML.</li> </ol>
Biological Effects of Schema Therapy
ClinicalTrials.gov study NCT06367907. IPD Sharing: Not stated. Countries: 1. Publications: 1.
The Influence of Treatment Format on Schema Therapy for Borderline Personality Disorder
ClinicalTrials.gov study NCT05986552. IPD Sharing: NO. Countries: 1. Publications: 4.
Clinical and Cost-effectiveness of Group Schema Therapy for Complex Eating Disorders: the GST-EAT Study
ClinicalTrials.gov study NCT05812950. IPD Sharing: NO. Countries: 1. Publications: 1.
The Influence of Collective Schemas on Individual Memory (MULTIBRAIN_1)
ClinicalTrials.gov study NCT02172677. IPD Sharing: Not stated. Countries: 1. Publications: 1.
Schema, Emotion, and Behavior Therapy for Children
ClinicalTrials.gov study NCT01784263. IPD Sharing: Not stated. Countries: 1. Publications: 2.
Schema, Emotion and Behavior-Based Therapy for School Children
ClinicalTrials.gov study NCT02010086. IPD Sharing: Not stated. Countries: 1. Publications: 2.
Early Maladaptative Schemas Among Euthymic Patients With Bipolar Disorder in the Versailles FondaMental Advanced Centers of Expertise for Bipolar Disorders Cohort
ClinicalTrials.gov study NCT03977909. IPD Sharing: NO. Countries: 1. Publications: 6.
The Influence of Collective Schemas on Individual Memory (MULTIBRAIN_2)
ClinicalTrials.gov study NCT02542800. IPD Sharing: Not stated. Countries: 1. Publications: 1.
Schema Therapy for Patients With Chronic Treatment Resistant Depression
ClinicalTrials.gov study NCT05833087. IPD Sharing: YES. Countries: 1. Publications: 21.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.