Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
110
datasets available to search
ShareScore release 0.9.0
Dataset results
110 results for “schema”
A Multiple Case Study of Schema Therapy for Difficult-to-treat Depression- DEPRE-ST*Case
ClinicalTrials.gov study NCT06655623. IPD Sharing: YES. Countries: 1. Publications: 4.
Validation of Schema Scale of Mental Health Service
ClinicalTrials.gov study NCT05952063. IPD Sharing: NO. Countries: 1. Publications: 1.
Effects of Schema Therapy vs. Cognitive Behavioral Therapy vs. Individual Supportive Therapy
ClinicalTrials.gov study NCT03287362. IPD Sharing: NO. Countries: 1. Publications: 6.
Mouse head schema
Drawing uploaded to scidraw.io on: 12 October 2019
SchemaPile: A Large Collection of Relational Database Schemas [OLD VERSION]
<p>[OLD VERSION -- Please visit <a href="../records/12682521"> https://zenodo.org/records/12682521</a> instead]</p> <p>Access to fine-grained schema information is crucial for understanding how relational databases are designed and used in practice, and for building systems that help users interact with them. Furthermore, such information is required as training data to leverage the potential of large language models (LLMs) for improving data preparation, data integration and natural language querying.<br><br>Existing single-table corpora such as GitTables provide insights into how tables are structured in-the-wild, but lack detailed schema information about how tables relate to each other, as well as metadata like data types or integrity constraints. On the other hand, existing multi-table (or database schema) datasets are rather small and attribute-poor, leaving it unclear to what extent they actually represent typical real-world database schemas. </p> <p>In order to address these challenges, we present SchemaPile, a corpus of 221,171 database schemas, extracted from SQL files on GitHub. It contains 1.7 million tables with 10 million column definitions, 700 thousand foreign key relationships, seven million integrity constraints, and data content for more than 340 thousand tables. We conduct an in-depth analysis on the millions of schema metadata properties in our corpus, as well as its highly diverse language and topic distribution. In addition, we showcase the potential of SchemaPile to improve a variety of data management applications, e.g., fine-tuning LLMs for schema-only foreign key detection, improving CSV header detection and evaluating multi-dialect SQL parsers. We publish the code and data for recreating SchemaPile and a permissively licensed subset SchemaPile-Perm. </p> <p>GitHub repo: <a href="http://github.com/amsterdata/schemapile">http://github.com/amsterdata/schemapile</a></p>
Unprocessed data for the "Deriving Semantics-Aware Fuzzers from Web API Schemas" paper.
<p>Unprocessed data for the "Deriving Semantics-Aware Fuzzers from Web API Schemas" paper.</p>
Workflow Trace Archive workflowhub_epigenomics_dataset-hep_grid5000_schema-0-2_epigenomics-hep-g5k-run001 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_epigenomics_dataset-ilmn_chameleon-cloud_schema-0-2_epigenomics-ilmn-100000-cc-run004 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_soykb_grid5000_schema-0-2_soykb-g5k-run002 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_montage_ti01-971107n_degree-2-0_osg_schema-0-2_montage-2-0-osg-run007 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_epigenomics_dataset-hep_futuregrid_schema-0-2_epigenomics-hep-fg-run001 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_epigenomics_dataset-hep_chameleon-cloud_schema-0-2_epigenomics-hep-100000-cc-run005 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_epigenomics_dataset-taq_chameleon-cloud_schema-0-2_epigenomics-taq-100000-cc-run002 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_montage_ti01-971107n_degree-4-0_osg_schema-0-2_montage-4-0-osg-run009 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_montage_dataset-02_degree-2-0_osg_schema-0-2_montage-2-0-osg-run007 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Workflow Trace Archive workflowhub_montage_dataset-02_degree-4-0_osg_schema-0-2_montage-4-0-osg-run009 trace
Workload downloaded from WorkflowHub, see http://workflowhub.isi.edu/.
Text-fig. 5. Schema of radial section of T. gypsaceum (sample Bečov 2). t – tracheid, r – ray, bp – bordered pit, tp – taxodioid pit, cp – cupressoid pit, c – crassulae. in New Fossil Woods From The Paleogene Of Doupovské Hory And České Středohoří Mts. (Bohemian Massif, Czech Republic)
Text-fig. 5. Schema of radial section of T. gypsaceum (sample Bečov 2). t – tracheid, r – ray, bp – bordered pit, tp – taxodioid pit, cp – cupressoid pit, c – crassulae.
Text-fig. 11. Schema of transversal section of Ulmoxylon cf. kersonianum (sample 1/3). v – vessel, r – ray, grb – one growth ring. in New Fossil Woods From The Paleogene Of Doupovské Hory And České Středohoří Mts. (Bohemian Massif, Czech Republic)
Text-fig. 11. Schema of transversal section of Ulmoxylon cf. kersonianum (sample 1/3). v – vessel, r – ray, grb – one growth ring.
Figure 1 from: Penev L, Lyal C, Weitzman A, Morse D, King D, Sautter G, Georgiev T, Catapano T, Agosti D (2011) XML schemas and mark-up practices of taxonomic literature. ZooKeys 150: 89-116. https://doi.org/10.3897/zookeys.150.2213
Figure 1 - Flowchart of mark-up, publication, dissemination and use of taxonomic information. Scratchpads (http://scratchpads.eu/) and EDIT Cybertaxonomy Platform (http://wp5.e-taxonomy.eu/) stand for the community-based collaborative platforms for taxonomists developed by the EDIT FP6 project (http://www.e-taxonomy.eu/).
Study schema: Kathleen Buchheit. (2023). Dupilumab and daily high-dose aspirin therapy synergistically reduce nasal eosinophil markers in NSAID-exacerbated respiratory disease.
<p>Study schema</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.