Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

1,045

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

1,045 results for “Generated Data”

Learn how ShareScore rates datasets ↗
zenodo44/100

Data and results of the example used in the SI-Hg D1 protocol for the SI-traceable calibration of elemental mercury (Hg0) gas generators used in the field

<p>During the SI-Hg project a metrological traceable protocol for the calibration of mercury gas generators used in the field&nbsp;was developed and validated. The SI-Hg calibration protocol specifies the procedures for establishing traceability to the SI units for the quantitative output of elemental mercury generators that are employed in regulatory applications for emission monitoring or testing. This protocol provides methods for</p><ul><li>the experimental procedures to compare the output of elemental mercury gas generators</li><li>the data processing for determination of mercury concentration and the expanded uncertainty of the mercury concentration obtained from the elemental mercury gas generator.</li></ul><p>In the protocol examples are given to explain the data processing, determining the mercury concentration and corresponding uncertainty. In this repository the raw data and results used for the example calculated with the data processing script can be found.</p>

opencc-by-4.0Nov 2023View details →
zenodo44/100

scGraph2Vec: a deep generative model for gene embedding augmented by Graph Neural Network and single-cell omics data

<p>This repository contains the training data and source code to reproduce the results of our paper:<br>scGraph2Vec: a deep generative model for gene embedding augmented by Graph Neural Network and single-cell omics data</p> <p>More description can be also found in GitHub (https://github.com/LPH-BIG/scGraph2Vec).</p>

opencc-zeroJun 2024View details →
zenodo44/100

Supplementary Data Files for the paper "Intrinsically disordered compositional bias in proteins: Sequence traits, region clustering, and generation of hypothetical functional associations"

<div> <div> <div> <div> <p><strong>Supplementary data files relating to <a href="https://doi.org/10.1177/11779322241287485">https://doi.org/10.1177/11779322241287485.&nbsp;</a></strong></p> <p><strong><span>Suppl. File 1: Protein Family Clusters.</span></strong></p> <p><strong><span>Suppl. File 2: Cluster GO enrichments/depletions. </span></strong></p> <p><strong><span>Suppl. File 3: The raw ID-CBR data with annotations. </span></strong></p> <p><strong><span>Suppl. File 4: &shy;ID-CBR Cluster membership.</span></strong></p> <p><strong><span>Each file has an explanatory header.&nbsp;</span></strong></p> <p>&nbsp;</p> </div> </div> </div> </div>

opencc-by-4.0Oct 2024View details →
zenodo44/100

Source Data and ambient ozone dataset generated in "Substantially underestimated global health risks of current ozone pollution"

<p>Existing assessments might have underappreciated ozone-related health impacts worldwide. Here our study assesses current global ozone pollution using the high-resolution (0.05&deg;) estimation from a geo-ensemble learning model, with key focuses on population exposure and all-cause mortality burden. Our model demonstrates strong performance, achieving a mean bias of less than -1.5 parts per billion against in-situ measurements. We estimate that 66.2% of the global population is exposed to excess ozone for short term (&gt; 30 days per year), and 94.2% suffers from long-term exposure. Furthermore, severe ozone exposure levels are observed in Cropland areas, particularly over Asia. Importantly, the all-cause ozone-attributable deaths significantly surpass previous recognition from specific diseases worldwide. Notably, mid-latitude Asia (30&deg;N) and the western United States show high mortality burden, contributing substantially to global ozone-attributable deaths. Our study highlights current significant global ozone-related health risks and may benefit the ozone-exposed population in the future.</p>

opencc-by-4.0Nov 2024View details →
zenodo44/100

Global Surface Ozone Concentration Dataset 1990-2017 Generated by Bayesian Maximum Entropy Data Fusion With RAMP Bias Correction

<p>This dataset reports estimates of surface ozone concentration at fine spatial resolution for 1990 to 2017, at 0.5 degree horizontal resolution.&nbsp; Also reported is the variance.&nbsp; Estimates correspond to this paper:</p> <p><span>Becker, J. S.</span><span>, DeLang, M. N., K.-L. Chang, M. L. Serre, O. R. Cooper, <u>H. Wang</u>, M. G. Schultz, S. Schroder, X. Lu, L. Zhang, M. Deushi, B. Josse, C. A. Keller, J.-F. Lamarque, M. Lin, J. Liu, V. Marecal, S. A. Strode, K. Sudo, S. Tilmes, L. Zhang, M. Brauer, and <span>J. J. West</span> (2023) Using Regionalized Air Quality Model Performance and Bayesian Maximum Entropy data fusion to map global surface ozone concentration, <em>Elementa Science of the Anthropocene</em>, 11: 1, doi: 10.1525/elementa.2022.00025.</span></p> <p>The dataset reports estimates of surface ozone for the OSDMA8 metric (the 6-month ozone-season average of the daily maximum 8-hr concentration), estimated through a data fusion of ozone observations from the Tropospheric Ozone Assessment Report (TOAR) database, and output from multiple global atmospheric models.&nbsp; Estimates are created in each year by a combination of M3Fusion to create a multi-model composite, Regional Air Quality Model Performance (RAMP) regional and nonlinear bias correction, and Bayesian Maximum Entropy (BME) data fusion in space and time.&nbsp; The estimates here are the final results using a weighted RAMP bias correction.&nbsp;</p>

opencc-by-4.0Jan 2024View details →
zenodo44/100

Efficient embryoid-based method to improve generation of optic vesicles from human induced pluripotent stem cells data

<p>Animal models have provided many insights into ocular development and disease, but they remain suboptimal for understanding human oculogenesis. Eye development requires spatiotemporal gene expression patterns and disease phenotypes can differ significantly between humans and animal models, with patient-associated mutations causing embryonic lethality reported in some animal models. The emergence of human induced pluripotent stem cell (hiPSC) technology has provided a new resource for dissecting the complex nature of early eye morphogenesis through the generation of three-dimensional (3D) cellular models. By using patient-specific hiPSCs to generate <em>in vitro </em>optic vesicle-like models, we can enhance the understanding of early developmental eye disorders and provide a pre-clinical platform for disease modelling and therapeutics testing. A major challenge of <em>in vitro </em>optic vesicle generation is the low efficiency of differentiation in 3D cultures. To address this, we adapted a previously published protocol of retinal organoid differentiation to improve embryoid body formation using a microwell plate. Established morphology, upregulated transcript levels of known early eye-field transcription factors and protein expression of standard retinal progenitor markers confirmed the optic vesicle/presumptive optic cup identity of <em>in vitro </em>models between day 20 and 50 of culture. This adapted protocol is relevant to researchers seeking a physiologically relevant model of early human ocular development and disease with a view to replacing animal models.</p>

opencc-by-4.0Feb 2022View details →
zenodo44/100

Generated Data for the Manuscript "Nonideality-Aware Training for Accurate and Robust Low-Power Memristive Neural Networks"

<p>The file contains&nbsp;data generated and referred to in the text and the figures of the manuscript.</p>

opencc-by-4.0Dec 2021View details →
zenodo44/100

Data to generate the figures of: "Symmetry breaking of azimuthal waves: Slow-flow dynamics on the Bloch sphere"

<p>The folder contains the data and scripts to generate all the figures of the paper, with detailed instructions.</p> <p>No experimental data was used for this article.</p>

opencc-by-4.0Mar 2022View details →
zenodo44/100

Data generated and analysed for Santos Neves, Lambert, Valente & Etienne 2022

<p>This repository contains the data and metadata for accompanying the publication of Santos Neves, P., Lambert, J. W., Valente, L., &amp; Etienne, R. S. (2022). The robustness of a simple dynamic model of island biodiversity to geological and sea-level change.</p> <p><br> All files apart from metadata contained within this repository were obtained via computation at University of Groningen Peregrine High Performance Computing Cluster (HPCC).<br> Data was generated using the pipeline implemented on the R package DAISIErobustness, which itself greatly depends on the R package DAISIE. The code for these packages is version controlled on GitHub and is freely available in open-source repositories. See the Related Identifiers section for links to relevant archived versions of both these packages.</p>

opencc-by-4.0May 2022View details →
zenodo44/100

Dataset of 30 energy customers with flexibility data, and distributed generation, considering residential, small commerce, large commerce, and industrial customers

<p>The dataset has 30 customers: ten residential, ten small commerce, five large commerce, and five industrial customers. The combination of several energy customer types allows the creation of a dataset with different types of consumption profiles, generation, and flexibility, and, therefore, different values of participation in demand response events.</p> <p>The residential profiles of the considered customers use the data available in the Working Group on Intelligent Data Mining and Analysis (IDMA): https://site.ieee.org/pes-iss/data-sets/</p> <p>The values represent a week period using 15 minutes reading periods. All the values are expressed in kWh and the matrixes were created as [customer x time_period].</p> <p>&nbsp;</p> <p>We would be grateful if you could acknowledge the use of this dataset in your publications. Please use the Zenodo publication to cite this work.</p>

opencc-by-4.0Jun 2022View details →
zenodo44/100

Data from microphone to measure the noise generated by the mobilefuge

<p>The two datasets uploaded are the measurement of noise generated when the mobilefuge is placed on table with a damping pad or without a damping pad.&nbsp;We found&nbsp;that with the use of the damping pad, the noise recorded in the microphone decreased by 13dB indicating the improved stable operation of the mobilefuge.</p>

opencc-by-4.0Aug 2022View details →
zenodo44/100

Hourly generation and supply data - current mix and future scenarios

<p>Dataset on hourly generation, imports and exports of electricity in Italy for 2018, 2019 and 2020 (current mix) and two future scenarios (2030).</p> <p>Modelling materials and methods are described in the paper &quot;Life-cycle assessment of current and future electricity supply in Italy: addressing average and marginal hourly demand&quot;.</p> <p>&nbsp;</p>

opencc-by-4.0Oct 2022View details →
zenodo44/100

Data for "Accelerating equilibrium spin-glass simulations using quantum annealers via generative deep learning"

<p>Datasets and material for replicating plots and results from the paper &quot;Accelerating equilibrium spin-glass simulations using quantum annealers via generative deep learning&quot; <a href="https://scipost.org/SciPostPhys.15.1.018">SciPost Phys. 15, 018 (2023)</a>.</p> <p>You will find three data&nbsp;files and a ReadMe.txt:</p> <ul> <li><strong>couplings.tar.gz&nbsp;</strong>contains the random couplings of the system&#39;s Hamiltonian&nbsp;<span class="math-tex">\(H = \sum_{\langle ij \rangle}{J_{ij} \sigma_i \sigma_j}\)</span>;</li> <li><strong>datasets.tar.gz&nbsp;</strong>contains all the datasets generated by the&nbsp;<a href="https://www.dwavesys.com/">D-Wave</a>&nbsp;quantum computer. They are already split&nbsp;into train and validation and divided for the type of model and annealing time;</li> <li><strong>data_for_fig.tar.gz&nbsp;</strong>contains files for reproducing the plots of the article, almost all of them are saved in double format, .csv and .npy or .npz.</li> </ul> <p>We encourage you to download the GitHub code linked below to open all the listed data.</p> <p>All the data are zip, so to unzip them using</p> <pre><code class="language-bash">tar -xvf datasets.tar.gz</code></pre> <p>The code for training the Neural Networks and reproducing all the results&nbsp;is open access at <a href="https://doi.org/10.5281/zenodo.7118502">zenodo.7118502</a>.</p>

opencc-by-4.0Oct 2022View details →
zenodo44/100

Data supporting "Transformer Model Generated Bacteriophage Genomes are Compositionally Distinct from Natural Sequences"

<p>Sequence and composition data supporting doi: <a href="https://doi.org/10.1101/2024.03.19.585716" target="_blank" rel="noopener">10.1101/2024.03.19.585716</a>.&nbsp;Uncompressed file size is ~5.8GB.</p> <p>Data in zip files is organized by sequence provenance (generRNA, natural, or transformer (megaDNA)). Common file types between folders include:</p> <ul> <li>Multi-record fasta file: Sequence data for all sequences of a given provenance. For generRNA sequences, these are found within the `seq` column of file "MFE_distribution_Fig4a.csv"</li> <li>Composition files: Individual sequence level compositional metrics for sliding 120 bp windows. Only structural metrics were used in this study.</li> <li>Genomad: Results from the genomad pipeline (https://portal.nersc.gov/genomad/)</li> <li>Stats: Aggregate statistics for all sequences of a given provenance.</li> </ul> <p>The natural folder also has a metadata file detailing the taxonomy for all natural sequences.<br><br>Figure datasets are the cleaned (sometimes aggregated) datasets that underly specific figures in the manuscript. The figure designations are based on the order in: https://www.biorxiv.org/content/10.1101/2024.03.19.585716v1.</p>

opencc-by-4.0May 2024View details →
zenodo44/100

Data for Dodds et al., The direction of core solidification in asteroids: implications for dynamo generation

<p>Numerical dataset for the data presented in Dodds et al., The direction of core solidification in asteroids: implications for dynamo generation, manuscript submitted to Icarus journal.</p>

opencc-by-4.0Jun 2024View details →
zenodo44/100

Data for: Generation of sanitation system options for urban planning considering novel technologies

<p>This data has been used (1) to quantify the appropriateness of a set of sanitation technologies for a small town (Katarnyia) in Nepal and (2) to generate sanitation system options from the appropriate technologies as an input into strategic sanitation planning using a structured decision making process. For (1), the appropriateness is quantified based on a set of criteria, also called screening criteria. These criteria include technical, socio-demographic, climatic, and institutional aspects and are quantified using uncertainty functions in order to account for the quality and quantity of available input information.</p> <p>The data contains raw data as well as modelling results. The raw data is a compilation of information collected from literature, information collected through a household survey in the small town, field observations. They are all used to describe the screening criteria for the studied sanitation technologies and the small town. Results include: (1) the outcome of the technology appropriateness assessment (technology appropriateness scores); and (2) the sanitation system options (all possible sanitation systems built from the appropriate technologies, and a smaller set of divers and highly appropriate sanitation system options as an input into decision-making).</p>

opencc-zeroDec 2017View details →
zenodo44/100

Evaluation Data of the Implementation of the Approach for Automatic Test Generation for Information-Flow Properties

<p>This data set contains the programs for which the automatic test generation approach of the KeY theorem prover was used to automatically generate noninterference tests.</p> <p>The approach is described in <a href="http://dx.doi.org/10.1145/3297280.3297500 ">http://dx.doi.org/10.1145/3297280.3297500&nbsp;</a></p> <p>DATA<br> ---------<br> The data folder contains the secure and insecure programs which were evaluated and the tests which were generated for them.</p> <p>Each program is in the folder &quot;program&quot; and is written in Java and specified in an extended version of the JML specification language. Check out <a href="http://dx.doi.org/10.5445/IR/1000046878">http://dx.doi.org/10.5445/IR/1000046878</a> for a reference on the used specification language.</p> <p>For each example we provide the tests that were generated. For the insecure examples we provide the tests generated with each of the two options of our approach. The tests generated with the option for searching for counterexamples is in the folder &quot;WithPost&quot; of each insecure example.</p> <p>&nbsp;</p>

opencc-by-4.0Jul 2019View details →
zenodo44/100

Data for: Strong bottom currents in large, deep Lake Geneva generated by higher vertical-mode Poincaré waves

<p>Combining entire summer season current and temperature observations and 3D numerical modeling, we demonstrate that previously undetected vertical mode-two and vertical mode-three Poincar&eacute; waves in 309-meter deep Lake Geneva (Switzerland/France) generate strong bottom-boundary layer currents at 300-m depth. The data include measurements from moored Acoustic Doppler Current Profilers (ADCPs), vertical thermistor lines, and the corresponding 3D modeling results. The three-dimensional model used in this study is based on the MIT General Circulation Model (MITgcm, <a href="http://mitgcm.org/">http://mitgcm.org/</a>, <a href="https://doi.org/10.1029/96JC02775">https://doi.org/10.1029/96JC02775</a>). The main MITgcm model configuration files are available online at <a href="https://doi.org/10.5281/zenodo.13144189">https://doi.org/10.5281/zenodo.13144189</a>.</p> <p>The related scientific publication can be found at <a href="https://doi.org/10.1038/s43247-024-01653-8">https://doi.org/10.1038/s43247-024-01653-8</a></p>

opencc-by-4.0Jul 2024View details →
zenodo44/100

Data used to generate figures for Trossman et al. for Phil. Trans. A in 2024

<p>These are .mat files for the first six figures of a manuscript intended for a special edition of Philosophical Transactions A and .data/.meta files in MITgcm format that can be read using rdmds.m for the seventh figure of the same manuscript.</p>

opencc-by-4.0Apr 2024View details →
zenodo44/100

Verification of library complexity in the HEK-Cas9 sublibraries - sequence data of the generated sublibraries A and B

<p>Sequence data of the generated HEK-Cas9 sublibraries A and B, linked to the manuscript 10.1128/mbio.01925-24: The <em>Bordetella</em> effector protein BteA induces host cell death by disruption of calcium homeostasis by Martin Zmuda, Eliska Sedlackova, Barbora Pravdova, Monika Cizkova, Marketa Dalecka, Ondrej Cerny, Tania Romero Allsop, Tomas Grousl, Ivana Malcova, and Jana Kamanova</p>

opencc-by-4.0Nov 2024View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record