Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

252

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

252 results for “Synthetic data”

Learn how ShareScore rates datasets ↗
zenodo32/100

Orthogonal light-activated DNA for patterned biocomputing within synthetic cells (Source Data)

<p>Source data for the published version of &quot;Orthogonal light-activated DNA for patterned biocomputing within synthetic cells&quot;: Preprint (https://chemrxiv.org/engage/chemrxiv/article-details/63b55bb6ff4651ef52429534)</p>

opencc-by-4.0Apr 2023View details →
dryad32/100

Data from: Classifying interactions in a synthetic bacterial community is hindered by inhibitory growth medium

<p>Predicting the fate of a microbial community and its member species relies on understanding the nature of their interactions. However, designing simple assays that distinguish between interaction types can be challenging. Here, we performed spent media assays based on the predictions of a mathematical model to decipher the interactions between four bacterial species: <em>Agrobacterium</em> <em>tumefaciens</em> (<em>At</em>), <em>Comamonas</em> <em>testosteroni</em> (<em>Ct</em>), <em>Microbacterium</em> <em>saperdae</em> (<em>Ms</em>) and <em>Ochrobactrum</em> <em>anthropi</em> (<em>Oa</em>). While most experimental results matched model predictions, the behavior of <em>Ct</em> did not: its lag phase was reduced in the pure spent media of <em>At</em> and <em>Ms</em>, but prolonged again when we replenished with our growth medium. Further experiments showed that the growth medium actually delayed the growth of <em>Ct</em>, leading us to suspect that <em>At</em> and <em>Ms</em> could alleviate this inhibitory effect. There was, however, no evidence supporting such "cross-detoxification" and instead, we identified metabolites secreted by <em>At</em> and <em>Ms</em> that were then consumed or "cross-fed" by <em>Ct</em>, shortening its lag phase. Our results highlight that even simple, defined growth media can have inhibitory effects on some species and that such negative effects need to be included in our models. Based on this, we present new guidelines to correctly distinguish between different interaction types, such as cross-detoxification and cross-feeding.</p>

opencc-zeroJun 2023View details →
zenodo32/100

Data for: Engineering tRNA abundances for synthetic cellular systems

<p>Data for publication:<strong>&nbsp;Engineering tRNA abundances for synthetic cellular systems</strong></p> <p><strong>Abstract</strong></p> <p>Routinizing&nbsp;the engineering of synthetic cells requires&nbsp;specifying&nbsp;determining&nbsp;beforehand how many of each molecule are needed. First-principles tools for specifying molecular abundances enabling whole-cell synthetic biology are missing. We use a colloidal dynamics simulator to make predictions for how tRNA abundances impact protein synthesis rates. We use rational design and direct RNA synthesis to make 21 synthetic tRNA surrogates from scratch. We use evolutionary algorithms within a computer aided design framework to&nbsp;design&nbsp;engineer&nbsp;translation systems predicted to work faster or slower depending on tRNA abundance differences. We build and test the so-specified synthetic systems and find&nbsp;that&nbsp;qualitative agreement between&nbsp;expected and observed systems&nbsp;performance matchqualitatively match.&nbsp;&nbsp;First-principles modeling combined with bottom-up experiments can help molecular-to-cellular scale synthetic biology realize &ldquo;design, build, work&rdquo; frameworks that transcend tinker-and-test.</p> <p><strong>Data description</strong></p> <p>The data here consists of (1) All&nbsp;Colloidal Smoldyn &amp; CD-CAD simulation input parameter and output files &amp; (2) experimental data used for the associated publication. Simulation data was produced using Colloidal Dynamics modeling and Colloidal Dyamics-CAD (CD-CAD) as described in the associated manuscript. Data folders should be used directly with modeling and analysis code provided on Github: https://github.com/EndyLab/tRNACAD.</p>

opencc-by-4.0May 2023View details →
zenodo32/100

Modified Fuchs et al. model Synthetic Data Sets

<p>Synthetic data sets used for machine learning of laser acceleration of protons</p>

opencc-by-4.0Aug 2023View details →
zenodo32/100

Statistical error estimation from residual statistics of multiple collocated datasets: Data from synthetic experiments

<p>This archive contains the data used in the paper &quot;How far can the statistical error estimation problem be closed by collocated data?&quot; by A.Vogel and R.Menard (preprint available at: https://doi.org/10.5194/egusphere-2022-996) accepted for publication in Nonlinear Processes in Geophysics (NPG).&nbsp;</p> <p>The data refers to the synthetic experiments in Sect.5 of the paper which demonstrate the general ability to estimate statistical error covariances and cross-statistics from residual covariances, as well as the effects of inaccurate assumptions with respect to different setups.</p> <p>Further information on the data can be found in the README.txt file.</p>

opencc-by-4.0Aug 2023View details →
zenodo32/100

Synthetic fuel scenario data

<p>Scenario data generated by&nbsp;AIM/Technology model for the global synfuel scenario&nbsp;analysis.</p>

opencc-by-4.0Oct 2023View details →
ClinicalTrials.gov32/100

Simulated and Synthetic Health Data: Improving Clinical Research on Rare Diseases. A Real-World Data Simulation of Autosomal Dominant Polycystic Kidney Disease (ADPKD) Trials. A Retrospective, Observa

ClinicalTrials.gov study NCT07016282. IPD Sharing: NO. Countries: 2. Publications: 27.

closedIPD-NOFeb 2026View details →
ClinicalTrials.gov32/100

Development of Synthetic Medical Data Generation Technology to Predict Postoperative Complications

ClinicalTrials.gov study NCT05986474. IPD Sharing: Not stated. Countries: 1. Publications: 1.

restrictedIPD-UNDECIDEDFeb 2026View details →
dryad32/100

Data from: Analyzing negative feedback using a synthetic gene network expressed in the Drosophila melanogaster embryo

Open the record for dataset details and reuse information.

publicAug 2017View details →
dryad32/100

Data from: Two-scale dispersal estimation for biological invasions via synthetic likelihood

Open the record for dataset details and reuse information.

publicJun 2017View details →
dryad32/100

Data from: Classifying interactions in a synthetic bacterial community is hindered by inhibitory growth medium

Open the record for dataset details and reuse information.

publicJun 2023View details →
dryad32/100

Data from: Increases and fluctuations in nutrient availability do not promote dominance of alien plants in synthetic communities of common natives

Open the record for dataset details and reuse information.

publicDec 2021View details →
dryad32/100

Data from: Subgenome dominance in an interspecific hybrid, synthetic allopolyploid, and a 140-year-old naturally established neo-allopolyploid monkeyflower

Open the record for dataset details and reuse information.

publicSep 2017View details →
dryad32/100

Data from: Varying the spatial arrangement of synthetic herbivore-induced plant volatiles and companion plants to improve conservation biological control

Open the record for dataset details and reuse information.

publicFeb 2019View details →
dryad32/100

Data for: Synthetic red supergiant explosion model grid for systematic characterization of Type II supernovae

Open the record for dataset details and reuse information.

publicMar 2023View details →
dryad32/100

Data from: Biomimicry of iridescent, patterned insect cuticles: comparison of biological and synthetic, cholesteric microcells using hyperspectral imaging

Open the record for dataset details and reuse information.

publicJul 2020View details →
dryad32/100

Data from: The effects of synthetic estrogen exposure on pre-mating and post-mating episodes of selection in sex-role-reversed Gulf pipefish

Open the record for dataset details and reuse information.

publicJul 2013View details →
dryad32/100

Data for: Grounding zone of Amery Ice Shelf, Antarctica, from differential synthetic-aperture radar interferometry

Open the record for dataset details and reuse information.

publicJan 2023View details →
dryad32/100

Automatic delineation of glacier grounding lines in differential interferometric synthetic-aperture radar data using deep learning

Open the record for dataset details and reuse information.

publicMar 2021View details →
zenodo28/100

Synthetic Data Set for Uplift Modeling (One Trial)

<p>This dataset is designed and simulated for evaluating uplift modeling and feature selection methods.</p> <p>This dataset contains 10,000 samples and 36 features (one trial).</p> <p>The samples are equally split for control and treatment group.</p> <p>The generated data has three types of features: (1) uplift features influencing the treatment effect on the conversion probability; (2) classification features affecting the conversion probability but independent of the treatment effect; and (3) irrelevant features that are independent of both conversion probability and the treatment effect. To model the relationship between uplift features and the treatment effect and classification features and outcome probability, we implement six types of association patterns in the data generation process: linear, quadratic, cubic, ReLU (Rectified Linear Unit), trigonometric function sine, and cosine.</p> <p>In this data set, there are 36 features in total, including 10 classification features, 6 uplift features, and 20 irrelevant features.</p> <p>Column names:</p> <ul> <li>Experiment group label: &#39;treatment_group_key&#39;</li> <li>Feature names: [&#39;x1_informative&#39;,<br> &#39;x2_informative&#39;,<br> &#39;x3_informative&#39;,<br> &#39;x4_informative&#39;,<br> &#39;x5_informative&#39;,<br> &#39;x6_informative&#39;,<br> &#39;x7_informative&#39;,<br> &#39;x8_informative&#39;,<br> &#39;x9_informative&#39;,<br> &#39;x10_informative&#39;,<br> &#39;x11_irrelevant&#39;,<br> &#39;x12_irrelevant&#39;,<br> &#39;x13_irrelevant&#39;,<br> &#39;x14_irrelevant&#39;,<br> &#39;x15_irrelevant&#39;,<br> &#39;x16_irrelevant&#39;,<br> &#39;x17_irrelevant&#39;,<br> &#39;x18_irrelevant&#39;,<br> &#39;x19_irrelevant&#39;,<br> &#39;x20_irrelevant&#39;,<br> &#39;x21_irrelevant&#39;,<br> &#39;x22_irrelevant&#39;,<br> &#39;x23_irrelevant&#39;,<br> &#39;x24_irrelevant&#39;,<br> &#39;x25_irrelevant&#39;,<br> &#39;x26_irrelevant&#39;,<br> &#39;x27_irrelevant&#39;,<br> &#39;x28_irrelevant&#39;,<br> &#39;x29_irrelevant&#39;,<br> &#39;x30_irrelevant&#39;,<br> &#39;x31_uplift_increase&#39;,<br> &#39;x32_uplift_increase&#39;,<br> &#39;x33_uplift_increase&#39;,<br> &#39;x34_uplift_increase&#39;,<br> &#39;x35_uplift_increase&#39;,<br> &#39;x36_uplift_increase&#39;]</li> <li>Outcome variable: &nbsp;&#39;conversion&#39;</li> <li>True underlying control conversion probability: &#39;control_conversion_prob&#39;</li> <li>True underlying treatment conversion probability: &#39;treatment1_conversion_prob&#39;</li> <li>True treatment effect: &nbsp;&#39;treatment1_true_effect&#39;</li> <li>Note columns names with &#39;_transformed&#39; suffix are feature variables used in the intermediate steps during the data generation, that should be excluded for model training.</li> </ul>

opencc-by-4.0Feb 2020View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record