Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

119

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

119 results for “missing data”

Learn how ShareScore rates datasets ↗
dryad32/100

Data from: Out-of-sample predictions from plant–insect food webs: robustness to missing and erroneous trophic interaction records

Open the record for dataset details and reuse information.

publicFeb 2015View details →
dryad32/100

Data from: Regional heritability mapping method helps explain missing heritability of blood lipid traits in isolated populations

Open the record for dataset details and reuse information.

publicNov 2015View details →
dryad32/100

Data from: Effects of growth rate, size, and light availability on tree survival across life stages: a demographic analysis accounting for missing values and small sample sizes

Open the record for dataset details and reuse information.

publicMar 2015View details →
dryad32/100

Data from: HIV testing in a South African emergency department: a missed opportunity

Open the record for dataset details and reuse information.

publicFeb 2019View details →
dryad32/100

Data from: Genetic differentiation of Alaska Chinook salmon: the missing link for migratory studies

Open the record for dataset details and reuse information.

publicDec 2010View details →
dryad32/100

Data from: Nest inheritance is the missing source of direct fitness in a primitively eusocial insect

Open the record for dataset details and reuse information.

publicNov 2011View details →
dryad32/100

Data from: A hierarchical Bayesian approach for handling missing classification data

Open the record for dataset details and reuse information.

publicMar 2019View details →
dryad32/100

Data from: Missed opportunities for HIV testing among patients newly presenting for HIV care at a Swiss university hospital: a retrospective analysis

Open the record for dataset details and reuse information.

publicMay 2018View details →
dryad32/100

Data from: Resolving the mesoscopic missing link: biophysical modeling of EEG from cortical columns in primates

Open the record for dataset details and reuse information.

publicSep 2022View details →
dryad32/100

Data from: Rewilded mammal assemblages reveal the missing ecological functions of granivores

Open the record for dataset details and reuse information.

publicJul 2018View details →
dryad32/100

Data from: Phenotypic selection favors missing trait combinations in coexisting annual plants

Open the record for dataset details and reuse information.

publicApr 2013View details →
dryad32/100

Repositories for taxonomic data: Where we are and what is missing

Open the record for dataset details and reuse information.

publicJun 2021View details →
dryad32/100

Data from: Misconceptions on missing data in RAD-seq phylogenetics with a deep-scale example from flowering plants

Open the record for dataset details and reuse information.

publicOct 2016View details →
dryad32/100

Data from: Correcting for missing and irregular data in home-range estimation

Open the record for dataset details and reuse information.

publicJan 2018View details →
dryad32/100

Data from: Missing the people for the trees: identifying coupled natural-human system feedbacks driving the ecology of Lyme disease

Open the record for dataset details and reuse information.

publicOct 2018View details →
dryad32/100

Missing data in sea turtle population monitoring: a Bayesian statistical framework accounting for incomplete sampling

Open the record for dataset details and reuse information.

publicJun 2022View details →
dryad32/100

Data from: RADcap: sequence capture of dual-digest RADseq libraries with identifiable duplicates and reduced missing data

Open the record for dataset details and reuse information.

publicJul 2016View details →
dryad32/100

Nonrandom missing data can bias PCA inference of population genetic structure

Open the record for dataset details and reuse information.

publicAug 2021View details →
zenodo28/100

GECCO Industrial Challenge 2015 Dataset: A heating system dataset for the 'Recovering missing information in heating system operating data' competition at the Genetic and Evolutionary Computation Conference 2015, Madrid, Spain

<p>Dataset &nbsp;of the &#39;Industrial Challenge: Recovering missing information in heating system operating data&#39; competition hosted at&nbsp;The Genetic and Evolutionary Computation Conference (GECCO)&nbsp;July 11th-15th 2015, Madrid, Spain</p> <p>&nbsp;</p> <p>The task of the&nbsp;competition was&nbsp;to recover (impute) missing information in heating system operation time series&#39;.</p> <p>&nbsp;</p> <p>Included in zenodo:&nbsp;</p> <p>- dataset of heating system operational time series with missing values</p> <p>- additional material and descriptions provided for the competition</p> <p>&nbsp;</p> <p>The competition was organized by:</p> <p>M. Friese, A. Fischbach, C. Schlitt, T. Bartz-Beielstein (TH K&ouml;ln)</p> <p>&nbsp;</p> <p>The dataset was provided&nbsp;by:</p> <p>Major German heating systems supplier (S. Moritz)</p> <p>&nbsp;</p> <p>&nbsp;</p> <p>Industrial Challenge: Recovering missing information in heating system operating data</p> <p>&nbsp;</p> <p>The Industrial Challenge will be held in the competition session at the Genetic and Evolutionary Computation Conference. It poses difficult real-world problems provided by industry partners from various fields. Highlights of the Industrial Challenge include interesting problem domains, real-world data and realistic quality measurement</p> <p>Overview</p> <p>In times of accelerating climate change and rising energy costs, increasing energy efficiency and reducing expenses becomes a high priority goal for businesses and private households alike. Modern heating systems record detailed operating data and report this data to a central system. Here, the operating data can be correlated and analyzed to detect potential optimization opportunities or anomalies like unusually high energy consumption. Due to various difficulties this data might be incomplete which makes accurate forecasting even harder.</p> <p>Goal of the GECCO 2015 Industrial Challenge is to develop capable procedures to recover missing information in heating system operating data. Adequate recovery of the missing data enables more accurate forecastings which allow for intelligent control of the heating systems, and therefore contributes to a positive energy balance and reduced expenses.</p> <p>&nbsp;</p> <p><strong>Submission deadline:</strong><br> June 22, 2015</p> <p><strong>Official Webpage:</strong><br> <a href="http://www.spotseven.de/gecco-challenge/gecco-challenge-2015/">www.spotseven.de/gecco-challenge/gecco-challenge-2015/</a></p> <p>&nbsp;</p>

opencc-by-4.0Apr 2015View details →
dryad28/100

Pleistocene persistence and expansion in tarantulas on the Colorado Plateau and the effects of missing data on phylogeographical inferences from RADseq

Few phylogeographical studies exist for taxa inhabiting the Colorado Plateau province. We combined mitochondrial and genomic data with species distribution modeling to test Pleistocene hypotheses for <i>Aphonopelma marxi</i>, a large tarantula endemic to the plateau region. Mitochondrial and genomic analyses revealed that the species comprises at least three main clades that diverged in the Pleistocene. A clade distributed along the Mogollon Rim appears to have persisted in place during the last glacial maximum, whereas the other two clades probably colonized the central and northeastern portion of the species' range from small refugial areas along river-carved canyons. Climate models support this hypothesis for the Mogollon Rim, but late glacial climate data appear too coarse to detect suitable areas in canyons. Locations of canyon refugia could not be inferred from genomic analyses due to missing data, encouraging us to explore the effect of missing loci in phylogeographical inferences using RADseq. In phylogenetic analyses, node support for major clades decreased with the addition of samples with significant amounts of missing data (more than 30%). Population genomic structure was greatly influenced by missing data, with the group membership of many taxa changing as samples with missing loci were added. Results from DAPC, a distance-based method, did not change as samples with significant amounts missing data were added. We conclude that the specific loci that are missing matters more than the number of missing loci, and that samples with missing data can still add information to RADseq-based analyses as long as results are interpreted cautiously.

opencc-zeroAug 2020View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record