Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
1,773
datasets available to search
ShareScore release 0.7.1
Dataset results
1,773 results for “predictive modeling”
'CellTrajectory' for cellular automata modelling of leukaemic stem cell dynamics in acute myeloid leukaemia: insights into predictive outcomes and targeted therapies
Open the record for dataset details and reuse information.
Airflow modelling predicts seabird breeding habitat across islands
Open the record for dataset details and reuse information.
Data from: Predicting primate-parasite associations with exponential random graph models
Open the record for dataset details and reuse information.
Data for: Predicting age and mass at maturity from feeding behavior and diet in M. sexta: An empirical test of a life history model
Open the record for dataset details and reuse information.
Prediction of the potentially suitable areas of Leonurus japonicus with the optimized MaxEnt model
Open the record for dataset details and reuse information.
UltraScan Solution Modeler (US-SOMO) hydrodynamic parameter, structural small angle scattering and SESCA circular dichroism (CD) calculations on AlphaFold predicted structures
Open the record for dataset details and reuse information.
Data from: Integrated species distribution models to account for sampling biases and improve range wide occurrence predictions
Open the record for dataset details and reuse information.
Predicting daily activity time through ecological niche modeling and microclimatic data
Open the record for dataset details and reuse information.
'Biosim' for cellular automata modelling of leukaemic stem cell dynamics in acute myeloid leukaemia: insights into predictive outcomes and targeted therapies
Open the record for dataset details and reuse information.
Main model fits and substitution rate predictions for: A quantitative genetic model of background selection in humans
Open the record for dataset details and reuse information.
Data from: Integrating genomic data and simulations to evaluate alternative species distribution models and improve predictions of glacial refugia and future responses to climate change
Open the record for dataset details and reuse information.
Code and data for Bayesian joint species distribution model selection for community-level prediction
Open the record for dataset details and reuse information.
Modelling the carbon balance in bryophytes and lichens: Presentation of PoiCarb 1.0, a new model for explaining distribution patterns and predicting climate-change effects
Open the record for dataset details and reuse information.
Predictive models for secondary Epilepsy in patients with acute Ischemic Stroke within one year
Open the record for dataset details and reuse information.
Assessing predictive performance of supervised machine learning algorithms for a diamond pricing model
Open the record for dataset details and reuse information.
Extending Grime’s CSR model to predict plant demographic responses across resource availability gradients: evidence from the Patagonian steppes
Open the record for dataset details and reuse information.
On the cross-population generalizability of gene expression prediction models
Open the record for dataset details and reuse information.
Vertebrate-habitat relationships: Logistic regression models predict probability of occurrence of bird and small mammal species in western Oregon
Logistic regression models predicting probability of occurrence of bird and of small-mammal species were produced using animal-habitat data sets from throughout western Oregon (Garman and Cole 1999 - Vertebrate Habitat Relationships Data Bank (VHRDB), Report to Coastal Landscape Analysis and Modeling Study). Regression coefficients, variables, and metrics related to model predictions are provided here under Entity 1, and in VHRDB as VERTLOGR.
GTEx v7 prediction models
<p>PrediXcan's prediction models on gene expression from GTEx v7.</p> <p>Also contains LD compilations for S-PrediXcan and S-MultiXcan.</p>
Reliably predicting pollinator abundance: challenges of calibrating process-based ecological models
<p>1. Pollination is a key ecosystem service for global agriculture but evidence of pollinator population declines is growing. Reliable spatial modelling of pollinator abundance is essential if we are to identify areas at risk of pollination service deficit and effectively target resources to support pollinator populations. Many models exist which predict pollinator abundance but few have been calibrated against observational data from multiple habitats to ensure their predictions are accurate.</p> <p>2. We selected the most advanced process-based pollinator abundance model available and calibrated it for bumblebees and solitary bees using survey data collected at 239 sites across Great Britain. We compared three versions of the model: one parameterised using estimates based on expert opinion, one where the parameters are calibrated using a purely data-driven approach and one where we allow the expert opinion estimates to inform the calibration process.</p> <p>3. All three model versions showed significant agreement with the survey data, demonstrating this model's potential to reliably map pollinator abundance. However, there were significant differences between the nesting/floral attractiveness scores obtained by the two calibration methods and from the original expert opinion scores.</p> <p>4. Our results highlight a key universal challenge of calibrating spatially-explicit, process-based ecological models. Notably, the desire to reliably represent complex ecological processes in finely mapped landscapes necessarily generates a large number of parameters, which are challenging to calibrate with ecological and geographical data that is often noisy, biased, asynchronous and sometimes inaccurate. Purely data-driven calibration can therefore result in unrealistic parameter values, despite appearing to improve model-data agreement over initial expert opinion estimates. We therefore advocate a combined approach where data-driven calibration and expert opinion are integrated into an iterative Delphi-like process, which simultaneously combines model calibration and credibility assessment. This may provide the best opportunity to obtain realistic parameter estimates and reliable model predictions for ecological systems with expert knowledge gaps and patchy ecological data.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.