Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

103

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

103 results for “regression modeling”

Learn how ShareScore rates datasets ↗
zenodo36/100

Factors Associated with Scientific Production Citations in Dentistry: Zero-inflated Negative Binomial Regression and Hurdle Modelling

<p><strong>Abstracto:</strong> La literatura cient&iacute;fica mundial en odontolog&iacute;a ha mostrado importantes avances en este campo, con importantes contribuciones que van desde el an&aacute;lisis de los aspectos epidemiol&oacute;gicos b&aacute;sicos de la prevenci&oacute;n hasta resultados especializados en el campo de los tratamientos dentales. La presente investigaci&oacute;n tiene como objetivo analizar el estado actual de la literatura cient&iacute;fica sobre odontolog&iacute;a alojada en la base de datos Web of Science. La metodolog&iacute;a incluye dos fases en el an&aacute;lisis de art&iacute;culos y revisiones indexadas en todas las &aacute;reas tem&aacute;ticas. Durante la primera fase, se analizan las siguientes variables: la producci&oacute;n cient&iacute;fica por parte del editor, la evoluci&oacute;n de la producci&oacute;n cient&iacute;fica publicada por los editores, los factores asociados al impacto de la producci&oacute;n cient&iacute;fica y la modelizaci&oacute;n del impacto de la producci&oacute;n cient&iacute;fica en odontolog&iacute;a. Durante la segunda fase, se analizan asociaciones, evoluciones y tendencias en el uso de palabras clave principales en la literatura cient&iacute;fica en odontolog&iacute;a. En conclusi&oacute;n, el estudio muestra que los temas m&aacute;s estudiados incluyen la asociaci&oacute;n de la educaci&oacute;n dental y el plan de estudios, la asociaci&oacute;n de la odontolog&iacute;a pedi&aacute;trica con la salud oral y el cuidado dental. Los hallazgos muestran que tambi&eacute;n destacan temas enfatizados m&aacute;s recientemente, como la odontolog&iacute;a basada en la evidencia, la pandemia, el control de infecciones y la endodoncia, as&iacute; como la necesidad de futuras investigaciones para ampliar el conocimiento actual basado en temas emergentes en la literatura cient&iacute;fica sobre odontolog&iacute;a.</p>

opencc-by-4.0Sep 2023View details →
dryad36/100

Unpacking the "black box": improving ecological interpretation of regression based models

Open the record for dataset details and reuse information.

publicApr 2023View details →
dryad36/100

Hyperspectral reflectance-based partial least squares regression models for predicting cotton leaf physiological traits

Open the record for dataset details and reuse information.

publicSep 2025View details →
dryad36/100

Data for: A new threshold selection method for species distribution models with presence-only data: extracting the mutation point of the P/E curve by threshold regression

Open the record for dataset details and reuse information.

publicMar 2024View details →
dryad36/100

Data for: Accurate sequence-to-affinity models for SH2 domains from multi-round peptide binding assays coupled with free-energy regression

Open the record for dataset details and reuse information.

publicAug 2025View details →
dryad36/100

Data for: Leveraging spatio-temporal genomic breeding value estimates of dry matter yield and herbage quality in ryegrass via random regression models

Open the record for dataset details and reuse information.

publicAug 2023View details →
dryad36/100

Supplementary Materials include results of simulation experiments to investigate the impact of phylogenetic regression with model violations.

Open the record for dataset details and reuse information.

publicMay 2024View details →
dryad36/100

Data from: Improving performance of hurdle models using rare-event weighted logistic regression: An application to maternal mortality data

Open the record for dataset details and reuse information.

publicOct 2022View details →
zenodo32/100

Data for "Function Space Optimization: A symbolic regression method for estimating parameter transfer functions for hydrological models"

<p>This repository contains all geo-physical catchment properties used in the publication &quot;Function Space Optimization: A symbolic regression method for estimating parameter transfer functions for hydrological models&quot;.</p>

opencc-by-4.0Feb 2020View details →
zenodo32/100

Imbalanced regressive neural network model for whistler-mode hiss waves: spatial and temporal evolution

<p>This dataset contains the whistler-mode hiss waves obtained from the Van Allen Probes. It is accompanied by the manuscript "<span>Imbalanced regressive neural network model for whistler-mode hiss waves: spatial and temporal evolution".&nbsp;</span></p>

opencc-by-4.0Apr 2024View details →
zenodo32/100

Multivariate Ordinary Least Squares (OLS) regression-based Seismic Hazard Model Data

<p>This dataset includes earthquake parameters, slab geometry, gravity anomalies, and fault proximities used for seismic hazard modeling in the Makran Subduction Zone (MSZ). Supplementary Table S1 contains earthquake data (location, depth, magnitude), slab properties (depth, dip, thickness, strike), and distances to key faults. Supplementary Table S2 provides intraslab seismicity, slab geometry, trench distances, and gravity data. The data are sourced from the USGS Earthquake Catalog, IRIS, Slab-2 model, GMRT, and other geophysical models.</p>

opencc-by-4.0Nov 2024View details →
zenodo32/100

Dataset for "Modeling Cell Populations Measured By Flow Cytometry With Covariates Using Sparse Mixture of Regressions" in the Annals of Applied Statistics

<p>This is the dataset to be used for the paper in the Annals of Applied Statistics&nbsp;titled:</p> <p><strong>&quot;Modeling Cell Populations Measured By Flow Cytometry With Covariates Using Sparse Mixture of Regressions&quot;</strong></p> <p>Download, unzip and place in the ./<strong>paper-data</strong>&nbsp;directory in the R package&nbsp;repository&nbsp;<a href="https://github.com/sangwon-hyun/flowmix">https://github.com/sangwon-hyun/flowmix</a>. Then, run the code in <strong>./paper-code</strong>&nbsp;to produce the figures and tables.</p>

opencc-by-4.0Apr 2022View details →
zenodo32/100

Compound data sets for support vector machine and regression modeling

<p>Provided are compound data sets used for support vector machine and support vector regression modeling and associated information.</p>

opencc-by-4.0Oct 2018View details →
zenodo32/100

Intravoxel incoherent motion model of diffusion weighted imaging and diffusion kurtosis imaging in differentiating of local colorectal cancer recurrence from scar/fibrosis tissue by multivariate logistic regression analysis

<p>We&nbsp;uploaded&nbsp;mean of diffusion coefficient (MD) and mean of diffusional Kurtosis values of 56 patients related to the manuscript:&nbsp;Fusco, Roberta, Vincenza Granata, Mario Sansone, Robert Grimm, Paolo Delrio, Daniela Rega, Fabiana Tatangelo, Antonio Avallone, Nicola Raiano, Giuseppe Totaro, Vincenzo Cerciello, Biagio Pecori, and Antonella Petrillo. 2020. &quot;Intravoxel Incoherent Motion Model of Diffusion Weighted Imaging and Diffusion Kurtosis Imaging in Differentiating of Local Colorectal Cancer Recurrence from Scar/Fibrosis Tissue by Multivariate Logistic Regression Analysis&quot; Applied Sciences 10, no. 23: 8609. https://doi.org/10.3390/app10238609</p>

opencc-by-4.0Nov 2020View details →
ClinicalTrials.gov32/100

Predicting Postoperative Pulmonary Infection in Elderly Patients Undergoing Major Surgery: a Study Based on Logistic Regression and Machine Learning Models

ClinicalTrials.gov study NCT06491459. IPD Sharing: Not stated. Countries: 1. Publications: 1.

restrictedIPD-UNDECIDEDFeb 2026View details →
dryad32/100

Data from: A comparison of regression methods for model selection in individual-based landscape genetic analysis

Open the record for dataset details and reuse information.

publicSep 2017View details →
dryad32/100

Code from: Testing for normality in regression models: mistakes abound (but may not matter)

Open the record for dataset details and reuse information.

publicMay 2025View details →
zenodo28/100

Datasets used for Multi-omics integration using Deep Learning and other state-of-the-art regression models

<p>This repository link contains the LIHC files that were downloaded using TCGA Assembler 2 and used in the publication for benchmarking DL and other state-of-the-art regression models.</p> <p>The contents are as follows.</p> <p>Gene level CNA , filename= &quot;<a href="https://zenodo.org/api/files/15943ba8-5f3d-4397-8eb2-98ed85693b79/LIHC__genome_wide_snp_6__GeneLevelCNA.txt">LIHC__genome_wide_snp_6__GeneLevelCNA.txt</a>&quot;</p> <p>DNA Methylation data around 1500 bp around TSS (450K) , filename= &quot;<a href="https://zenodo.org/api/files/15943ba8-5f3d-4397-8eb2-98ed85693b79/LIHC_Methylation450__SingleValue__TSS1500__Both.txt">LIHC_Methylation450__SingleValue__TSS1500__Both.tx</a>t&quot;</p> <p>RNASeq&nbsp; data, filename= &quot;<a href="https://zenodo.org/api/files/15943ba8-5f3d-4397-8eb2-98ed85693b79/LIHC_RNASeq__illuminahiseq_rnaseqv2__GeneExp.txt">LIHC_RNASeq__illuminahiseq_rnaseqv2__GeneExp.txt</a>&quot;</p>

opencc-by-4.0Mar 2020View details →
zenodo28/100

Regression committee machine and petrophysical model jointly driven parameter reservoirs prediction from wireline logs for tight sandstone

<pre>This data comes from this study: &quot;Regression committee machine and petrophysical model jointly driven parameter reservoirs prediction from wireline logs for tight sandstone&quot;. It is the intelligent prediction result of porosity, permeability and water saturation of two wells in the Ordos Basin, China</pre>

opencc-by-4.0Aug 2020View details →
dryad28/100

Data from: Improving estimates of environmental change using multilevel regression models of Ellenberg indicator values

Ellenberg indicator values (EIVs) are a widely used metric in plant ecology comprising a semi-quantitative description of species' ecological requirements. Typically, point estimates of mean EIV scores are compared to infer differences in the environmental conditions structuring plant communities – particularly in resurvey studies with no historical environmental data available. However, the use of point estimates as a basis for inference does not take into account variance among species EIVs within sampled plots, and gives equal weighting to means calculated from sites with differing numbers of species. We present a set of multilevel models – fitted with and without group-level predictors – to improve precision and accuracy of site mean EIV scores, and to provide more reliable inference on changing environmental conditions over spatial and temporal gradients in re-visitation studies. We compare multilevel model performance to GLMM's fitted to point estimates of site mean EIVs. We also test the reliability of this method to improve inferences with incomplete species lists in some or all sample sites. Hierarchical modelling led to more accurate and precise estimates of site-level differences in mean EIV scores between time-periods, particularly for datasets with incomplete records of species occurrence. They also revealed directional environmental change within ecological habitat types, which estimates from GLMM's were inadequate to detect. Multilevel models also highlighted a prominent role of hydrological differences as a driver of community change in our case study, which traditional use of EIVs failed to reveal. We have demonstrated that multilevel modelling of EIVs allows for a nuanced estimation of environmental change underlying ecological communities from plant assemblage data, leading to a better understanding of temporal dynamics of ecosystems. Further, the ability of these methods to perform well with missing data should increase the total set of historical data which can be used to this end.

opencc-zeroDec 2017View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record