Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
1,943
datasets available to search
ShareScore release 0.9.0
Dataset results
1,943 results for “machine learning”
Health Status, Infection and Disease in California Sea Lions (Zalophus californianus) studied using a canine microarray platform and machine-learning approaches
GEO Series GSE29497. Zalophus californianus; Canis lupus familiaris. 73 samples. Type: Expression profiling by array.
SDePER: a hybrid machine learning and regression method for cell-type deconvolution of spatial barcoding-based transcriptomic data
GEO Series GSE231385. Homo sapiens. 1 samples. Type: Expression profiling by high throughput sequencing.
Improved Prediction of Venetoclax and Azacitidine Efficacy in Acute Myeloid Leukemia through Machine Learning [AML patients]
GEO Series GSE289786. Homo sapiens. 142 samples. Type: Expression profiling by high throughput sequencing.
Predicting gene level sensitivity to JAK-STAT signaling perturbation using a mechanistic-to-machine learning framework II
GEO Series GSE231344. Mus musculus. 20 samples. Type: Expression profiling by high throughput sequencing.
Machine Learning Classifiers for Endometriosis Using Transcriptomics and Methylomics Data [Transcriptomics]
GEO Series GSE134056. Homo sapiens. 38 samples. Type: Expression profiling by high throughput sequencing.
Machine Learning Analysis of Circulating Alternative Spliced RNAs Reveal Segregation of Retained Intron Species with New-onset Type 1 Diabetes
GEO Series GSE277133. Homo sapiens. 48 samples. Type: Other.
Machine learning analysis to identify endometrial transcriptomic biomarkers predictive of pregnancy success following artificial insemination in dairy cows
GEO Series GSE248266. Bos taurus. 193 samples. Type: Expression profiling by high throughput sequencing.
EpiGe: A cytosine methyl-genotyping PCR-based machine-learning tool for rapid classification of medulloblastoma
GEO Series GSE210723. Homo sapiens. 74 samples. Type: Methylation profiling by array.
Improved prediction of bacterial CRISPRi guide efficiency from depletion screens through mixed-effect machine learning and data integration
GEO Series GSE196911. Escherichia coli. 8 samples. Type: Other.
Signatures of GVHD and Relapse after Post-Transplant Cyclophosphamide Revealed by Immune Profiling and Machine Learning [17941R-18740R]
GEO Series GSE182612. Homo sapiens. 38 samples. Type: Expression profiling by high throughput sequencing.
Multi-omics and Machine Learning Accurately Predicts Clinical Response to Adalimumab and Etanercept Therapy in Patients with Rheumatoid Arthritis [RNA-Seq]
GEO Series GSE138746. Homo sapiens. 240 samples. Type: Expression profiling by high throughput sequencing.
2015 Urban Extents from VIIRS and MODIS for the Continental U.S. Using Machine Learning Methods
The 2015 Urban Extents from VIIRS and MODIS for the Continental U.S. Using Machine Learning Methods data set models urban settlements in the Continental United States (CONUS) as of 2015. When applied to the combination of daytime spectral and nighttime lights satellite data, the machine learning methods achieved high accuracy at an intermediate-resolution of 500 meters at large spatial scales. The input data for these models were two types of satellite imagery: Visible Infrared Imaging Radiometer Suite (VIIRS) Nighttime Light (NTL) data from the Day/Night Band (DNB), and Moderate Resolution Imaging Spectroradiometer (MODIS) corrected daytime Normalized Difference Vegetation Index (NDVI). Although several machine learning methods were evaluated, including Random Forest (RF), Gradient Boosting Machine (GBM), Neural Network (NN), and the Ensemble of RF, GBM, and NN (ESB), the highest accuracy results were achieved with NN, and those results were used to delineate the urban extents in this data set.
Methylation profiling and machine learning distinguishes primary lung squamous cell carcinomas from head and neck metastases
GEO Series GSE124052. Homo sapiens. 61 samples. Type: Methylation profiling by array.
Using integrative bioinformatics approaches and machine‑learning strategies to Identify potential signatures for atrial fibrillation
GEO Series GSE282504. Homo sapiens. 30 samples. Type: Expression profiling by high throughput sequencing.
Predicting gene level sensitivity to JAK-STAT signaling perturbation using a mechanistic-to-machine learning framework
GEO Series GSE231345. Mus musculus. 60 samples. Type: Expression profiling by high throughput sequencing.
A machine learning approach to integrate big data for precision medicine in acute myeloid leukemia [array]
GEO Series GSE107465. Homo sapiens. 30 samples. Type: Expression profiling by array.
An unbiased machine learning exploration reveals gene sets predictive of allograft tolerance after kidney transplantation
GEO Series GSE166865. Homo sapiens. 0 samples. Type: Expression profiling by array; Third-party reanalysis.
Machine learning predicts individual cancer patient responses to therapeutic drugs with high accuracy
GEO Series GSE112798. Homo sapiens. 31 samples. Type: Expression profiling by array.
Generating automated kidney transplant biopsy reports combining molecular measurements with ensembles of machine learning classifiers
GEO Series GSE124203. Homo sapiens. 1745 samples. Type: Expression profiling by array.
Open Chromatin Guided Interpretable Machine Learning Reveals Cancer-Specific Chromatin Features in Cell-free DNA
GEO Series GSE279542. Homo sapiens. 20 samples. Type: Genome binding/occupancy profiling by high throughput sequencing.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.