Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
127
datasets available to search
ShareScore release 0.9.0
Dataset results
127 results for “biomarker discovery”
Discovery of South African plant-based biomarkers as potential flagships for SARS- CoV-2 receptor
<h1><span>Table S1: </span><span>Distinguished metabolites in the extracts of <em><span>Artemisia annua </span></em><span>and <em>Artemisia afra </em></span>using UPLC-MS/MS set in positive ionization mode.</span></h1> <h1><span>Table S2: Identified and docked compound-based biomarkers from <em><span>Artemisia annua </span></em><span>and <em>Artemisia afra </em></span>(ESI+ scan).</span></h1>
Selected articles from the scoping review of biomarker discovery studies for the EU project on "Personalised Medicine Trials" (PERMIT)
<p>This dataset provides the data extracted for the scoping review of the literature on biomarker discovery studies for patient stratification using machine learning analysis of omics data, as part of the EU project on “Personalised Medicine Trials” (PERMIT). It covers the references for all articles selected as part of the scoping review, as well as information on the study type and methodology, the outcome measures, the validation type, and representative sentences extracted from each article on the main results and key findings of the corresponding biomarker study.</p>
Discovery of sparse, reliable omic biomarkers with Stabl
<p><span>Adoption of high-content omic technologies in clinical studies, coupled with computational </span><span>methods, have yielded an abundance of candidate biomarkers. However, translating such find</span><span>ings into bona fide clinical biomarkers remains challenging.</span> <span>To facilitate this process, we </span><span>introduce Stabl, a general machine learning framework that identifies a sparse, reliable set </span><span>of biomarkers by integrating noise injection and a data-driven signal-to-noise threshold into </span><span>multivariable predictive modeling.</span> <span>Evaluation of Stabl on synthetic datasets and five inde</span><span>pendent clinical studies demonstrates improved biomarker sparsity and reliability compared to </span><span>commonly used sparsity-promoting regularization methods while maintaining predictive per</span><span>formance; it distills datasets containing 1,400 to 35,000 features down to 4 to 34 candidate </span><span>biomarkers. Stabl extends to multi-omic integration tasks, enabling biological interpretation of </span><span>complex predictive models, as it hones in on a shortlist of proteomic, metabolomic, and cyto</span><span>metric events predicting labor onset, microbial biomarkers of preterm birth, and a pre-operative </span><span>immune signature of post-surgical infections.</span></p>
Discovery of sparse, reliable omic biomarkers with Stabl
Open the record for dataset details and reuse information.
Preprocessing scripts and data for study: Pathway-Based Subnetworks Enable Cross-Disease Biomarker Discovery
<p>Supplementary File 1 containing preprocessing scripts and data for converting pathway databases into subnetworks.</p>
Fruit and Vegetable Biomarker Discovery
ClinicalTrials.gov study NCT05621863. IPD Sharing: YES. Countries: 1. Publications: 1.
Integrated Discovery of New Immuno-Molecular Actionable Biomarkers for Tumors With Immune-suppressed Environment
ClinicalTrials.gov study NCT03706625. IPD Sharing: YES. Countries: 1. Publications: 0.
Data from: Profiling extracellular long RNA transcriptome in human plasma and extracellular vesicles for biomarker discovery
<p>The recent discovery of extracellular RNAs in blood, including RNAs in extracellular vesicles (EVs), combined with low-input RNA-sequencing advances have enabled scientists to investigate their role in human disease. To date, most studies have been focusing on small RNAs, and methodologies to optimize long RNAs measurement are lacking. We used plasma RNA to assess the performance of six long RNA sequencing methods, at two different sites, and we report their differences in reads (%) mapped to the genome/transcriptome, number of genes detected, long RNA transcript diversity, and reproducibility. Using the best performing method, we further compare the profile of long RNAs in the EV- and no-EV-enriched RNA plasma compartments. These results provide insights on the performance and reproducibility of commercially available kits in assessing the landscape of long RNAs in human plasma and different extracellular RNA carriers that may be exploited for biomarker discovery.</p>
Multi-omics data for ischemic stroke etiology biomarker discovery
<p>Multi-omics data for ischemic stroke etiology biomarker discovery</p>
Fox Investigation for New Discovery of Biomarkers
ClinicalTrials.gov study NCT01705327. IPD Sharing: Not stated. Countries: 1. Publications: 1.
Omic Technologies Applied to the Study of B-cell Lymphoma for the Discovery of Diagnostic and Prognosis Biomarkers
ClinicalTrials.gov study NCT05834426. IPD Sharing: NO. Countries: 1. Publications: 7.
Targeted Therapy With or Without Nephrectomy in Metastatic Renal Cell Carcinoma: Liquid Biopsy for Biomarkers Discovery
ClinicalTrials.gov study NCT02535351. IPD Sharing: Not stated. Countries: 1. Publications: 9.
A Longitudinal Multi-Center Molecular Biomarker Discovery Registry for Patients With Hematologic Malignancies
ClinicalTrials.gov study NCT07154823. IPD Sharing: Not stated. Countries: 1. Publications: 12.
Non-invasive Biomarker Discovery for Pre-cervical or/and Cervical Cancer-HPV DNA and Other Biomarkers in Urine
ClinicalTrials.gov study NCT06261892. IPD Sharing: NO. Countries: 1. Publications: 4.
Biomarker Discovery in Patients With Advanced Biliary Tract Cancer
ClinicalTrials.gov study NCT04871321. IPD Sharing: NO. Countries: 1. Publications: 8.
Deep Phenotyping of Hearing Instability Disorders: Cohort Establishment, Biomarker Identification, Development of Novel Phenotyping Measures, and Discovery of Therapeutic Targets
ClinicalTrials.gov study NCT04806282. IPD Sharing: Not stated. Countries: 1. Publications: 1.
TIDE Project: Biomarker Discovery for Chronic Tinnitus Diagnosis
ClinicalTrials.gov study NCT06520865. IPD Sharing: UNDECIDED. Countries: 5. Publications: 0.
Non-invasive Biomarker Discovery for Pre-cervical or/and Cervical Cancer--ACTN4 and Other Biomarkers in Menstrual Blood
ClinicalTrials.gov study NCT06261879. IPD Sharing: NO. Countries: 1. Publications: 4.
Metabolomics for Biomarker Discovery in Children With EoE
ClinicalTrials.gov study NCT03107819. IPD Sharing: NO. Countries: 1. Publications: 16.
Discovery and Validate of Multi-genetic Biomarkers for Capecitabine in Chinese Colorectal Patients
ClinicalTrials.gov study NCT03030508. IPD Sharing: NO. Countries: 1. Publications: 11.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.