Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
3,481
datasets available to search
ShareScore release 0.9.0
Dataset results
3,481 results for “data set”
An exploratory study to examine the effects of maternal choline supplementation during mouse pregnancy on placental epigenetic markers (miRNA-seq data set)
GEO Series GSE111295. Mus musculus. 11 samples. Type: Non-coding RNA profiling by high throughput sequencing.
An exploratory study to examine the effects of maternal choline supplementation during mouse pregnancy on placental epigenetic markers (mRNA-seq data set)
GEO Series GSE111294. Mus musculus. 12 samples. Type: Expression profiling by high throughput sequencing.
Developmental trajectory of pre-hematopoietic stem cell formation from endothelium (scATAC-seq data set)
GEO Series GSE137115. Mus musculus. 1 samples. Type: Genome binding/occupancy profiling by high throughput sequencing.
Protocol matters – reproducibility and rigor of DNA methylation data sets [RNA-seq]
GEO Series GSE164800. Rattus norvegicus. 17 samples. Type: Expression profiling by high throughput sequencing.
Integrating multimodal data sets into a mathematical framework to describe and predict therapeutic resistance in cancer
GEO Series GSE154932. Homo sapiens. 3 samples. Type: Expression profiling by high throughput sequencing.
Early commitment genes drive fast, irreversible commitment to human embryonic stem cell differentiation (RNA-seq data set 1)
GEO Series GSE127935. Homo sapiens. 30 samples. Type: Expression profiling by high throughput sequencing.
Water Transfer Data set
Open the record for dataset details and reuse information.
Visitor patterns for n-ary data sets
Open the record for dataset details and reuse information.
Data from: Development of a gridded meteorological data set over the Java island, Indonesia 1985-2014
Open the record for dataset details and reuse information.
Data from: Feasibility of a blended group intervention (bGT) for major depression: uncontrolled interventional study in a university setting
Open the record for dataset details and reuse information.
Data from: Assessment of a storage system to deliver uninterrupted therapeutic oxygen during power outages in resource-limited settings
Open the record for dataset details and reuse information.
Data from: Assimilating MODIS data-derived minimum input data set and water stress factors into CERES-Maize model improves regional corn yield predictions
Open the record for dataset details and reuse information.
Data from: Scaling up DNA barcoding - primer sets for simple and cost efficient arthropod systematics by multiplex PCR and Illumina amplicon sequencing
Open the record for dataset details and reuse information.
Data from: A nuclear DNA barcode for eastern North American oaks and application to a study of hybridization in an Arboretum setting
Open the record for dataset details and reuse information.
Data from: Fully automated sequence alignment methods are comparable to, and much faster than, traditional methods in large data sets: an example with hepatitis B virus
Open the record for dataset details and reuse information.
Data from: A resource-rational theory of set size effects in human visual working memory
Open the record for dataset details and reuse information.
DATA set for PLOS One Hsieh et al. 2016
Open the record for dataset details and reuse information.
In silico characterization of miRNA and long non-coding RNA interplay in multiple myeloma (30 PCL miRNA data sets)
GEO Series GSE106744. Homo sapiens; synthetic construct. 34 samples. Type: Non-coding RNA profiling by array.
SEQC Toxicogenomics Study: RNA-Seq data set
GEO Series GSE55347. Rattus norvegicus. 116 samples. Type: Expression profiling by high throughput sequencing.
Systematic discovery of uncharacterized transcription factors in Escherichia coli K-12 MG1655 (RNA-seq data set)
GEO Series GSE111094. Escherichia coli str. K-12 substr. MG1655. 28 samples. Type: Expression profiling by high throughput sequencing.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.