Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

1,079

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

1,079 results for “source data”

Learn how ShareScore rates datasets ↗
geo20/100

Expression data from HepaRG human progenitor hepatic cells treated with different sources of Vitamin E and wheat germ oil.

GEO Series GSE195619. Homo sapiens. 14 samples. Type: Expression profiling by array.

openGEO-OpenJan 2022View details →
geo20/100

Differential RNA-Seq (dRNA-seq) data from Sphingopyxis granuli strain TFA grown in two different carbon sources and RNA-seq from Hfq-coIP experiment.

GEO Series GSE111181. Sphingopyxis granuli. 6 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenAug 2018View details →
geo20/100

The transcriptome comparison data between Saccharomyces cerevisiae cells in xylose consumption phase after glucose depleted in glucose-xylose co-fermentation and when xylose was the sole carbon source

GEO Series GSE95076. Saccharomyces cerevisiae. 8 samples. Type: Expression profiling by high throughput sequencing.

openGEO-OpenApr 2018View details →
geo20/100

Expression data after irradiating mMSCs for 2 hours with single frequency terahertz laser source

GEO Series GSE41085. Mus musculus. 6 samples. Type: Expression profiling by array.

openGEO-OpenJul 2013View details →
geo20/100

Ileal expression data of mice fed with diet containing protein from various sources

GEO Series GSE84442. Mus musculus. 33 samples. Type: Expression profiling by array.

openGEO-OpenAug 2016View details →
zenodo20/100

Source data for PointNovo

<p>Source data used to generate the analysis in PointNovo paper</p>

opencc-by-4.0Aug 2020View details →
zenodo20/100

CBA source data

<p>CBA: Cluster-guided batch alignment for single-cell RNA-seq</p>

opencc-by-4.0Dec 2020View details →
zenodo20/100

Data for "Uncovering natural sources of LAB bacteriocins through the use of metagenomics"

Open the record for dataset details and reuse information.

opencc-by-4.0Dec 2023View details →
zenodo20/100

Source Data File_Zebrafish

<p>Raw data of zebrafish experiments</p>

opencc-by-4.0Dec 2023View details →
zenodo20/100

Data of water sources rehydration experiments

<p>This is the dataset for manuscrit which submitted to GRL.</p>

opencc-by-4.0Mar 2022View details →
zenodo20/100

Figure Source Data for "Breaking through safety performance stagnation in autonomous vehicles with dense learning"

Open the record for dataset details and reuse information.

openJul 2024View details →
zenodo20/100

Source data for: Direct retino-iridal projections and intrinsic iris contraction mediate the pupillary light reflex in early vertebrates (Communications Biology)

Open the record for dataset details and reuse information.

opencc-by-4.0Jul 2024View details →
zenodo20/100

Rendering Scripts and Source Data for Figures in "Why East Asian Monsoon Anomalies Are More Robust in Post El Niño than in Post La Niña Summers"

Open the record for dataset details and reuse information.

opencc-by-4.0Aug 2024View details →
zenodo20/100

Open source data for microservice partitioning

<p>Open-source datasets for microservice partitioning.</p>

opencc-by-4.0Aug 2021View details →
zenodo20/100

Weakly hydrated anions bind to polymers but not monomers in aqueous solutions - Source data for Figure 2

<p>Source data that was used to create Figure 2.&nbsp;</p>

opencc-by-4.0Dec 2020View details →
zenodo20/100

Source data for "Fast multi-source nanophotonic simulations using augmented partial factorization"

<p>Source data for Fig. 3b-c and Fig. 5a-b</p>

opencc-by-4.0Nov 2022View details →
zenodo20/100

Source data for role of sesquiterpenes in biogenic new particle formation

<p>Data generated from experiments performed within the CLOUD chamber at CERN.&nbsp;</p>

opencc-by-4.0Jul 2023View details →
nasa20/100

Classification of Mars Terrain Using Multiple Data Sources

Classification of Mars Terrain Using Multiple Data Sources Alan Kraut1, David Wettergreen1 ABSTRACT. Images of Mars are being collected faster than they can be analyzed by planetary scientists. Automatic analysis of images would enable more rapid and more consistent image interpretation and could draft geologic maps where none yet exist. In this work we develop a method for incorporating images from multiple instruments to classify Martian terrain into multiple types. Each image is segmented into contiguous groups of similar pixels, called superpixels, with an associated vector of discriminative features. We have developed and tested several classification algorithms to associate a best class to each superpixel. These classifiers are trained using three different manual classifications with between 2 and 6 classes. Automatic classification accuracies of 50 to 80% are achieved in leave-one-out cross-validation across 20 scenes using a multi-class boosting classifier.

restrictednotspecifiedMar 2025View details →
nasa20/100

Comet Data Compilations from Published Sources

The collections in this archive contain results that have been reported in the literature, or in some cases made available by private communication, for various comet properties. Individual collections focus on one or several closely related properties or targets of high interest.

restrictedus-pdApr 2025View details →
nasa20/100

LOFAR 2-Meter Sky Survey Preliminary Data Release Source Catalog

The Low Frequency Array (LOFAR) Two-metre Sky Survey (LoTSS) is a deep 120-168 MHz imaging survey that will eventually cover the entire Northern sky. Each of the 3,170 pointings will be observed for 8 hours, which, at most declinations, is sufficient to produce ~5-arcsec resolution images with a sensitivity of ~0.1 mJy/beam and accomplish the main scientific aims of the survey which are to explore the formation and evolution of massive black holes, galaxies, clusters of galaxies and large-scale structure. Due to the compact core and long baselines of LOFAR, the images provide excellent sensitivity to both highly extended and compact emission. For legacy value, the data are archived at high spectral and time resolution to facilitate sub-arcsecond imaging and spectral line studies. In this paper, The authors provide an overview of the LoTSS. They outline the survey strategy, the observational status, the current calibration techniques, a preliminary data release, and the anticipated scientific impact. The preliminary images that they have released were created using a fully-automated but direction-independent calibration strategy and are significantly more sensitive than those produced by any existing large-area low-frequency survey. In excess of 44,000 sources are detected in the images that have a resolution of 25-arcseconds, typical noise levels of less than 0.5 mJy/beam, and cover an area of 381 square degrees in the region of the HETDEX Spring Field (Right Ascension 10&lt;sup&gt;h&lt;/sup&gt; 45&lt;sup&gt;m&lt;/sup&gt; 00&lt;sup&gt;s&lt;/sup&gt; to 15&lt;sup&gt;h&lt;/sup&gt; 30^m ^00&lt;sup&gt;s&lt;/sup&gt; and Declination +45&lt;sup&gt;o&lt;/sup&gt; 00&#39; 00&quot; to +57&lt;sup&gt;o&lt;/sup&gt; 00&#39; 00&quot;). Source detection on the mosaics that are centered on each pointing was performed with PyBDSM (See &lt;a href="http://www.astron.nl/citt/pybdsm/"&gt;http://www.astron.nl/citt/pybdsm/&lt;/a&gt; for more details). In an effort to minimize contamination from artifacts, the catalog was created using a conservative 7-sigma detection threshold. Furthermore, as the artifacts are predominantly in regions surrounding bright sources, the authors utilized the PyBDSM functionality to decrease the size of the box used to calculate the local noise when close to bright sources, which has the effect of increasing the estimated noise level in these regions. Their catalogs from each mosaic are merged to create a final catalogue of the entire HETDEX Spring Field region. During this process, the authors remove multiple entries for sources by only keeping sources that are detected in the mosaic centered on the pointing to which the source is closest to the center. In the catalog, they provide the type of source, for which they used PyBDSM to distinguish isolated compact sources, large complex sources, and sources that are within an island of emission that contains multiple sources. In addition, they attempted to distinguish between sources that are resolved and unresolved in their images. The authors have provided a preliminary data release from the LOFAR Two-metre Sky Survey (LoTSS). This release contains 44,500 sources which were detected with a signal in excess of seven times the local noise in their 25&quot; resolution images. The noise varies across the surveyed region but is typically below 0.5 mJy/beam and the authors estimate the catalog to be 90% complete for sources with flux densities in excess of 3.9 mJy/beam. This table was created by the HEASARC in February 2017 based on &lt;a href="https://cdsarc.cds.unistra.fr/ftp/cats/J/A+A/598/A104"&gt;CDS Catalog J/A+A/598/A104&lt;/a&gt; file lotss.dat. This is a service provided by NASA HEASARC .

restrictednotspecifiedApr 2025View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record