Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
1,221
datasets available to search
ShareScore release 0.9.0
Dataset results
1,221 results for “Aggregators”
National Aggregates of Geospatial Data Collection: Population, Landscape, And Climate Estimates, Version 2 (PLACE II)
The National Aggregates of Geospatial Data Collection: Population, Landscape, And Climate Estimates, Version 2 (PLACE II) data set contains estimates of national-level aggregations of territorial extent and population size by biome, climate zone, coastal proximity zone, elevation zone, and population density zone, a compendium of nearly 300 variables for 228 countries. This data set is produced by the Columbia University Center for International Earth Science Information Network (CIESIN).
Population Exposure Estimates in Proximity to Nuclear Power Plants, Country-Level Aggregates
The Population Exposure Estimates in Proximity to Nuclear Power Plants, Country-Level Aggregates data set consists of country-level estimates of total, urban, and rural populations and land area, country-wide, that are in proximity to a nuclear power plant. This data set was created using a global data set of point locations of nuclear power plants, with buffer zones at 30km, 75km, 150km, 300km, 600km, and 1200km, and the Global Population Count Grid Time Series Estimates, Version 1 to estimate the population within each buffer zone for the years 1990, 2000, and 2010. Global Rural-Urban Mapping Project, Version 1 (GRUMPv1) Land and Geographic Unit Area Grids were used to estimate land area within each buffer zone. The GRUMPv1 Urban Extents Grid was used to further delineate population and land area estimates within urban and rural areas. All grids used for population, land area, and urban mask were of 1 km (30 arc-second) resolution.
Lunar Orbiter Laser Altimeter (LOLA) one-way Laser Ranging Full Rate Data (all ranges collected, ground stations, aggregate of normal points daily) from NASA CDDIS
Lunar Orbiter Laser Altimeter (LOLA) one-way laser ranging full rate data. These files contain the full rate data (all ranges collected) as delivered from the ground stations participating in one way ranging. Each file is an aggregate of full rate data collected for every station on a particular day. Note that this does not constitute the official data delivered by the LOLA mission; for these data, please visit the LOLA Planetary Data System listed in the reference. The ground station only data may be useful for those who wish to do their own transmit-receive pairing from onboard spacecraft data.
National Aggregates of Geospatial Data Collection: Population, Landscape, And Climate Estimates (PLACE)
The National Aggregates of Geospatial Data Collection: Population, Landscape, And Climate Estimates (PLACE) data set contains estimates of national-level aggregations of territorial extent and population size by biome, climate zone, coastal proximity, elevation and slope, a compendium of nearly 300 variables for 222 countries. This data set is produced by the Columbia University Center for International Earth Science Information Network (CIESIN).
Cyanobacteria Aggregated Manual Labels
Continuous monitoring for cyanobacteria blooms in small, inland water bodies via in-situ sampling and analysis can be challenging not only due to the number and locations of water bodies to cover, but also due to the dynamic nature of algal growth and toxin production. Detection targets vary with cyanobacteria strains as well as physical, chemical, and biological factors. Ground monitoring also lacks consistency as sampling methods, frequency, and analytical techniques vary from region to region. However, remote sensing allows systematic data collection over a large area to identify regions with potential harmful algal growth. We introduce the Cyanobacteria Aggregated Manual Labels (CAML), a large dataset of in-situ cyanobacteria measurements for investigations of cyanobacteria detection and severity classification in inland water bodies across the United States. Relevant satellite imagery from publicly available endpoints are applicable to use when applying the CAML dataset to models. The dataset labels ground measurements of cyanobacteria cell counts at 23,570 points in U.S. inland water bodies over 2013 2021. Algorithms trained on this data could be used to estimate cyanobacteria cell counts in water bodies for timely water quality and public health interventions and to gain an understanding of environmental and anthropogenic factors associated with cyanobacteria incidence and proliferation. Data is provided in a comma-separated values (CSV) format.
Expression data from follicular lymphoma cells cultured either in suspension either as Multicellular aggregates of lymphoma cells (MALC)
GEO Series GSE41851. Homo sapiens. 6 samples. Type: Expression profiling by array.
FUS/TLS acts as an aggregation-dependent modifier of polyglutamine disease model mice (I)
GEO Series GSE80004. Mus musculus. 12 samples. Type: Expression profiling by array.
An atlas of amyloid aggregation: the impact of substitutions, insertions, deletions and truncations on amyloid beta fibril nucleation
GEO Series GSE193837. Saccharomyces cerevisiae. 30 samples. Type: Other.
RNAseq of rabbit lungs infected with Mycobacterium tuberculosis as single cells or aggregates
GEO Series GSE176139. Oryctolagus cuniculus. 9 samples. Type: Expression profiling by high throughput sequencing.
RNA splicing and aggregate gene expression differences between lung squamous cell carcinoma from patients of West African and European ancestry
GEO Series GSE137291. Homo sapiens. 40 samples. Type: Expression profiling by array.
Sox11 induces cell aggregation and growth suppression in early progenitor B-cells
GEO Series GSE108419. Mus musculus. 16 samples. Type: Expression profiling by array.
Global expression of mouse hepatocytes cultured as monolayers or as three-dimensional aggregates
GEO Series GSE53054. Mus musculus. 15 samples. Type: Expression profiling by array.
FUS/TLS acts as an aggregation-dependent modifier of polyglutamine disease model mice
GEO Series GSE80109. Mus musculus. 24 samples. Type: Expression profiling by array.
FUS/TLS acts as an aggregation-dependent modifier of polyglutamine disease model mice (II)
GEO Series GSE80093. Mus musculus. 12 samples. Type: Expression profiling by array.
A multidimensional analysis reveals distinct immune phenotypes and tertiary lymphoid structure-like aggregates in pediatric acute myeloid leukemia.
GEO Series GSE248597. Homo sapiens. 143 samples. Type: Other.
Engineering Human Bone Marrow-derived Mesenchymal Stromal Cell Aggregates for Enhanced Extracellular Vesicle Secretion in a Vertical-Wheel Bioreactor
GEO Series GSE316471. Homo sapiens. 6 samples. Type: Non-coding RNA profiling by high throughput sequencing.
Control of expression noise, fitness and gene aggregation by a component of the yeast pheromone response pathway
GEO Series GSE17583. Saccharomyces cerevisiae. 16 samples. Type: Genome binding/occupancy profiling by genome tiling array.
Transcriptome profile in the human synovial MSC-aggregates
GEO Series GSE31980. Homo sapiens. 6 samples. Type: Expression profiling by array.
Histone H3K4 methylation regulation and three-dimensional aggregating distribution of a subset of genomic loci mediated by CHD3 homolog PKL in Arabidopsis [Hi-C]
GEO Series GSE305652. Marchantia polymorpha. 2 samples. Type: Other.
CRISPR/Cas12 genome editing of TXNIP in human pluripotent stem cells to generate hepatocyte-like cells and insulin-producing islet-like aggregates.
GEO Series GSE284753. Homo sapiens. 12 samples. Type: Expression profiling by high throughput sequencing.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.