Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
3,481
datasets available to search
ShareScore release 0.9.0
Dataset results
3,481 results for “data set”
Data from: Enriching the ant tree of life: enhanced UCE bait set for genome-scale phylogenetics of ants and other Hymenoptera
Open the record for dataset details and reuse information.
Data from: Protein Set Transformer: A protein-based genome language model to power high diversity viromics
Open the record for dataset details and reuse information.
Data from: Let's stick together: infection enhances preferences for social settings in a songbird species
Open the record for dataset details and reuse information.
Scripts and data sets associated with: On testing homogeneity of the evolutionary process using alignments of homologous sequences
Open the record for dataset details and reuse information.
Online appendix and simulated data sets for assesment of Birth-Death Exposed-Infectious (BDEI) phylodynamic model estimators
Open the record for dataset details and reuse information.
Data for: "Generative prediction of causal gene sets responsible for complex traits"
Open the record for dataset details and reuse information.
Data from: Unravelling the distinctive craniomandibular morphology of the Plio‐Pleistocene Eumysops in the evolutionary setting of South American octodontoid rodents (Hystricomorpha)
Open the record for dataset details and reuse information.
Tassie BRUV: A benchmark data set for computer vision and movement quantification algorithms
Open the record for dataset details and reuse information.
Biogeochemistry data set for soil waters, streams, and lakes near Toolik on the North Slope of Alaska, 2011.
Data file describing the biogeochemistry of samples collected at various sites near Toolik Lake, North Slope of Alaska. Sample site descriptors include a unique assigned number (sortchem), site, date, time, depth, distance (downstream), elevation, treatment, date-time, category, and water type (lake, surface, soil). Physical measures collected in the field include temperature (water, soil, well water), conductivity, pH, average thaw depth, well height, discharge, stage height, and light (lakes). Chemical analysis for the sample include alkalinity; dissolved inorganic and organic carbon (DIC and DOC); inorganic and total dissolved nutrients (NH4, PO4, NO3, TDN, TDP); particulate carbon, nitrogen, and phosphorus (PC, PN, and PP); cations (Ca, Mg, Na, K, and Si); anions (SO4 and Cl); and oxygen.
Physical site characteristics for the ARCSS/TK stream dissolved organic carbon biodegradability (2011) data set.
The (ARCSSTK) did extensive research during 2009-2011 field seasons in Arctic Alaska. The objective of this data set was to measure the quantity and biodegradability of DOC from headwater streams and rivers across three geographic regions and across four natural ‘treatments’ (reference; thermokarst-; burned-, and thermokarst + burned-impacted streams) to evaluate which factors most strongly influence DOC quantity and biodegradablity at a watershed scale. This table provides physical site characteristics for the locations sampled for stream water biodegradability.
Bacterial production and respiration data set for NSF Arctic Photochemistry project on the North Slope of Alaska.
Data file describing the bacterial production and bacterial respiration of water samples collected at various sites near Toolik Lake on the North Slope of Alaska. Sample site descriptors include site, date, time, depth, and category representing severity of thermokarst disturbance. A synthesis of the data presented here is published in Cory et al. 2013, PNAS 110:3429-3434, and in Cory et al. 2014, Science 345:925-928.
Biogeochemistry data set for NSF Arctic Photochemistry project on the North Slope of Alaska.
Data file describing the biogeochemistry of samples collected at various sites near Toolik Lake on the North Slope of Alaska. Sample site descriptors include a unique assigned number (sortchem), site, date, time, depth, and category (level of thermokarst disturbance). Physical measures collected in the field include temperature, electrical conductivity, and pH. Chemical analyses include alkalinity; dissolved organic carbon (DOC); inorganic and total dissolved nutrients (NH4, PO4, NO3, TDN, TDP); particulate carbon, nitrogen, and phosphorus (PC, PN, and PP); cations (Ca, Mg, Na, K); anions (Cl, SO4); and silica. A synthesis of much of the data presented here is published in Cory et al. 2013, PNAS 110:3429-3434; Cory et al. 2014, Science 345:925-928; and Page et al. 2013, Environment, Science, & Technology 47:12860−12867.
Light profile data set for NSF Photochemistry project on the North Slope of Alaska.
Data file containing the irradiance profile with depth in two rivers on the North Slope of Alaska near Toolik Lake . Variables include site, depth, and wavelength. A synthesis of the data presented here is published in Cory et al. 2013, PNAS 110:3429-3434, and in Cory et al. 2014, Science 345:925-928.
Apparent quantum yield data set for NSF Photochemistry project on the North Slope of Alaska.
Data file describing the apparent quantum yield of photo-oxidation, photo-mineralization, and photo-stimulated microbial respiration of dissolved organic carbon in water samples collected at various sites near Toolik Lake on the North Slope of Alaska. A synthesis of the data presented here is published in Cory et al. 2013, PNAS 110:3429-3434, and in Cory et al. 2014, Science 345:925-928.
Photochemistry data set for NSF Photochemistry project on the North Slope of Alaska.
Data file containing optical characterization of colored dissolved organic matter (CDOM). Data include CDOM absorption coefficients, water column light attenuation coefficients, specific UV light absorbance (SUVA254), spectral slope ratio, and fluorescence index from waters near Toolik Lake on the North Slope of Alaska. A synthesis of the data presented here is published in Cory et al. 2013, PNAS 110:3429-3434, and in Cory et al. 2014, Science 345:925-928.
Scripts and Data sets used in Study
<p>Evaluation of Machine Learning Algorithms for Breast Cancer Malignancy Prediction Based on Mammogram Data. the provided files represent scripts and datasets used in evaluation </p>
Data set for initializing and forcing of high-resolution local area model implementations in two ATLAS case study areas, Rockall Bank and Condor Seamount
<p>The data set includes all necessary data for set up, initial and boundary conditions of high-resolution local area model implementations using the ROMS-AGRIF model in two case study areas, Rockall Bank and Condor Seamount, conducted as part of the EU ATLAS project. The data include all computational grids, initialization fields (temperature, salinity) and boundary conditions (temperature salinity, currents, sea surface height) for each case study area.</p>
Data set related to the manuscript "Ionic liquids under confinement: From systematic variations of the ion and pore sizes towards an understanding of structure and dynamics in complex porous carbons"
<p>Graphical files in the agr format for all the figures in the main text of the manuscript entitled "Ionic liquids under confinement: From systematic variations of the ion and pore sizes towards an understanding of structure and dynamics in complex porous carbons" (10.1021/acsami.9b16740). Example of input file for one of the systems simulated.</p>
Table S2.- ADC phylogeny data set.
<p>Data set for phylogenetic analysis, we used the Pfam v31.0 database to determine which proteinortho clusters represent the ADC proteins. A total of 37 homologs belonging to alphaproteobacteria sequences were tested with a group of nine external sequences were aligned against Muscle v3.8.31. The resulting dataset containing 46 putative ADC homologs was used to infer the evolutionary relationships. We used ProtTest3 v3.4.2 for the evolutionary model and, the best result was LG+G model; using amino acid alignment. The phylogenetic analysis was performed with PhyML v3.3.20170530 under (-d aa -m LG -a e -o ltr) parameters. </p>
Data set for the article "Formation of Néel-type skyrmions in an antidot lattice with perpendicular magnetic anisotropy" DOI: 10.1103/PhysRevB.100.144435
<p>This dataset provides experimental data and linear fit reported in the Fig2b of the article </p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.