Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

1,298

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

1,298 results for “Archive”

Learn how ShareScore rates datasets ↗
zenodo40/100

Indian energy: Archive materials

<p>This collection houses a number of reports and datasets concerning India's energy situation. In many cases these reports have been difficult to obtain, so I am releasing them here to make similar work easier for others.&nbsp;Note that many of the PDF documents here have already been scraped, and key data are available in machine-readable format here:&nbsp;<a href="https://robbieandrew.github.io/india/">https://robbieandrew.github.io/india/</a></p> <p>The collection includes:</p> <ul> <li>Monthly reports by the Ministry of Coal to Cabinet, from April 2015</li> <li>Monthly summary reports by the Ministry of Coal to Cabinet, from September 2014</li> <li>Monthly statistical reports by the Ministry of Coal to Cabinet, from April 2020</li> <li>Coal India Ltd's monthly reports, from February 2013</li> <li>SCCL's monthly reports, from April 2016</li> <li>Energy Statistics Yearbooks, from 2007</li> <li>Monthly Gas reports from PPAC, from December 2015</li> <li>Annual Provisional Coal Statistics reports, from 2005-06</li> <li>Annual Coal Directory reports, from 2007-08</li> <li>Monthly trade by principal commodities from DGCIS, from April 2004</li> <li>CEA annual reports, from 2010</li> <li>CEA daily renewable generation reports, from January 2021</li> <li>CEA monthly OPM01 reports, from January 2008</li> <li>CEA monthly OPM16&nbsp;reports, from January 2008</li> <li>CEA monthly Executive Summary reports, from January 2006</li> <li>CEA daily DGR17 reports, from January 2013 (there are gaps)</li> </ul> <p>The collection is grouped into ZIP archives, and within each ZIP archive the files are named to indicate the time period to which they apply. This might be the fiscal year, such as "2012-13", or the month in YYYYMM format, such as "200110".</p> <p>Related datasets:</p> <ul> <li>"Background data for: Timely estimates of India's annual and monthly fossil CO2 emissions" <a href="https://doi.org/10.5281/zenodo.3894394">https://doi.org/10.5281/zenodo.3894394</a></li> <li>"Monthly, state-level electricity generation (from 2007) and capacity (from 2008) in India: Thermal, Natural Gas, Nuclear, and Large Hydro" <a href="https://zenodo.org/record/4542834">https://zenodo.org/record/4542834</a></li> </ul> <p>&nbsp;</p>

opencc-by-4.0Mar 2023View details →
zenodo40/100

Modeling archive of How does humidity data impact land surface modeling of hydrothermal regimes at a permafrost site in Utqiaġvik, Alaska?

<p>Modeling archive contains&nbsp;the meteorological forcings, model input files, and Jupyter notebooks used to generate model meshes and figures for the paper entitled &quot;How does humidity data impact land surface modeling of hydrothermal regimes at a permafrost site in Utqiaġvik, Alaska?&quot;</p>

opencc-by-4.0Sep 2023View details →
dryad40/100

Data archive for: Fire-regime variability and ecosystem resilience over four millennia in a Rocky Mountain subalpine watershed

<p>Wildfires strongly influence forest ecosystem processes, including carbon and nutrient cycling and vegetation dynamics. As fire activity increases under changing climate conditions, the ecological and biogeochemical resilience of many forest ecosystems remains unknown. To investigate the resilience of forest ecosystems to changing climate and wildfire activity over decades to millennia, we developed a 4800-yr high-resolution lake-sediment record from Silver Lake, Montana, USA (47.360° N, 115.566° W). Charcoal particles, pollen grains, element concentrations, and stable isotopes of C and N serve as proxies of past changes in fire, vegetation, and ecosystem processes such as nitrogen cycling and soil erosion, within a small subalpine forest watershed. A published lake-level history from Silver Lake provides a local record of paleohydrology. A trend toward increased effective moisture over the late Holocene coincided with a distinct shift in the pollen assemblage c. 1900 yr BP, resulting from increased subalpine conifer abundance. Fire activity, inferred from peaks in macroscopic charcoal, decreased significantly after 1900 yr BP, from one fire event every 126 yr (83–184 yr, 95% CI) from 4800–1900 yr BP, to one event every 223 yr (175–280 yr) from 1900 yr BP to present. Across the record, individual fire events were followed by two distinct decadal-scale biogeochemical responses, reflecting differences in ecosystem impacts of fires on watershed processes. These distinct biogeochemical responses were interpreted as reflecting fire severity, highlighting (i) erosion, likely from large or high-severity fires, and (ii) nutrient transfers and enhanced within-lake productivity, likely from lower-severity or patchier fires. Biogeochemical and vegetation proxies returned to pre-fire values within decades regardless of the nature of fire effects. Paleo records of fire and ecosystem responses provide a novel view revealing past variability in fire effects, analogous to spatial variability in fire severity observed within contemporary wildfires. Overall, the paleo record highlights ecosystem resilience to fire across long-term variability in climate and fire activity. Higher fire frequencies in past millennia relative to the 20th and 21st centuries suggest that northern Rocky Mountain subalpine ecosystems could remain resilient to future increases in fire activity, provided continued ecosystem recovery within decades.</p>

opencc-zeroSep 2023View details →
zenodo40/100

Monarch Initiative Data Archive

<p>The Monarch Initiative is an extensive knowledge graph and ecosystem of tools made for the benefit of clinicians, researchers, and scientists. The knowledge graph consists of millions of entities &ndash; genes, diseases, phenotypes, and many more &ndash; imported from dozens of sources.</p>

openbsd-licenseSep 2023View details →
zenodo40/100

benmarwick/binford: archive on Zenodo

<p>Datasets used in Binford&#39;s 2001 book &quot;Constructing Frames of Reference: An Analytical Method for Archaeological Theory Building Using Ethnographic and Environmental Data Sets&quot;</p>

openother-openOct 2023View details →
zenodo40/100

"It was recorded on Sunday, morning of the 28th of September as some of the slower runners of the Berlin Marathon made it past Torstrasse near my flat. Iwas out to buy some bread for breakfast, but Iusually bring a camera and my Edirol R-1 recorder whenever Igo out. Since Iwas freshly returned to Berlin Iguess Iwas sensitive to the more antiquated sounds which still survive there, like that of the organ grinder. Iam generally interested in how human beings are replacing the presence of Nature with an artificial environment made entirely by human hands (and thus far more understandable, it is hoped). In this new Human Nature, the sounds of Nature are also Human made. Iwrite about these things, but Ialso use the sounds in my videos and my interactive and generative media work, so generally Iam wandering around building up my archive of media documents for use as material in future works." [Baruch/ gottlieb]17 in Collecting Sounds. Online Sharing of Field Recordings as Cultural Practice

"It was recorded on Sunday, morning of the 28th of September as some of the slower runners of the Berlin Marathon made it past Torstrasse near my flat. Iwas out to buy some bread for breakfast, but Iusually bring a camera and my Edirol R-1 recorder whenever Igo out. Since Iwas freshly returned to Berlin Iguess Iwas sensitive to the more antiquated sounds which still survive there, like that of the organ grinder. Iam generally interested in how human beings are replacing the presence of Nature with an artificial environment made entirely by human hands (and thus far more understandable, it is hoped). In this new Human Nature, the sounds of Nature are also Human made. Iwrite about these things, but Ialso use the sounds in my videos and my interactive and generative media work, so generally Iam wandering around building up my archive of media documents for use as material in future works." [Baruch/ gottlieb]17

opencc-by-4.0Dec 2019View details →
dryad40/100

COVID information commons archive

Open the record for dataset details and reuse information.

publicSep 2025View details →
dryad40/100

Data from: Integrating tracking and resight data enables unbiased inferences about migratory connectivity and winter range survival from archival tags

Open the record for dataset details and reuse information.

publicMar 2022View details →
dryad40/100

Russian Arctic Vegetation Archive – a new database of plant community composition and environmental conditions

Open the record for dataset details and reuse information.

publicJan 2024View details →
dryad40/100

Alaska compendium of ocean profile data [ACOD]: Archival CTD and nutrient hydrography from NOAA's EcoFOCI, EMA, and predecessor programs

Open the record for dataset details and reuse information.

publicOct 2025View details →
dryad40/100

A novel method to assess the integrity of frozen archival DNA samples: Alpha-diversity ratios of short and long-read 16S rRNA gene sequences

Open the record for dataset details and reuse information.

publicAug 2024View details →
dryad40/100

Data archive for: Fire-regime variability and ecosystem resilience over four millennia in a Rocky Mountain subalpine watershed

Open the record for dataset details and reuse information.

publicSep 2023View details →
edi40/100

North Temperate Lakes LTER: Crayfish Abundance 1981 - current (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/287/3, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-ntl/3/29. The abstract below was extracted from the Level 0 data package and is included for context: Crayfish data include crayfish catch in cylindrical minnow traps baited with beef liver and occasional occurrence in other gear used to sample fish. Traps are placed at fyke net locations in nine study lakes (Allequash, Big Muskellunge, Crystal, Sparkling, Trout, Mendota, Monona, Wingra and Fish). Crayfish traps have been eliminated as gear in the Madison area lakes (Mendota, Monona, Wingra, and Fish) after 2003. Individuals are identified to species and counted. In Trout and Sparkling Lake more detailed surveys have been conducted during the summer on an ad hoc basis to track distribution and abundance of the invading species Orconectes rusticus. Additional data sets consist of pre-LTER sets (initiated in late June 1972) gathered by Capelli (Ph.D. dissertation) and Lorman (Ph.D. dissertation). Most of pre-LTER data is detailed distribution in Trout Lake, and community composition in other area lakes. Sampling Frequency: annually Number of sites: 9 Note that 2020 data does not exist due to insufficient sampling.

openCC (other)Jul 2021View details →
edi40/100

North Temperate Lakes LTER: Benthic Macroinvertebrates 1981 - current (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/290/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-ntl/11/35. The abstract below was extracted from the Level 0 data package and is included for context: Macroinvertebrates are collected from selected shoreline and deep water locations in the seven primary lakes (Allequash, Big Muskellunge, Crystal, Sparkling, and Trout lakes, and unnamed lakes 27-02 [Crystal Bog], and 12-15 [Trout Bog]) in the Trout Lake area using modified Hester-Dendy samplers. Samplers are placed at fyke net and gill net locations in August and retrieved 3-4 weeks later. Macroinvertebrates are preserved in ethanol. This dataset contains counts of various groups of macroinvertebrates identified from specific samples. The majority of the identifications are at the genus level. The data table "Benthic Macroinvertebrate Codes" identifies the taxonomic group represented by each group code. Taxonomic references: Ecology and Classification of North American Freshwater Invertebrates, Edited by James H Thorp and Alan P Covich, Academic Press, Inc, 1991; Aquatic Insects of Wisconsin, William L Hilsenhoff, Natural History Museums Council, University of Wisconsin-Madison (1995). Sampling Frequency: annually Number of sites: 7

openCC (other)Jul 2021View details →
edi40/100

North Temperate Lakes LTER: Fish Abundance 1981 - current (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/303/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-ntl/7/39. The abstract below was extracted from the Level 0 data package and is included for context: This data set is a derived data set based on fish catch data. Data are collected annually to enable us to track the fish assemblages of eleven primary lakes (Allequash, Big Muskellunge, Crystal, Sparkling, Trout, bog lakes 27-02 [Crystal Bog] and 12-15 [Trout Bog], Mendota, Monona, Wingra and Fish). Sampling on Lakes Monona, Wingra, and Fish started in 1995; sampling on other lakes started in 1981. Sampling is done at six littoral zone sites per lake with seine, minnow or crayfish traps, and fyke nets; a boat-mounted electrofishing system samples three littoral transects. Vertically hung gill nets are used to obtain two pelagic samples per lake from the deepest point. A trammel net samples across the thermocline at two sites per lake. In the bog lakes only fyke nets and minnow traps are deployed. Parameters measured include species-level identification and lengths for all fish caught, and weight and scale samples from a subset. Derived data sets include species richness, catch per unit effort, and size distribution by species, lake, and year. Dominant species vary from lake to lake. Perch, rockbass, and bluegill are common, with walleye, large and smallmouth bass, northern pike and muskellunge as major piscivores. Cisco have been present in the pelagic waters of four lakes, and the exotic species, rainbow smelt, is present in two. The bog lakes contain mudminnows. Protocol used to generate data: Day seines were only used in 1981 and have been eliminated from this da

openCC0Jul 2021View details →
edi40/100

CGR02 Sweep sampling of Grasshoppers on Konza Prairie LTER watersheds (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/341/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-knz/29/18. The abstract below was extracted from the Level 0 data package and is included for context: Sweep samples were taken for grasshoppers (Acrididae) at two sites for each of 14 Konza Prairie LTER watersheds. Samples are taken in late July to early August. At each site on each occasion, 10 sets of 20 sweeps (200 sweeps total) are taken. Stored data include for each site on each occasion: total number of each species (all instars combined) collected and total number for each instar for each species (200 sweeps combined).

openCC0Aug 2021View details →
edi40/100

CSM01 Seasonal Summary of Numbers of Small Mammals on 14 LTER Traplines in Prairie Habitats at Konza Prairie (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/343/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-knz/88/8. The abstract below was extracted from the Level 0 data package and is included for context: Data set contains seasonal summaries (spring and autumn) of the number of individuals of each species of small mammal captured (relative abundance) on each grassland trapline. Each record contains year, season, trapline and number of individuals captured of each species. These live trap records are based on daily captures during two 4-day trapping periods in spring (late February to early April) and autumn (early October to mid-November) for each of 14 permanent traplines established on seven fire-grazing treatments (two traplines per treatment). These seven fire-grazing treatments include three sites that are grazed by bison (1 unburned, 1 annual burn and 1 4-year burn) and four sites that are not grazed by bison (1 unburned, 1 annual burn and 2 4-year burn).

openCC0Aug 2021View details →
edi40/100

Biodiversity - Fauna - Bird Survey (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/191/4, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-bes/543/170. The abstract below was extracted from the Level 0 data package and is included for context: This dataset is associated with BES Bird Monitoring Bird Monitoring Project: ================= The BES Bird Monitoring Project is a breeding bird survey designed to find out what birds are found in the breeding season in Baltimore and where. Our monitoring efforts will show associations among block group socioeconomic variables, land cover, land use, and habitat features with breeding bird abundance, to provide information for land managers on possible consequences of land use changes on bird communities. A distinguishing feature of the bird monitoring at BES LTER, relative to other urban bird work, is the capacity for long-term monitoring of features at multiple scales through links to other parts of the project. Different processes influence habitat for birds at different scales, e.g. ongoing household level human decision-making at lot scale vs. block or neighborhood scale abandonment/re-development. Our project seeks to understand how these processes impact bird occurrence, abundance, and composition differ at the lot, block and neighborhood scale. The database consists of four tables. Sites, Surveys, Taxalist, and Birds. Sites records thje sites and their characteristics. Surveys describe the actual outings or sampling sessions. They describe the weather, the temperature, the sites visited. Taxalist provides the integration of speciaies abbreviations and common names, and Birds describes the actual sightings, linking to the other three tables. Attribute info

openCustomAug 2021View details →
edi40/100

Meiobenthos abundance. Long-term variability and dynamics of estuarine meiobenthic populations for North Inlet Estuary, South Carolina, from 1972 to 1992, North Inlet LTER (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/350/3, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-nin/6/1. The abstract below was extracted from the Level 0 data package and is included for context: The original purpose of this research was to determine if natural meiobenthic assemblages exhibited continuity over time and to monitor several physical variables to determine if these influenced long-term temporal patterns. The most recent study focused on variation and the relations of meiobenthos abundance with environmental factors over 11 years. Typically marine benthic community studies are limited temporally and the majority of previously published 'longterm' meiofauna results (all taxa) were based on about a year's duration.

openOpenAug 2021View details →
edi40/100

Zooplankton Data for North Inlet Estuary, South Carolina, from 1981 to 1992, North Inlet LTER (Reformatted to a Darwin Core Archive)

This data package is formatted as a Darwin Core Archive (DwC-A, event core). For more information on Darwin Core see https://www.tdwg.org/standards/dwc/. This Level 2 data package was derived from the Level 1 data package found here: https://pasta.lternet.edu/package/metadata/eml/edi/352/2, which was derived from the Level 0 data package found here: https://pasta.lternet.edu/package/metadata/eml/knb-lter-nin/2/1. The abstract below was extracted from the Level 0 data package and is included for context: This data package consists of Zooplankton Data for North Inlet Estuary, South Carolina, from 1981 to 1992, North Inlet LTER. The purpose of the long term monitoring of zooplankton was to characterize the fauna in the water column larger than or equal to 153 microns and to obtain some basic information on each of the taxa encountered there. A sampling regime of collections made at regular biweekly intervals was implemented to provide the best quantitative assessment of long term changes in the zooplankton population dynamics.

openOpenAug 2021View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record