Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
1,298
datasets available to search
ShareScore release 0.9.0
Dataset results
1,298 results for “Archive”
Beetle Team Tribolium Data Archive
Open the record for dataset details and reuse information.
Data from: Evaluating genotyping-in-thousands by sequencing as a genetic monitoring tool for a climate sentinel mammal using non-invasive and archival samples
Open the record for dataset details and reuse information.
Sequencing data for: Tracking climate-change induced biological invasions over 4 decades by metabarcoding archived natural eDNA samplers
Open the record for dataset details and reuse information.
A data archive including processed spiking data, raw EMG datasets, and video data during locomotion behavior from six mice
Open the record for dataset details and reuse information.
Archived data for: Balancing selection, genetic drift, and human mediated-introgression interplay to shape MHC (functional) diversity in Mediterranean brown trout
Open the record for dataset details and reuse information.
ArXiV Archive
Open the record for dataset details and reuse information.
Nitrate concentration and 15N signal of Streamwater and Precipitation from Archived Samples: Watershed 3
This data set includes the analysis of 18 Oxygen isotopes in precipitation and streamwater archived samples from watershed 3. These data were gathered as part of the Hubbard Brook Ecosystem Study (HBES). The HBES is a collaborative effort at the Hubbard Brook Experimental Forest, which is operated and maintained by the USDA Forest Service, Northern Research Station.
Preliminary metadata for the UHMRL digital tape archive pilot
<p>The dataset (.csv) consists of preliminary metadata for pilot testing digitization and research-driven metadata production of the University of Helsinki Music Research Laboratory and Electronic Music Studio tape archive. PI and contact information: Mikko Ojanen / https://orcid.org/0000-0002-7833-9659</p>
Archive of JPG image files of Herbarium specimens from Columbia used in the BRAVO project.
<p>Archive of JPG image files of Herbarium specimens used in the BRAVO project.</p> <p>The individual images of herbarium specimens will also be included in the Zenodo project (in time)</p> <p>The project will pull these files into the workbench <a href="http://bravo.rbge.info/">http://bravo.rbge.info/</a></p> <p>The metadata for these images is included in the Zenodo project.</p>
Data Appendices for "Textual Scholarship and Contemporary Literary Studies: Jennifer Egan's Editorial Processes and the Archival Edition of Emerald City"
<p>This dataset documents the version variants between Jennifer Egan's early short stories, "The Stylist" and "Sacred Heart". It is a supplement to the article "Textual Scholarship and Contemporary Literary Studies: Jennifer Egan’s Editorial Processes and the Archival Edition of <em>Emerald City</em>" in <em>LIT: Literature Interpretation Theory</em>.</p>
Hall-of-Apps: The Top Android Apps Metadata Archive
<p>The amount of Android apps available for download is constantly increasing, exerting a continuous pressure on developers to publish outstanding apps. Google Play (GP) is the default distribution channel for Android apps, which provides mobile app users with metrics to identify and report apps quality such as rating, amount of downloads, previous users comments, etc. In addition to those metrics, GP presents a set of top charts that highlight the outstanding apps in different categories. Both metrics and top app charts help developers to identify whether their development decisions are well valued by the community. Therefore, app presence in these top charts is a valuable information when understanding the features of top-apps. In this paper we present <strong>Hall-of-Apps</strong>, a dataset containing top charts' apps metadata extracted (weekly) from GP, for 4 different countries, during 30 weeks. The data is presented as (i) raw HTML files, (ii) a MongoDB database with all the information contained in app's HTML files (e.g., app description, category, general rating, etc.), and (iii) data visualizations built with the D3.js framework. A first characterization of the data along with the urls to retrieve it can be found in our online appendix: <a href="https://thesoftwaredesignlab.github.io/hall-of-apps-tools/">https://thesoftwaredesignlab.github.io/hall-of-apps-tools/</a></p>
Project Panormos Archaeological Survey: Photograph Archive (survey-photo-archive)
<p>This forms part of the preliminary open data release for the Project Panormos archaeological survey.</p> <p>The panormos/survey-photo-archive repository contains the source archive of digital images (primarily photographs) created during the course of the fieldwork for Project Panormos excavations and survey. This includes field photographs taken along with tract and POI data, unprocessed photographs of finds, and occasionally edited, colour- or balance-corrected photographs in an appropriate quality for (re-)publication. Some included photographs are of the general region and although produced by team members are not a direct result of the survey work.</p> <p>There are some restrictions on the distribution of images from the original source archive; these images will not be available in the distribution of "Core" released photographs found here, but may be made available on request. Basic metadata from these images may be available in the associated metadata files nonetheless.</p> <p><strong>Release 0.2.0</strong> includes data from the 2015, 2017 and 2019 seasons. It is a pre-publication or observation version. No derivative works should be made until the expiry of the observation phase: please see enclosed LICENSE file for details.</p> <p><strong>Release 0.1.0</strong> includes data from the 2015 season. It is a pre-publication or observation version. No derivative works should be made until the expiry of the observation phase: please see enclosed LICENSE file for details.</p> <p><strong>Releases below 1.0.0 represent preprint working datasets before final publication. </strong>Although every effort has been made to reduce errors and make the datasets available in a form that should be easy to navigate, the status of the data as a form of "beta" should be borne in mind.</p>
MAIRY Photo Archive
<p>ENG: Since 1980, the Italian Archaeological Mission in Yemen (MAIRY) has collected a remarkable documentation (drawings, photographs, etc.) relating to its scientific activity in the country (surveys, discoveries, excavations, restoration, etc). Thanks to a valuable three-year project (2017-2020) for the digitisation of the documentary archive of the MAIRY, generously financed by <a href="https://whitelevy.fas.harvard.edu/">The Shelby White and Leon Levy Program for Archaeological Publications</a>, the images have been entered in a database. This can be accessed by the name of archaeological site and year of photograph. Due to technical issues, in some cases it has been necessary to change the original ID of the sites to a standard Access Number composed of three alphabetic letters (see Access Number). In the Access Number, the letter 'n' is the initial of 'negative' (b&w and colour), 's' stands for slide, 'p' for plate, 'o' for object, and 'm' for map.<br> Most of the material in this archive has been published in books and academic articles. The remaining images are part of current studies and not published yet. Therefore, these images, which bear the MAIRY watermark, cannot be copied. All images from MAIRY’s archive are owned by MAIRY itself and Monumenta Orientalia, and therefore protected by copyright. To request the use of the images, you should write to Monumenta Orientalia e-mail address [monumenta(dot)orientalia(at)yahoo(dot)it] giving the Access number, and a brief reason for the request. All images from MAIRY’s archive are owned by MAIRY itself and Monumenta Orientalia, and therefore protected by copyright. To request the use of the images, you should write to Monumenta Orientalia e-mail address [monumenta(dot)orientalia(at)yahoo(dot)it] giving the Access number, and a brief reason for the request.</p> <p>ITA: La Missione Archeologica Italiana in Yemen (MAIRY) ha raccolto dal 1980 una notevole documentazione grafica e fotografica relativa alla propria attività scientifica nel Paese (ricognizioni, scoperte, scavi, restauri, ecc.). Grazie ad un prezioso progetto triennale (2017-2020), generosamente finanziato da <a href="https://whitelevy.fas.harvard.edu/">The Shelby White and Leon Levy Program for Archaeological Publications</a>, l’archivio documentario della MAIRY è stato digitalizzato e inserito in un database. Qui di seguito è consultabile per nome del sito e anno di scatto della fotografia. Per esigenze tecniche in fase di inventariazione, molte sigle originarie dei siti sono state opportunamente ‘adeguate’, uniformandole tutte a tre lettere (v. Access number). Nei numeri di inventario (Access number), la lettera ‘n’ sta per ‘negative’ (negativo in b&n e colori), ‘s’ slide (diapositiva), ‘p’ plate (tavola), ‘o’ object (oggetto), ‘m’ map (carta, mappa). Gran parte del materiale di questo archivio è pubblicato; il resto è in corso di studio o in corso di stampa. Per questo si è ritenuto opportuno apporre la filigrana del logo MAIRY. Tutte le immagini dell’archivio della MAIRY sono di proprietà della Missione stessa e di Monumenta Orientalia, e quindi soggette a copyright. Per richiedere l’utilizzo delle foto occorre scrivere alla casella di posta elettronica di Monumenta Orientalia [monumenta(dot)orientalia(at)yahoo(dot)it] indicando il numero di inventario dell’immagine (Access number), motivando la richiesta con un breve testo. Tutte le immagini dell’archivio della MAIRY sono di proprietà della Missione stessa e di Monumenta Orientalia, e quindi soggette a copyright. Per richiedere l’utilizzo delle foto occorre scrivere alla casella di posta elettronica di Monumenta Orientalia [monumenta(dot)orientalia(at)yahoo(dot)it] indicando il numero di inventario dell’immagine (Access number), motivando la richiesta con un breve testo.</p>
Archived GBIF Dataset: Endemic plants of the Congo basin
<p>This dataset was archived from GBIF on 2020-09-10. More up to date data on the specimens inside can be found in the <a href="https://doi.org/10.15468/wrthhx">Meise Botanic Garden Herbarium dataset</a>.</p> <p><strong>Description</strong></p> <p>Specimens of plants endemic to either Burundi, Democratic Republic of the Congo<br> or Rwanda</p> <p><strong>Temporal scope</strong></p> <ul> <li>January 1, 1890 - December 31, 2014</li> </ul> <p><strong>Geographic scope</strong></p> <p>Rwanda, Democratic Republic of the Congo & Burundi</p>
Workflow Trace Archive Galaxy trace
Traces from different biomedical research workflows, executed on the public Galaxy server in Europe.
Archive data for: Loss of predation risk from apex predators can exacerbate marine tropicalization caused by extreme climatic events
<p>1. Extreme climatic events (ECEs) and predator removal represent some of the most widespread stressors to ecosystems. Though species interactions can alter ecological effects of climate change (and vice versa), it is less understood whether, when, and how predator removal can interact with ECEs to exacerbate their effects. Understanding the circumstances under which such interactions might occur is critical because predator loss is widespread and ECEs can generate rapid phase shifts in ecosystems which can ultimately lead to tropicalization.</p> <p>2. Our goal was to determine whether loss of predation risk may be an important mechanism governing ecosystem responses to extreme events, and whether the effects of such events, such as tropicalization, can occur even when species range shifts do not. Specifically, our goal was to experimentally simulate loss of an apex predator, the tiger shark (<i>Galeocerdo cuvier</i>) effects on a recently damaged seagrass ecosystem of Shark Bay, Australia by applying documented changes to risk sensitive grazing of dugong (<i>Dugong dugon</i>) herbivores. </p> <p>3. Using a 16-month field experiment established in recently disturbed seagrass meadows, we used previous estimates of risk-sensitive dugong foraging behavior to simulate altered risk-sensitive foraging densities and strategies of dugongs consistent with apex predator loss, and tracked seagrass responses to the simulated grazing.</p> <p>4. Grazing treatments targeted and removed tropical seagrasses, which declined. However, like in other mixed-bed habitats where dugongs forage, treatments also incidentally accelerated temperate seagrass losses, revealing that herbivore behavioral changes in response to predator loss can exacerbate ECE effects and promote tropicalization, even without range expansions or introductions of novel species. </p> <p>5. Our results suggest that changes to herbivore behaviors triggered by loss of predation risk can undermine ecological resilience to ECEs, particularly where long lived herbivores are abundant. By implication, ongoing losses of apex predators may combine with increasingly frequent ECEs to amplify climate change impacts across diverse ecosystems and large spatial scales.</p>
Data from: An archive of longitudinal recordings of the vocalizations of adult Gombe chimpanzees
Studies of chimpanzee vocal communication provide valuable insights into the evolution of communication in complex societies, and also comparative data for understanding the evolution of human language. One particularly valuable dataset of recordings from free-living chimpanzees was collected by Frans X. Plooij and the late Hetty van de Rijt-Plooij at Gombe National Park, Tanzania (1971–73). These audio specimens, which have not yet been analysed, total over 10 h on 28 tapes, including 7 tapes focusing on adult individuals with a total of 605 recordings. In 2014 the first part of that collection of audio specimens covering the vocalizations of the immature Gombe chimpanzees was made available. The data package described here covers the vocalizations of the adult chimpanzees. We expect these recordings will prove useful for studies on topics including referential signalling and the emergence of dialects. The digitized sound recordings were stored in the Macaulay Library and the Dryad Repository. In addition, the original notes on the contexts of the calls were translated and transcribed from Dutch into English.
Data from: Resurrecting an extinct salmon evolutionarily significant unit: archived scales, historical DNA, and implications for restoration
Archival scales from 603 sockeye salmon (Oncorhynchus nerka), sampled from May to July 1924 in the lower Columbia River, were analyzed for genetic variability at 12 microsatellite loci, and compared to 17 present-day O. nerka populations—exhibiting either anadromous (sockeye salmon) or non-anadromous (kokanee) life histories—from throughout the Columbia River Basin, including areas upstream of impassable dams built subsequent to 1924. Statistical analyses identified four major genetic assemblages of sockeye salmon in the 1924 samples. Two of these putative historical groupings were found to be genetically similar to extant evolutionarily significant units (ESUs) in the Okanogan and Wenatchee rivers (pairwise FST = 0.004 and 0.002, respectively) and assignment tests were able to allocate 77% of the fish in these two historical groupings to the contemporary Okanogan River and Lake Wenatchee ESUs. A third historical genetic grouping was most closely aligned with contemporary sockeye salmon in Redfish Lake, Idaho, although the association was less robust (pairwise FST = 0.060). However, a fourth genetic grouping did not appear to be related to any contemporary sockeye salmon or kokanee population, assigned poorly to the O. nerka baseline, and had distinctive early return migration-timing suggesting that this group represented a putative historical ESU originating in headwater lakes in British Columbia that was likely extirpated sometime after 1924. The lack of a contemporary O. nerka population possessing the genetic legacy of this extinct ESU indicates that efforts to reestablish early-migrating sockeye salmon to the headwater lakes region of the Columbia River will be difficult.
Data from: Extracting DNA from 'jaws': high yield and quality from archived tiger shark (Galeocerdo cuvier) skeletal material
Archived specimens are highly valuable sources of DNA for retrospective genetic/genomic analysis. However, often limited effort has been made to evaluate and optimize extraction methods, which may be crucial for downstream applications. Here, we assessed and optimized the usefulness of abundant archived skeletal material from sharks as a source of DNA for temporal genomic studies. Six different methods for DNA extraction, encompassing two different commercial kits and three different protocols, were applied to material, so-called bio-swarf, from contemporary and archived jaws and vertebrae of tiger sharks (Galeocerdo cuvier). Protocols were compared for DNA yield and quality using a qPCR approach. For jaw swarf, all methods provided relatively high DNA yield and quality, while large differences in yield between protocols were observed for vertebrae. Similar results were obtained from samples of white shark (Carcharodon carcharias). Application of the optimized methods to 38 museum and private angler trophy specimens dating back to 1912 yielded sufficient DNA for downstream genomic analysis for 68% of the samples. No clear relationships between age of samples, DNA quality and quantity were observed, likely reflecting different preparation and storage methods for the trophies. Trial sequencing of DNA capture genomic libraries using 20 000 baits revealed that a significant proportion of captured sequences were derived from tiger sharks. This study demonstrates that archived shark jaws and vertebrae are potential high-yield sources of DNA for genomic-scale analysis. It also highlights that even for similar tissue types, a careful evaluation of extraction protocols can vastly improve DNA yield.
Data from: Return of a giant: DNA from archival museum samples helps to identify a unique cutthroat trout lineage formerly thought to be extinct
Currently one small, native population of the culturally and ecologically important Lahontan cutthroat trout (Oncorhynchus clarkii henshawi, LCT, Federally listed) remains in the Truckee River watershed of northwestern Nevada and northeastern California. The majority of populations in this watershed were extirpated in the 1940's due to invasive species, overharvest, anthropogenic water consumption and changing precipitation regimes. In 1977, a population of cutthroat trout discovered in the Pilot Peak Mountains in the Bonneville basin of Utah, was putatively identified as the extirpated LCT lacustrine lineage native to Pyramid Lake in the Truckee River basin based upon morphological and meristic characters. Our phylogenetic and Bayesian genotype clustering analyses of museum specimens collected from the large lakes (1872-1913) and contemporary samples collected from populations throughout the extant range provide evidence in support of a genetically distinct Truckee River basin origin for this population. Analysis of museum samples alone identified three distinct genotype clusters and historical connectivity among water bodies within the Truckee River basin. Baseline data from museum collections indicate that the extant Pilot Peak strain represents a remnant of the extirpated lacustrine lineage. Given the limitations on high quality data when working with a sparse number of preserved museum samples, we acknowledge that, in the end, this may be a more complicated story. However, the paucity of remnant populations in the Truckee River watershed in combination with data on the distribution of morphological, meristic and genetic data for Lahontan cutthroat trout, suggest that recovery strategies, particularly in the large lacustrine habitats should consider this lineage as an important part of the genetic legacy of this species.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.