Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

23,351

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

23,351 results for “Comparative”

Learn how ShareScore rates datasets ↗
zenodo44/100

Aphidinae comparative genomics resource

<p>Here we provide early access to 18 new genome assemblies, including 8 assembled to chromosome-scale, for aphids from the subfamily Aphidinae.&nbsp;For consistency and to aid comparative analysis, all genomes have been annotated using the same repeat masking and RNA-seq-based gene prediction pipeline.&nbsp;Using this pipeline we also provide new annotations for three previously published genome assemblies.</p> <p>The genome assemblies and annotations are made freely available without restriction, we only request that this Zenodo resource is cited when using the data. Raw sequence data upload to NCBI is underway and full details of all accessions will be given in an updated version of this resource. Manuscripts are in preparation describing the individual genome assemblies in detail and larger comparative genome analyses and we will update this resource with additional citation information as papers are published.</p> <p>Full details of all genome assemblies and annotations included in this release are given in the attached &quot;Data_Description.pdf&quot; document.&nbsp;</p> <p><strong>Aphid species included in this release (bold type = chromosome-scale assembly):</strong></p> <p><em><strong>Aphis fabae</strong><br> Aphis glycines </em>(updated annotation)<br> <em><strong>Aphis gossypii</strong><br> Aphis thalictri<br> Aphis rumicis<br> Brachycaudus cardui<br> Brachycaudus helichrysi<br> Brachycaudus klugkisti<br> <strong>Brevicoryne brassicae</strong><br> Diuraphis noxia<br> <strong>Macrosiphum albifrons</strong><br> Metopolophium dirhodum<br> Myzus cerasi&nbsp;</em>(updated annotation)<br> <em>Myzus ligustri<br> Myzus lythri<br> Myzus varians<br> Pentalonia nigronervosa&nbsp;</em>(updated annotation)<br> <em><strong>Phorodon humuli</strong><br> <strong>Rhopalosiphum padi<br> Sitobion avenae<br> Sitobion miscanthi</strong></em></p>

opencc-by-4.0Feb 2022View details →
zenodo44/100

Data from: Radial stem growth of the clonal shrub Alnus alnobetula at treeline is constrained by summer temperature and winter desiccation and differs in carbon allocation strategy compared to co-occurring Pinus cembra

<p><strong>Data are documented in the following article:</strong></p> <p>Oberhuber W., G Wieser, F. Bernich, A. Gruber (2022) Radial stem growth of the clonal shrub <em>Alnus alnobetula</em> at treeline is constrained by summer temperature and winter desiccation and differs in carbon allocation strategy compared to co-occurring <em>Pinus cembra</em>. Forests 2022, 13, 440. doi: 10.3390/f13030440.</p> <p>&nbsp;</p> <p><strong>Summary:</strong></p> <p>Global change is affecting species areal distribution in many regions. A better understanding of how land-use change and climate warming affects shrub growth is essential for improved predictions of forest dynamics at the alpine treeline. Evaluation of radial stem growth of the clonal shrub <em>Alnus alnobetula</em> (= <em>Alnus viridis</em>) and the co-occurring tree species Swiss stone pine (<em>Pinus cembra</em>) within an alpine treeline ecotone revealed that mean ring width of nitrogen fixing <em>A. alnobetula</em> was about four times lower compared to <em>P. cembra</em>. Our findings are based on ring width data from <em>A. alnobetula</em> and <em>P. cembra</em> stems sampled at the alpine treeline ecotone on Mt. Patscherkofel (47&deg;12&rsquo;N, 11&deg;27&rsquo;E, Central European Alps, Austria, elevation range 2050 to 2190 m asl). Ring width time series include 86 radii from 51 stems of <em>A. alnobetula</em> (stems had mean age of 18&plusmn;7 yrs) and 24 radii from 16 stems of <em>P. cembra </em>(18&plusmn;4 yrs). We explain our findings by different carbon allocation strategies, i.e., preference of &ldquo;vertical&rdquo; stem growth in late successional <em>P. cembra</em> vs. favoring &ldquo;horizontal&rdquo; spread in the pioneer shrub<em> A. alnobetula.</em> By favouring clonal propagation over individual stem growth <em>A. alnobetula</em> is able to quickly spread at the alpine treeline ecotone.</p>

opencc-by-4.0Mar 2022View details →
zenodo44/100

data set to bioRxiv preprint 'Persistent cross-species SARS-CoV-2 variant infectivity predicted via comparative molecular dynamics simulation

<p>This is supporting data and software code for the following preprint in bioRxiv</p> <p><strong>Persistent cross-species SARS-CoV-2 variant infectivity predicted via comparative molecular dynamics simulation</strong></p> <p>https://www.biorxiv.org/content/10.1101/2022.04.18.488629v1</p>

opencc-by-4.0Apr 2022View details →
zenodo44/100

The HumBug Challenge: ComParE 2022

<p><strong>A large-scale multi-species dataset of acoustic recordings</strong></p> <p>Dataset compatible with two papers:</p> <ul> <li>The&nbsp;<strong>Computational Paralinguistics ChallengE (ComParE): Mosquito Event Detection Task</strong><strong>&nbsp;</strong><a href="https://github.com/EIHW/ComParE2022/tree/MOS-C">https://github.com/EIHW/ComParE2022/tree/MOS-C</a></li> <li>An update to:&nbsp;<em>HumBugDB: a large-scale acoustic mosquito dataset:</em> <ul> <li><a href="https://arxiv.org/abs/2110.07607">NeurIPS 2021 Paper</a></li> <li><a href="https://github.com/HumBug-Mosquito/HumBugDB">https://github.com/HumBug-Mosquito/HumBugDB</a>.</li> </ul> </li> </ul> <p>A large-scale multi-species dataset containing recordings of mosquitoes collected from multiple locations globally, as well as via different collection methods.&nbsp;In total, we present&nbsp;20 hours&nbsp;of labelled mosquito data with 15 hours&nbsp;of corresponding background noise, recorded at the sites of 8 experiments.&nbsp;&nbsp;Of these, 64,843 seconds contain species metadata, consisting of 36 species (or species complexes).</p> <p>This repository contains:</p> <ul> <li>Audio files to be extracted into <em>audio/data/</em><em>train</em>&nbsp;and&nbsp;<em>audio/data/</em><em>dev/{a/b} respectively</em></li> <li>Metadata in <em>csv</em> format:&nbsp;<a href="https://zenodo.org/record/6478589/files/humbugdb_zenodo_0_0_2.csv?download=1">neurips_2021_zenodo_0_0_2.csv</a></li> </ul> <p>&nbsp;</p>

opencc-by-4.0Apr 2022View details →
zenodo44/100

Dataset of - Comparative architectural analysis of Cerberiopsis genus (Apocynaceae) -

<p>This dataset was used for a comparative architectural analysis of the genus <em>Cerberiopsis</em> (Apocynaceae), endemic to New Caledonia.<br> The data were collected on several areas of the territory (GPS coordinates provided) from 11/05/2021 to 28/07/2021.<br> For each individual described an architectural drawing is associated.</p>

opencc-by-4.0Jul 2022View details →
zenodo44/100

Intraspecific variation in the sensitivity of bees to pesticides: a comparative analysis in Bombus terrestris and Osmia bicornis

<p>These files describe the archived CSV files associated with the publication "Intra-specific variation in sensitivity of Bombus terrestris and Osmia bicornis to three pesticides"</p> <p>By Alberto Linguadoca, Margret J&uuml;rison, Sara Hellstr&ouml;m, Edward A. Straw1, Peter &Scaron;ima, Reet Karise, Cecilia Costa, Giorgia Serra, Roberto Colombo, Robert J. Paxton, Marika M&auml;nd, Mark J. F. Brown<br>&nbsp;</p>

opencc-by-4.0Jun 2022View details →
zenodo44/100

Comparing recent PTA results on the nanohertz stochastic gravitational wave background - full noise and GWB parameter comparison plots

<p>A full collection of plots comparing the noise properties of individual pulsars and gravitational wave background parameters discussed in the companion paper <em>Comparing recent PTA results on the nanohertz stochastic gravitational wave background</em> (IPTA 2024).</p> <p><code>Section4_GWB_comparison.zip</code> supplements and expands section 4.1, "Comparing the published GWB measurements," of IPTA (2024). It contains parameter difference distributions for GWB model parameters.&nbsp; There are four different models included. The HD correlated powerlaw (PL) model make up the basis for Figure 2.&nbsp; Additionally, there are three comparisons not included in IPTA (2024).&nbsp; First, comparisons the common uncorrelated red noise (CURN) PL model are included.&nbsp; Finally,&nbsp; comparisons of two free spectral (FS) models (HD and CURN) are included.&nbsp; These comparisons fit the HD and CURN FS posteriors using the <code>ceffyl</code> software package, and then compare the parameters of the resulting powerlaw fits.</p> <p><code>Section5_Noise_comparison.zip</code> supplements section 5, "Comparing Pulsar Noice Properties," of IPTA (2024).&nbsp; It contains plots for 27 pulsars timed by more than one PTA collaboration, including the plots for PSR J1012+5307, which are presented in Figure 7.&nbsp; The plots include noise parameter posteriors, time domain GP realizations, TOA residuals, and TOA radio frequency.</p>

opencc-by-4.0Apr 2024View details →
zenodo44/100

A study of comparative (2019-2023) trends and current acceleration in Particulate Matter (PM2.5) concentration in India

<p><span>&nbsp;For PM<sub>2.5</sub><span>&nbsp; </span>monitoring model, the data was procured from the Central Pollution Control Board&rsquo;s functional and selected air monitoring stations. The data is available online at the&nbsp;<span> Central Pollution Control Board but in form of daily trends with numerous air quality monitoring stations in an area; monthly and Annual average level especially PM2.5 trends processed from the original data.&nbsp;&nbsp;</span></span></p>

opencc-by-4.0May 2024View details →
zenodo44/100

Data from: BioEncoder: a metric learning toolkit for comparative organismal biology

<p><strong>BioEncoder: a metric learning toolkit for comparative organismal biology</strong></p> <p><strong>Abstract </strong>- In the realm of biological image analysis, deep learning (DL) has become a core toolkit, e.g., for segmentation and classification. However, conventional DL methods are challenged by large biodiversity datasets characterized by unbalanced classes and hard-to-distinguish phenotypic differences between them. Here we present BioEncoder, a user-friendly toolkit for metric learning, which overcomes these challenges by focussing on learning relationships between individual data points rather than on the separability of classes. BioEncoder is released as a Python package, created for ease of use and flexibility across diverse datasets. It features taxon-agnostic data loaders, custom augmentation options, and simple hyperparameter adjustments through text-based configuration files. The toolkit's significance lies in its potential to unlock new research avenues in biological image analysis while democratizing access to advanced deep metric learning techniques. BioEncoder focuses on the urgent need for toolkits bridging the gap between complex DL pipelines and practical applications in biological research.</p> <p><strong>Dataset&nbsp;</strong>- This data repository includes two things: a snapshot of the BioEncoder package (BioEncoder-main.zip, version 1.0.0, downloaded from https://github.com/agporto/BioEncoder on 2024-07-19 at 17:20), and the damselfly dataset used for the case study presented in the paper (bioencoder_data.zip). The dataset archive also encompasses the configuration files and the final model checkpoints from the case study, as well as a script to reproduce the results and figures presented in the paper.</p> <p><strong>How to use - </strong>Get started by consulting the&nbsp;<a href="https://github.com/agporto/BioEncoder?tab=readme-ov-file#quickstart">GithHub repository</a> for information on how to install BioEncoder, then download the <a href="../records/10909614/files/BioEncoder-data.zip?download=1&amp;preview=1">data archive</a> and run the script. Some parts of the script can be executed using the model checkpoints, for orther parts the training rountine needs to be run.&nbsp; &nbsp; &nbsp;</p>

opencc-by-4.0Apr 2024View details →
zenodo44/100

CLDF dataset derived from Tolmie and Dawson's "Comparative Vocabulary of the Indigenous Peoples in British Columbia" from 1884

<p>Cite the source of the dataset as:</p> <blockquote> <p>Tolmie, Fraser W. and Dawson, George M. (1884). Comparative vocabularies of the Indian tribes of British Columbia, with a map illustrating distribution. Montreal: Dawson Brothers.</p> </blockquote>

opencc-by-4.0Jul 2024View details →
zenodo44/100

CLDF dataset derived from Birchall et al.'s "A Combined Comparative and Phylogenetic Analysis of the Chapacuran Language Family" from 2016

<p>Cite the source of the dataset as:</p> <blockquote> <p>Birchall J, Dunn M, &amp; Greenhill SJ. 2016. A Combined Comparative and Phylogenetic Analysis of the Chapacuran Language Family. International Journal of American Linguistics 82(3). 255–284.</p> </blockquote>

opencc-by-4.0Jul 2021View details →
zenodo44/100

CLDF dataset derived from de Carvalho's "Comparative reconstruction of Proto-Purus" from 2021

<p>Cite the source of the dataset as:</p> <blockquote> <p>de Carvalho, F. O. (2021): A comparative reconstruction of Proto-Purus (Arawakan) segmental phonology. IJAL. 87.1. 49-108</p> </blockquote>

opencc-by-4.0Jul 2021View details →
zenodo44/100

CLDF dataset derived from Z'graggen's "Madang Comparative Wordlists" from 1980

<p>Cite the source of the dataset as:</p> <blockquote> <p>Z&#x27;graggen, J A. (1980) A comparative word list of the Northern Adelbert Range Languages, Madang Province, Papua New Guinea. Canberra: Pacific Linguistics.</p> </blockquote>

opencc-by-4.0Jul 2021View details →
zenodo44/100

Comparing the integration of bone cells on an even and a nano structured surface

<p>Bone cells develop better on a nano structured surface.</p> <p>Successful cellular integration is extremely important to the long-term viability of dental implants. The sooner cells can attach and surround the implant, the faster the patient recovers, and the lower the incidence of infection and site contamination. The specific biological response of the surrounding tissues depends enormously on the surface characteristics of the particular biomaterial. Moreover mouth infections are currently regarded as the main reason why dental implants fail. Therefore, antibacterial properties are another requirement for preventing potential bacterial infection.</p> <p>The multi-beam optical module developed within LASER4SURF, will be able to obtain functionalized metallic surfaces with textures around 1&mu;m or less, enabling the required tolerances to reach the best cell adhesion and antibacterial properties. Thus, it will lead to extended implant life, reduced rejection and improved overall quality of life for patients.</p>

opencc-by-4.0Jun 2018View details →
zenodo44/100

Chapter 3: Scale Separation Reliability: What Does it Mean in the Context of Comparative Judgement?

<p>This is the supplementary material for Chapter 3 of the dissertation &quot;Beyond a Mere Rank Order: The Method, the Reliability and the Efficiency of Comparative Judgment&quot; and the article &nbsp;&quot;Scale separation reliability: What does it mean in the context of comparative judgment?&quot; published in&nbsp;&quot;Applied Psychological Measurement&quot;.</p>

opencc-by-4.0Dec 2017View details →
zenodo44/100

Comparative wordlist prompts for Australian languages

<p>An aligned version of three wordlists, Sutton &amp; Walsh, Curr, and Bates. See also David Nash&#39;s item (https://zenodo.org/record/1476467#.W9uXR3ozbUI) of Various Australian wordlist schemes.</p>

opencc-by-sa-4.0Nov 2018View details →
zenodo44/100

ACTIV-ES: a comparable Spanish corpus comprised of film dialogue from Argentine, Mexican and Spanish productions

<p><strong>DESCRIPTION</strong>: ACTIV-ES is a comparable Spanish corpus comprised of film dialogue from Argentine, Mexican and Spanish productions. Titles for each of these three countries were seeded from the Internet Movie Database, subtitle data for the hearing impaired was provided by Opensubtitles.org and was post-processed to correct/remove subtitle, OCR and diacritic artifacts and annotated for part-of-speech.</p> <p>The data is available in two main formats: 1) running text for each document and 2) 1:5 gram aggregate files. Each format includes a plain text and part-of-speech annotated version. Document names reflect the language code, country, year, title, type, genre (first genre listed in the IMDb), and IMDb ID.</p> <p>For more information about the development and evaluation of these resources and to cite this work refer to:</p> <p>Francom, J., Hulden, M. and Ussishkin, A.. (2014) ACTIV-ES: a comparable, cross-dialect corpus of &#39;everyday&#39; Spanish from Argentina, Mexico, and Spain. In Proceedings of the Ninth Annual Language Resources and Evaluation Conference, Reykjavik, Iceland. European Language Resources Association (ELRA).</p> <p>In <strong>version .02</strong> of the tagged running format corpus in the /eagles directory has been added which includes the EAGLES tagset. This tagset is much more fleshed out than the simplified tagset in the /tagged directory. For information on the tagset refer here: <a href="http://nlp.lsi.upc.edu/freeling/doc/tagsets/tagset-es.html">http://nlp.lsi.upc.edu/freeling/doc/tagsets/tagset-es.html</a>.</p>

opengpl-2.0Nov 2018View details →
zenodo44/100

Comparative Dataset on Migration

<p>The purpose of this dataset is to provide a systematic set of standardised contextual (economic,&nbsp;socio-political, cultural and legal) indicators in order to identify and measure on a comparative&nbsp;basis those contextual factors that have an (beneficial or inhibiting) impact on European, but not&nbsp;exclusively, responses to mass migration. Attention has been paid to existing socio-economic&nbsp;conditions and to national policies related to immigrants and asylum seekers. In this respect, the&nbsp;dataset comprises a set of both macro-level indicators measuring the socio-economic, political and&nbsp;institutional context of migration and cultural &ndash; or individual-level &ndash; indicators addressing ordinary&nbsp;citizens&rsquo; subjective attitudes, behaviours and perceptions about migration related-phenomena (e.g.&nbsp;perceived discrimination on ethnic grounds; immigration being bad or good for a country&#39;s&nbsp;economy; a country&#39;s cultural life being undermined or enriched by immigration).</p>

opencc-by-4.0Dec 2017View details →
zenodo44/100

Data used in paper "A comparative study of calibration methods for low-cost ozone sensors in IoT platforms"

<p>Data used in paper &quot;A comparative study of calibration methods for low-cost ozone sensors in IoT platforms&quot;, submitted for publication. The data consists of: (i) raw data from three nodes with four MICS 2614 metal-oxide ozone sensors deployed in Spain, summer 2017, and (ii) raw data of five alphasense OX-B431 and NO2-B43F electro-chemical sensors, four deployed in Italy and one in Austria, summers 2017 and 2018. Moreover, we have added the calibrated data using four machine learning methods: Multiple Linear Regression (MLR), K-Nearest Neighbors (KNN), Random Forest (RF) and Support Vector Regression (SVR).</p>

opencc-by-4.0Dec 2018View details →
zenodo44/100

Comparatives, Quantifiers, Proportions

<p>The present work investigates whether different quantification mechanisms (set comparison, vague quantification, and proportional estimation) can be jointly learned from visual scenes by a multi-task computational model. The motivation is that, in humans, these processes underlie the same cognitive, non-symbolic ability, which allows an automatic estimation and comparison of set magnitudes. We show that when information about lower complexity tasks is available, the higher-level proportional task becomes more accurate than when performed in isolation. Moreover, the multi-task model is able to generalize to unseen combinations of target/non-target objects. Consistently with behavioral evidence showing the interference of absolute number in the proportional task, the multi-task model no longer works when asked to provide the number of target objects in the scene.</p>

opencc-by-4.0Jun 2018View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record