Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

1,582

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

1,582 results for “manuscript”

Learn how ShareScore rates datasets ↗
zenodo36/100

Raw data for the submitted manuscript: The Influence of Soil Organic Matter Content on the Toxicity of Pesticides to the Springtail Folsomia candida

<p>Raw data obtained from toxicity tests with the springtail Folsomia candida exposed for 28 days to chlorpyrifos, lindane, cyproconazole, carbendazim and imidacloprid in artificial soils containing 10%, 5%, 2.5% sphagnum peat, and LUFA 2.2 soil. Tests were performed following OECD guideline 232. The file includes data on springtail survival and reproduction.</p>

opencc-by-4.0Oct 2024View details →
zenodo36/100

Raw data for the manuscript: The influence of soil organic matter content and substance lipophilicity on the toxicity of pesticides to the earthworm Eisenia andrei

<p>Raw data obtained from toxicity tests with the earthworm <em>Eisenia andrei</em> exposed for 56 days to chlorpyrifos, lindane, cyproconazole, carbendazim and imidacloprid in artificial soils containing 10%, 5%, 2.5% sphagnum peat, and LUFA 2.2 soil. Tests were performed following OECD guideline 222. The file includes data on earthworm starting and ending weights, survival, and reproduction.</p>

opencc-by-4.0Oct 2024View details →
zenodo36/100

Data for manuscript: An extrapolation algorithm for estimating river bed grain size distributions across basins

<p>Pebble counts collected and used for the analysis presented in the manuscript: An extrapolation agorithm for estimating river bed grain size distributions across drainage basins.</p>

opencc-by-4.0Nov 2024View details →
zenodo36/100

Supplementary Table 1-9 of the manuscript: Magmatic Cl-H2O contents, fluid extraction and porphyry fertility: Evidence from zircon and its apatite inclusions

<p><strong><span>Table DR1</span></strong><span> Major element composition of biotite from the three intrusions in the ZOF</span></p> <p><strong><span>Table DR2</span></strong><span> Laser Raman spectra and trace element compositions of zircon grains from the Cretaceous intrusions in the ZOF</span></p> <p><strong><span>Table DR3</span></strong><span> Major element composition of plagioclase grains from the Cretaceous intrusions in the ZOF</span></p> <p><strong><span>Table DR4</span></strong><span> Sr isotopic composition of the plagioclase from the Cretaceous intrusions in the ZOF</span></p> <p><strong><span>Table DR5</span></strong><span> Zircon Lu-Hf isotopic composition of the Sifang granodiorite, Luoboling granodiorite porphyry and Zhongliao granodiorite</span></p> <p><strong><span>Table DR6 </span></strong><span>Zircon water contents and O isotopic compositions by SIMS of the Cretaceous intrusions in the ZOF</span></p> <p><strong><span>Table DR7 </span></strong><span>Major and volatile element composition of the zircon-host apatite from the Cretaceous intrusions in the ZOF </span></p> <p><strong><span>Table DR8</span></strong><span> The top 20 best fitting runs</span></p> <p><strong><span>Table DR9</span></strong><span> Summarization of the salinity calculations</span></p>

opencc-by-4.0Nov 2024View details →
zenodo36/100

Data Files for Tresoldi/Robinson article on spelling variation in Canterbury Tales manuscripts

<p>This data is at <a href="https://github.com/peterrobinson/CTSpellingArticle2024">https://github.com/peterrobinson/CTSpellingArticle2024</a>. It is contained in three folders, each folder corresponding to one of the three sets of data used in this analysis, as follows:</p> <p><strong>&nbsp;</strong></p> <ol> <li> <p>&ldquo;sorted by regularization&rdquo;. This folder contains spelling data and results derived from the regularization process, where (for example) spellings of forms regularized to &ldquo;goode&rdquo; are distinguished from spellings of forms regularized to &ldquo;god&rdquo;;</p> </li> <li> <p>&ldquo;sorted by part of speech&rdquo;. This folder contains spelling data and results derived from a lemmatization and part-of-speech identification process, where (for example) spellings of forms lemmatized&nbsp; to &ldquo;goode&rdquo; singular adjective are distinguished from spellings of forms regularized to &ldquo;goode&rdquo; plural adjective, and forms lemmatized to &ldquo;gode&rdquo; singular noun nominative case are distinguished from &ldquo;gode&rdquo; singular noun oblique case (as in &ldquo;to gode&rdquo;).</p> </li> <li> <p>&ldquo;all spellings unsorted&rdquo;. This folder contains spelling data as undifferentiated counts of &ldquo;bags of words&rdquo;: for each witness: so many occurrences of &ldquo;good&rdquo;, so many of &ldquo;goode&rdquo;, so many of &ldquo;god&rdquo;, so many of &ldquo;gode&rdquo;.</p> </li> </ol> <p><strong>&nbsp;</strong></p> <p>Each folder contains the following files (under various names):</p> <p><strong>&nbsp;</strong></p> <ol> <li> <p>A . json file holding all the data, structured according to its categorization. The &ldquo;sorted by part of speech&rdquo; folder contains two .json files, one with spellings organized by headword lemma, the other organized by part-of-speech;</p> </li> <li> <p>Two .nex Nexus files containing all the data. In the &ldquo;sorted by regularization&rdquo; and&nbsp; &ldquo;sorted by part of speech&rdquo; folders one Nexus file groups spellings by variant sites within each line, the other Nexus file groups spellings&nbsp; by words within each line. In the &ldquo;all spellings unsorted&rdquo; folder one Nexus file contains all the spellings organized by spelling; the second holds a Nexus distance matrix with distances created according to the Manhattan distance algorithm;</p> </li> <li> <p>A .dst distance matrix file, containing a distance matrix constructed with distacnes calculated by the Manhattan distance algorithm;</p> </li> <li> <p>A &ldquo;features&rdquo; file, containing a spreadsheet ranking each variant site according to its impact on the analysis&nbsp;</p> </li> <li> <p>Multiple .pdf files visualizing the results of our analysis, with the names reflecting the analysis each contains. Files with names including &ldquo;Splits&rdquo; were created using the SplitsTree algorithm and software <a href="https://www.zotero.org/google-docs/?peaEav">(Huson and Bryant 2006; &ldquo;SplitsTree | Universit&auml;t T&uuml;bingen,&rdquo; n.d.)</a></p> </li> </ol> <p><strong>&nbsp;</strong></p> <p>The &ldquo;sorted by regularization&rdquo; folder also contains a single image file, &ldquo;tiagoplot1.jpg&rdquo;, visualizing the results of PCA analysis on the &ldquo;sorted by regularization&rdquo; data.</p> <p>&nbsp;</p>

opencc-by-4.0Nov 2024View details →
zenodo36/100

Simulation cases in the manuscript

Open the record for dataset details and reuse information.

opencc-by-4.0Nov 2024View details →
zenodo36/100

Data for "Shear Strain Evolution Spanning the 2020 Mw6.8 Elazığ and 2023 Mw7.8/Mw7.6 Kahramanmaraş Earthquake Sequence along the East Anatolian Fault Zone" manuscript

<p>Data necessary to support the analysis presented in "Shear Strain Evolution Spanning the 2020 Mw6.8 Elazığ and 2023 Mw7.8/Mw7.6 Kahramanmaraş Earthquake Sequence along the East Anatolian Fault Zone" manuscript</p>

opencc-by-4.0Nov 2024View details →
zenodo36/100

File data and Code for submitted manuscript to New Phytologist

<p>Original data from submitted manuscript: Nutrient resorption of leaves and roots coordinates with root nutrient-acquisition strategies in a temperate forest by Wang, Freschet, McCormack, Lambers, and Gu, submitted for publication in <em>New phytologist</em>.</p> <p>&nbsp;</p> <p>The data includes the leaf and root resorption traits (NRE_L, PRE_L, N_Ls, P_Ls, NRE_R, PRE_R, N_Rs, P_Rs), and root economic traits (RN, RP, RD, SRL, RTD) of 34 woody species in temperate forest.</p> <p>&nbsp;</p> <p>Abbreviations of the data:</p> <p>NRE_L, Nitrogen resorption efficiency of leaf (%)</p> <p>PRE_L, Phosphorus resorption efficiency of leaf (%)</p> <p>N_Ls, Nitrogen concentration of senesced leaf (mg g<sup>-1</sup>)</p> <p>P_Ls, Phosphorus concentration of senesced leaf (mg g<sup>-1</sup>)</p> <p>NRE_R, Nitrogen resorption efficiency of root (%)</p> <p>PRE_R, Phosphorus resorption efficiency of root (%)</p> <p>N_Rs, Nitrogen concentration of senesced root (mg g<sup>-1</sup>)</p> <p>P_Rs, Phosphorus concentration of senesced root (mg g<sup>-1</sup>)</p> <p>RN, root nitrogen concentration (mg g<sup>-1</sup>);</p> <p>RP, root phosphorus concentration (mg g<sup>-1</sup>);</p> <p>RD, root diameter (mm);</p> <p>SRL, specific root length (m g<sup>-1</sup>);</p> <p>RTD, root tissue density (g cm<sup>-3</sup>);</p> <p>&nbsp;</p> <p>The R files used to generate the results are provided respectively.</p>

opencc-by-4.0Oct 2024View details →
zenodo36/100

Supplementary material for Gender & Personality & Tears manuscript (Gadea et al. 2025)

<p>Supplementary material for the manuscript "The impact of gender stereotypes and observer personality on the perceived authenticity of tears". Consists of:&nbsp;</p><p>- The anonymized databases (options A &amp; B): only the age &amp; gender of the participants are shown as demographic data, together with the results of the Salamanca questionnaire and the answers to the questions regarding perceptions about type of emotion, sincerity, and intensity for each photograph presented. &nbsp;</p><p>- The 10 photographs that were used in the study (men and women), both the originals and the ones that were edited with the tear digitally erased.</p>

opencc-by-4.0Dec 2023View details →
zenodo36/100

Data associated with the Jaziri et al. 2022 manuscript

<p>Data used in Jaziri et al. 2022. 1D and 3D models for a methane abundance of 1e-4 and oxygen abundances of 1e-3, 1e-4, 1e-5, 1e-6 and 1e-7.</p>

opencc-by-4.0Sep 2022View details →
zenodo36/100

Dataset supporting the manuscript "Comprehensive evaluation of phosphoproteomic-based kinase activity inference"

<p>Datasets involved in the benchmarking of kinase activity inference as presented in the manuscript "Comprehensive evaluation of phosphoproteomic-based kinase activity inference".</p>

opencc-by-4.0Jun 2024View details →
zenodo36/100

Dataset related to manuscript: Rapid iPSC inclusionopathy models shed light on formation, consequence and molecular subtype of a-synuclein inclusions

<p>Key Resources Table and tabular data related to Lam, Ndayisaba et al 2024.</p> <p>Code for generating graphs can be found at: doi:10.5281/zenodo.12574231 (version 1), doi:10.5281/zenodo.12574230 (all versions)</p>

opencc-by-4.0Jun 2024View details →
zenodo36/100

Supplementary data for manuscript "Genetic risk converges on regulatory networks mediating early type 2 diabetes"

<p>Supplementary data for manuscript "Genetic risk converges on regulatory networks mediating early type 2 diabetes" Nature 624, 621&ndash;629 (2023). <a href="https://doi.org/10.1038/s41586-023-06693-2">https://doi.org/10.1038/s41586-023-06693-2</a></p> <p>Brief description of the included files is given below. Please visit the manuscript website for latest updates:&nbsp;<a href="http://theparkerlab.org/manuscripts/2021_islet-rfx6/">http://theparkerlab.org/manuscripts/2021_islet-rfx6/</a></p>

openMay 2022View details →
zenodo36/100

The Repository for the Manuscript "Temperature and Precipitation Dominate Seasonal Variations in Seismic Velocity and Attenuation in Deserts"

<p><strong><span>Overview</span></strong></p> <p><span>This dataset contains the essential code and data for calculating the Horizontal-to-Vertical Spectral Ratio (HVSR), analyzing vehicle-generated seismic events, retrieving Q-values, and comparing them with meteorological data. It also includes waveform data from 20 seismic events.</span></p> <p><span>The seismic data originate from a temporary broadband seismic array deployed in the Tarim Basin, from July 2017 to October 2019 (Zuo et al., 2022). This dataset focuses on three seismic stations: T12, T52, and T23. Stations T12 and T23 recorded data from July 2017 to October 2019, while station T52 recorded from November 2018 to October 2019.</span></p> <p><span>&nbsp;</span></p> <p><strong><span>Code</span></strong></p> <p><span>The dataset includes Python scripts for calculating HVSR and retrieving Q-values. The HVSR calculation follows Li et al., (2023), while forward modeling is based on Antonio Garc&iacute;a-Jerez et al., (2016).</span></p> <p><span>The codes for Q-value estimation are stored in &lsquo;Retrieving Q-value&rsquo; folder. The Q-value estimation process, demonstrated for station T12 in Jupyter Notebook, involves extracting single vehicle signals from continuous data, time-frequency spectrogram calculations, two-dimensional correlation coefficient of their time-frequency amplitude calculations, using hierarchical clustering algorithm to classify vehicle signals, vehicle speed estimation, and performing Q-value inversion.</span></p> <p><span>&nbsp;</span></p> <p><strong><span>Dataset </span></strong></p> <p><span>HVSR variations over time for three stations are calculated from continuous seismic recordings and are stored in the <em>&lsquo;HVSR&rsquo;</em> folder under each station directory. </span></p> <p><span>Time-frequency spectrograms for Q-value estimation are stored in the <em>&lsquo;Spectrogram&rsquo;</em> folder, with filenames indicating the record time of each vehicle signal. The Q-value is inverted using these signals, and for stability, we stacked every 100 individual results, which are stored in the 'Q-values' folder under the corresponding station name folder. Due to interference from wind and other sources, Q-value inversion using vehicle signals was unreliable for T23, so Q-values are only provided for T12 and T52.</span></p> <p><span>Meteorological data (temperature and soil water content) are stored in the <em>&lsquo;temperature&rsquo;</em> and <em>&lsquo;soil water content&rsquo;</em> folders under each station directory.</span></p> <p><span>Seismic event waveforms for 20 selected strong earthquakes are stored in the <em>&lsquo;events&rsquo;</em> folder, with filenames indicating the start and end times of the events.</span></p> <p><span>&nbsp;</span></p>

opencc-by-4.0Sep 2024View details →
zenodo36/100

Code and Data for Manuscript "How raw milk-based adjunct cultures influence the microbial diversity in cheese"

<p>This repository contains the data and code used in the study "How raw milk-based adjunct cultures influence the microbial diversity in cheese" by Dreier et al. (2024). The study investigates the impact of raw milk-based adjunct cultures (NMAC) on the microbial diversity of cheese. The data includes 16S rRNA gene amplicon sequences from cheese samples, as well as metadata and code used for data analysis.</p>

openmit-licenseOct 2024View details →
zenodo36/100

Data for Manuscript "Deep-Water Near-Inertial Waves and Turbulence on a Continental Slope in the South China Sea during Typhoon Mangkhut (2018)"

<p>Data for manuscript &quot;Deep-Water Near-Inertial Waves and Turbulence on a Continental Slope in the South China Sea during Typhoon Mangkhut (2018)&quot;.</p>

opencc-by-4.0Oct 2021View details →
zenodo36/100

In search for Old Greek readings: The Old Latin manuscript Palimpsestus Vindobonensis (La115) in 2 Samuel - the collation file for SBL 2021

<p>Collected cases for my SBL 2021 paper &quot;In search for Old Greek readings: The Old Latin manuscript Palimpsestus Vindobonensis (La115) in 2 Samuel&quot;</p>

opencc-by-4.0Nov 2021View details →
zenodo36/100

Data in support of manuscript "Impacts of storm surge barriers on drag, mixing, and exchange flow in a partially mixed estuary" submitted to JGR-Oceans

<p>Data set in support of manuscript &quot;Impacts of storm surge barriers on drag, mixing, and exchange flow in a partially mixed estuary&quot; submitted to JGR-Oceans in November 2021.&nbsp; Matlab script (makeFigs_barDragMix_upload.m) is used to generate the figures from the manuscript.&nbsp; Data files (*.mat) correspond with each figure (*.png).&nbsp; For questions or additional information please contact&nbsp;D. Ralston.</p>

opencc-by-4.0Nov 2021View details →
zenodo36/100

Data for manuscript "Reciprocal Radicalization: The Rise of Culture War Terminology in British and American News Coverage"

<p>This data set contains frequency counts of target words in 16&nbsp;million news and opinion articles from 10 popular news media outlets in the United Kingdom: The Guardian, The Times, The Independent, The Daily Mirror, BBC, Financial Times, Metro, Telegraph, The and The Daily Mail plus a few additional American-based outlets used for comparison reference. The target words are listed in the associated manuscript and are mostly words that denote some type of prejudice, social justice related terms or counterreaction to it. A&nbsp;few additional&nbsp;words are also available since they are used in the manuscript for illustration purposes.</p> <p>The textual content of news and opinion articles from the outlets listed in Figure 3&nbsp;of the main manuscript is available in the outlet&#39;s online domains and/or public cache repositories such as Google cache (https://webcache.googleusercontent.com), The Internet Wayback Machine (https://archive.org/web/web.php), and Common Crawl (https://commoncrawl.org). We derived relative frequency counts from these sources. Textual content included in our analysis is circumscribed to articles headlines and main body of text of the articles and does not include other article elements such as figure captions.</p> <p>Targeted textual content was located in HTML raw data using outlet specific xpath expressions.&nbsp;Tokens were lowercased prior to estimating frequency counts.&nbsp;To prevent outlets with sparse text content for a year from distorting aggregate frequency counts, we only include outlet frequency counts from years for which there is at least 1&nbsp;million words of article content from an outlet.&nbsp;</p> <p>Yearly frequency usage of a target word in an outlet in any given year was estimated by dividing the total number of occurrences of the target word in all articles of a given year by the number of all words in all articles of that year. This method of estimating frequency accounts for variable volume of total article output over time.</p> <p>The list of compressed files in this data set is listed next:</p> <p>-analysisScripts.rar contains the analysis scripts used in the main manuscript&nbsp;</p> <p>-targetWordsInArticlesCounts.rar contains counts of target words in outlets articles as well as total counts of words in articles</p> <p>-targetWordsInArticlesCountsGuardianExampleWords&nbsp;contains counts of target words in outlets articles as well as total counts of words in articles for illustrative Figure 1 in main manuscript</p> <p>Usage Notes</p> <p>In a small percentage of articles, outlet specific XPath expressions can fail&nbsp;to properly capture the content of the article due to the heterogeneity of HTML elements and CSS styling combinations with which articles text content is arranged in outlets online domains. As a result, the total and target word counts metrics for a small subset of articles are not precise. In a random sample of articles and outlets, manual estimation of target words counts overlapped with the automatically derived counts for over 90% of the articles.</p> <p>Most of the incorrect frequency counts were minor deviations from the actual counts such as for instance counting the word &quot;Facebook&quot; in an article footnote encouraging article readers to follow the journalist&rsquo;s Facebook profile and that the XPath expression mistakenly included as the content of the article main text. To conclude, in a data analysis of 16 million articles, we cannot manually check the correctness of frequency counts for every single article and hundred percent accuracy at capturing articles&rsquo; content is elusive due to the small number of difficult to detect boundary cases such as incorrect HTML markup syntax in online domains. Overall however, we are confident that our frequency metrics are representative of word prevalence in print news media content (see Figure 1 of main manuscript for supporting evidence).</p>

opencc-by-4.0Nov 2021View details →
zenodo36/100

Datasets to produce figures of the manuscript "Quantum Critical Points and the Sign Problem"

<p>Datasets to produce figures of the manuscript &quot;Quantum Critical Points and the Sign Problem&quot;</p> <p>Contents:</p> <p>All figures are divided by folders, Main_text (Figs. 1 to 4) and Supplemental Material (Figs. S1 to S11).&nbsp;</p> <p>In each folder the data files (.dat) and scripts (.py) that generate the figures are included. Data files have easily readable names, with relevant parameters explicitly given.</p> <p>Using Python 3+ with an updated Matplotlib can directly reproduce the figures in the manuscript. We further include the Figures (.pdf or .png) in the corresponding folders for convenience.</p> <p>Data can be reproduced using the QUEST: QUantum Electron Simulation Toolbox, freely available at https://www.cs.ucdavis.edu/~bai/QUEST_public/</p> <p>The geometry files (.geom) and example input files (.in) for the two types of lattices used are given in this repository.</p>

opencc-by-4.0Aug 2021View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record