Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
1,582
datasets available to search
ShareScore release 0.9.0
Dataset results
1,582 results for “manuscript”
EEG Dataset and Processing script Associated with the Manuscript, "Using the Time-varying Drift Rate, the Signal Suppression and the Utility Maximisation to Account for Road Crossing Decisions".
<p>The repository archived the behavioural data (zip files) associated with the EEG experiment and the MATLAB script for EEG processing. BDFs stored the down-sampled EEG. </p>
Seq2science manuscript supplementary data
<p>Supplementary data for the seq2science manuscript.</p> <p>Includes a description on how to install seq2science, and for each re-analysis its configuration, samples file, and QC report.</p>
Database Figure 6 manuscript GCB 23-1473.R1 P_Pacheco
<p>Figure 6 of the article corresponding to annual mean values of NEE (g CO2 m-2 y-1) obtained from various studies in relation to different direct anthropogenic disturbances performed on Sphagnum, and management actions that favor moss reestablishment </p>
Raw Dataset for Mg2V12 manuscript
<p>Experimental datasets for manuscript "Coupled reaction equilibria enable the light-driven formation of metal-functionalized molecular vanadium oxides"</p>
Raw data for manuscript: Use of immunology in news and YouTube videos in the context of COVID-19: politicization and information bubbles
<p>Coding of newsarticles and videos related to immunology and COVID-19 in Italian and English</p>
Dataset for the manuscript "Helical and nonhelical dynamos in thin accretion discs"
<p>Dataset and post-processing scripts for the manuscript "Helical and nonhelical dynamos in thin accretion discs".</p>
Data used in the recently submitted AGU manuscript "Global Mantle Conductivity Imaging using 3-D GDS Inversion with Real Earth Surface Conductivity Constraint"
<p>Main data used in the recently submitted AGU manuscript "Global Mantle Conductivity Imaging using 3-D GDS Inversion with Real Earth Surface Conductivity Constraint"</p>
Zip archive containing datasets described in the manuscript entitled "Distribution-agnostic Deep Learning Enables Accurate Single‐Cell Data Recovery and Transcriptional Regulation Interpretation"
<p>The datasets used in the manuscript entitled "Distribution-agnostic Deep Learning Enables Accurate Single‐Cell Data Recovery and Transcriptional Regulation Interpretation". These datasets encompass all the experiments conducted in the manuscript, including simulation experiments, downsampling experiments, clustering, differential expression analysis, enrichment analysis, trajectory inference, batch correction, and clinical case discovery.</p> <p>The open-source software is available at https://github.com/XuYuanchi/Bis.</p>
The Middle Dutch Manuscripts Surviving from the Carthusian Monastery of Herne (14th century)
<p>This repository contains the dataset described in the following conference paper:</p><blockquote><p>Wouter Haverals & Mike Kestemont, "The Middle Dutch Manuscripts Surviving from the Carthusian Monastery of Herne (14th century): Constructing an Open Dataset of Digital Transcriptions". CHR 2023: Computational Humanities Research Conference. December 6-8, 2023, Paris, France.</p></blockquote><p>The dataset consists of (automatically created) hyper-diplomatic, digital transcriptions of 18 Middle Dutch manuscripts that survive from the carthusian monastery in Herne in nowadays Belgium (or manuscripts which have meaningful ties with the charterhouse). These manuscripts primarily date to the second half of the fourteenth century and offer exciting possibilities for the analysis of authorship, translatorship and scribal practices in the history of the Low Countries. The transcriptions have been (partly) automated through the use of handwritten text recognition (on the Transkribus platform). This dataset is licensed under a CC-BY 4.0 licence, encouraging the further re-use of this data for all purposes, provided an unambiguous scholarly reference to the paper above is given.</p><p><strong>Content</strong></p><p>Transcriptions for the following 18 manuscripts are included in various formats:</p><ul><li>Brussels, RL, 1805-1808</li><li>Brussels, RL, 2485</li><li>Brussels, RL, 2849-51</li><li>Brussels, RL, 2877-78</li><li>Brussels, RL, 2879-80</li><li>Brussels, RL, 2905-09</li><li>Brussels, RL, 2979</li><li>Brussels, RL, 3091</li><li>Brussels, RL, 3093-95</li><li>Ghent, UL, 1374</li><li>Ghent, UL, 941</li><li>Paris, Bibl. Mazarine, 920</li><li>Paris, Bibl. de l'Arsenal, 8224</li><li>Saint Petersburg, BAN, O 256</li><li>Vienna, ÖNB, SN 12.857</li><li>Vienna, ÖNB, SN 12.905</li><li>Vienna, ÖNB, Cod. 13.708</li><li>Vienna, ÖNB, SN 65</li></ul><p>The contents of the repository have been structured as follows:</p><ul><li><i>transcriptions</i>: transcriptions of the 18 manuscripts in various formats (hyper-diplomatic; i.e. without brevigraph expansion):<ul><li>pagexmls: One file per folium, encoded in the PAGEXML format as outputted by Transkribus. One zip-file per manuscript folder.</li></ul></li><li><i>spreadsheets.zip</i>: detailed metadata on various aspects of the data in spreadsheat format.<ul><li>silent_voices_summary.xlsx: summary statistics at the codex-level (cf. Table 2 in the paper)</li><li>codex_info.xlsx: folium-level metadata</li><li>manuscript_data_metadata.xlsx: text region-level metadata</li><li>manuscript_data_metadata_rich.xlsx: contains the most convenient and complete version of the dataset, including the texts with automatically expanded abbreviations and the linguistic enrichment (lemma's and part-of-speech tags).</li></ul></li><li><i>code</i>: Python notebooks (requiring Python >= 3.8).<ul><li>transduction.ipynb: the notebook for the replication of the abbreviation expansion experiments described in the paper. (See also the configuration file for there tagger norm.json.</li><li>enrich.ipynb: the notebook used for the linguistic enrichment of the expanded texts, on the basis of the PIE(-NLP) lemmatizer. (See also the PIE model file herne-norm.tar, which is used in the enrichment.)</li><li>requirements.txt: third-party dependencies for running the code in these notebooks. Note: enrich.ipynb will require you the Middle Dutch (DUM) model for nlp-pie.</li></ul></li></ul><p><strong>Related data</strong></p><ul><li>The final Transkribus model used to generate the transcriptions will be make publicly available on the platform.</li><li>The accompanying images are released in a separate, restricted access repository on Zenodo, because we were unable to clear the copyright on some of the facsimiles. We will only be able to share these images under very strict conditions.</li></ul><p><strong>Acknowledgments</strong></p><p>Thanks to Anouck Kuypers, Sam Verellen and Frans de Jonge for their work on the transcriptions. The transcription of Brussels, RL, 3093-95 was contributed by Dr. Ine Kiekens. We acknowledge the help of Renée Gabriël and Peter Boot in previous collaborations that relate to the present paper. Finally, we would like to thank Caroline Vandyck who has helped with the finalization of the dataset.</p><p><strong>Funding statement</strong></p><p>This work has been funded by the Flemish Research Agency (FWO) in the context of the project "Silent voices: A Digital Study of the Herne Charterhouse as a Textual Community (ca. 1350-1400)".</p>
Fake Manuscript 5
https://www.artbreeder.com/beta/image/dd09d35e8d4023a2e7505972c6ad Source: Objaverse 1.0 / Sketchfab
Raw ptychography data for manuscript with title 'Multiplexing limitis in ptychography'
<p>- Given are the HDF files containing the raw diffraction patterns, and the reconstructed files.<br>- The metadata for the reconstruction of each dataset is stored in the raw files, and in the reconstructed files<br>- For each scenario, the Fourier ring correlation (FRC) was calculated against the reconstruction with a single beam as reference, and against the corresponding region for the simulated dataset.</p> <p>The reconstruction of these datasets was carried out with ptylab:<br>Loetgering, L., Du, M., Boonzajer Flaes, D., Aidukas, T., Wechsler, F., Penagos Molina, D. S., Rose, M., Pelekanidis, A., Eschen, W., Hess, J., Wilhein, T., Heintzmann, R., Rothhardt, J., & Witte, S. (2023). PtyLab.m/py/jl: a cross-platform, open-source inverse modeling toolbox for conventional and Fourier ptychography. In Optics Express (Vol. 31, Number 9, pp. 13763–13797).<br>Zenodo. https://doi.org/10.5281/zenodo.8287047<br>Github: https://github.com/PtyLab/PtyLab.py</p> <p>EDIT:<br>The first version of this repository corresponds to the datasets used in the preprint:<br>Molina, Daniel Penagos; Eschen, Wilhelm; Liu, Chang; Limpert, Jens; Rothhardt, Jan (2024). Multiplexing information limits in ptychography. Optica Open. Preprint. https://doi.org/10.1364/opticaopen.25941427.v2</p> <p>A second set of datasets was added. These correspond to the experimental results in the final version of the published paper:<br>Daniel S. Penagos Molina, Wilhelm Eschen, Chang Liu, Jens Limpert, and Jan Rothhardt, "Multiplexing information limits in multi-beam ptychography," Opt. Express 33, 12925-12938 (2025)</p> <p>- Data were measured at the Institute of Applied Physics in Jena.</p> <p>Please contact me (santiago.penagos@uni-jena.de) for additional support.</p>
Data and Code for replicating the manuscript submitted to World Development
<p>Data and Code for replicating the paper "A Data-Driven Approach Improves Food Insecurity Crisis Prediction"</p>
The influence of the global COVID-19 pandemic on manuscript submissions and editor and reviewer performance at six ecology journals
Open the record for dataset details and reuse information.
Dataset for manuscript entitled: Switchgrass cropping systems affect soil carbon and nitrogen and microbial diversity and activity on marginal lands
Open the record for dataset details and reuse information.
Colistin kills bacteria by targeting lipopolysaccharide in the cytoplasmic membrane - primary data for all experiments described in the manuscript
Open the record for dataset details and reuse information.
Pharmacokinetics parameters, bioanalytical method underlying the manuscript: The effect of morning versus evening administration of empagliflozin on its pharmacokinetics and pharmacodynamics characteristics in healthy adults: a two-way crossover non-randomised trial
Open the record for dataset details and reuse information.
HIDRA simulations and post-processing scripts for JGR: SP manuscript: characterization of N+ abundances in the terrestrial polar wind using the multiscale atmosphere-geospace environment
Open the record for dataset details and reuse information.
Demographic data, reporting guidelines underlying the manuscript: The effect of morning versus evening administration of Empagliflozin on its pharmacokinetics and pharmacodynamics characteristics in healthy adults: a two way cross over, non-randomized trial
Open the record for dataset details and reuse information.
RNAseq data relating to manuscript: Prior Fc Receptor activation primes macrophages for increased sensitivity to IgG via long term and short term mechanisms.
Open the record for dataset details and reuse information.
Pharmacodynamics parameters underlying the manuscript: The effect of morning versus evening administration of empagliflozin on its pharmacokinetics and pharmacodynamics characteristics in healthy adults: a two-way crossover, non-randomised trial
Open the record for dataset details and reuse information.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.