Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

281

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

281 results for “source code”

Learn how ShareScore rates datasets ↗
dryad36/100

Source code for dynamic models and simulations of mate sampling behavior

Open the record for dataset details and reuse information.

publicApr 2022View details →
dryad36/100

Code from: Source–sink dynamics explains the coexistence of the invasive pest <em>Dryocosmus kuriphilus</em> and its biological control agent <em>Torymus sinensis</em> across French Eastern Pyrenees

Open the record for dataset details and reuse information.

publicDec 2025View details →
dryad36/100

Codes and source data files for: Proximity labeling identifies LOTUS domain proteins that promote the formation of perinuclear germ granules in C. elegans

Open the record for dataset details and reuse information.

publicNov 2021View details →
dryad36/100

Data and source code from: Contingency and selection in mitochondrial genome dynamics

Open the record for dataset details and reuse information.

publicMay 2022View details →
zenodo32/100

Dataset and Code for "Mining and Predicting Micro-Process Patterns of Issue Resolution for Open Source Software Projects"

<p>Dataset and Code for &quot;Mining and Predicting Micro-Process Patterns of Issue Resolution for Open Source Software Projects&quot; with README included</p>

opencc-by-4.0Jan 2020View details →
zenodo32/100

Open source physiological data and physiological-based kinetic model code for the chicken (Gallus gallus domesticus)

<p>This excel file and mode code (DOI:10.5281/zenodo.3603114) provides:</p> <p>1. Physiological parameters and associated inter-individual variability (sample size, mean, coefficient of variation,) for chicken (<em>Gallus gallus domesticus</em>). These physiological parameters were estimated based on the results of extensive literature searches and specific experimental data described in Lautz et al., (2020).</p> <p>2. An R code for the generic chicken physiologically based model as well as the &ldquo;soboljansen&rdquo; code to carry out sensitivity analysis using sobol plots. The code for the generic model allows to run:</p> <p>a. A deterministic PBK model which represents only a single animal.</p> <p>b. A probabilistic PBK model to simulate individual differences in physiological parameters within a population. Sensitivity analyses can be performed to identify which parameters have the most impact on the model&rsquo;s outputs. Predictions can be compared with experimental data. The model can be used to assess the influence of physiological parameters on the kinetics of chemicals. For PBK modelling purposes, species and chemical specific kinetics (e.g clearance, absorption rate, etc&hellip;) should be provided by the user.</p> <p>The full data collection and implementation of the models using case studies are described in (Lautz et al., 2020).</p> <p><strong>The dataset providing the physiological parameters is available in Excel.<br> The R code is presented as meta data to be implemented in R.</strong></p>

opencc-by-4.0Jan 2020View details →
zenodo32/100

Datasets for: Semantic Robustness of Models of Source Code

<p>Datasets for Semantic Robustness of Models of Source Code.</p> <p>Includes the c2s/java-small, csn/java, csn/python, and sri/py150 in the following representations:</p> <ol> <li>Raw [in raw.tar.gz]</li> <li>Normalized [in normalized.tar.gz]</li> <li>Pre-processed (<em>ast-paths and tokens</em>) [in preprocessed.tar.gz]</li> <li>Transformed [in transformed.tar.gz] <ol> <li>Normalized <ol> <li>transforms.All</li> <li>transforms.ShuffleLocalVariables</li> <li>transforms.ShuffleParameters</li> <li>transforms.RenameLocalVariables</li> <li>transforms.RenameFields</li> <li>transforms.RenameParameters</li> <li>transforms.ReplaceTrueFalse</li> <li>transforms.InsertPrintStatements</li> <li>transforms.Identity</li> </ol> </li> <li>Pre-processed (<em>ast-paths and tokens</em>) <ol> <li>transforms.Identity</li> <li>transforms.InsertPrintStatements</li> <li>transforms.ReplaceTrueFalse</li> <li>transforms.RenameParameters</li> <li>transforms.RenameFields</li> <li>transforms.RenameLocalVariables</li> <li>transforms.ShuffleParameters</li> <li>transforms.ShuffleLocalVariables</li> <li>transforms.All</li> </ol> </li> </ol> </li> </ol>

opencc-by-4.0Feb 2020View details →
zenodo32/100

Replication Package: A Study on the Accuracy of OCR Engines for Source Code Transcription from Programming Screencasts

<p>The replication package of the paper &quot;A Study on the Accuracy of OCR Engines for Source Code Transcription from Programming Screencasts&quot;&nbsp;including the dataset, results and tools</p>

opencc-by-4.0Oct 2020View details →
zenodo32/100

Model input and output, performance measures and modifications in the source code for PALM simulations on Mäkelänkatu in Helsinki, Finland

<p>This dataset is for air quality simulations conducted on M&auml;kel&auml;nkatu in Helsinki, Finland, using the PALM model system 6.0.</p> <p>By default, simulations use modelled data as boundary conditions. This includes modelled meteorological data from MEPS (MetCoOp Ensemble Prediction System) and air pollutant background concentrations from the ADCHEM model. Alternatively, measured meteorology from the Kivenlahti mast in Espoo, Finland, and aerosol size distribution from the SMEAR III station in Kumpula, Helsinki, is applied.</p> <p>Simulations:</p> <ul> <li>9 June morning: <ul> <li>0609_morning: use modelled boundary conditions for meteorology and air pollutants</li> <li>0609_morning_allmet_smear: use measured boundary conditions for meteorology and aerosol size distribution</li> <li>0609_morning_smear: use modelled boundary conditions for meteorology and measured for aerosol size distribution</li> <li>0609_morning_wd_smear: use modelled boundary conditions for meteorology, but modify the wind direction by using the measured wind direction at Kivenlahti. For aerosol size distribution, use the measured boundary conditions.</li> <li>0609_morning_wdk_smear: use modelled boundary conditions for meteorology, but modify the wind direction by using the measured wind direction at SMEAR III. For aerosol size distribution, use the measured boundary conditions.</li> <li>precursor_0609_morning: precursor with&nbsp;modelled boundary conditions for meteorology</li> <li>precursor_0609_morning_allmet: precursor with&nbsp;measured boundary conditions for meteorology</li> <li>precursor_0609_morning_wd: precursor with&nbsp;modelled boundary conditions, but&nbsp;the wind direction is modified by using the measured wind direction at&nbsp;Kivenlahti.</li> <li>precursor_0609_morning_wdk: precursor with&nbsp;modelled boundary conditions, but&nbsp;the wind direction is modified by using the measured wind direction at SMEAR III.</li> </ul> </li> <li>9 June evening: <ul> <li>0609_evening: use modelled boundary conditions for meteorology and air pollutants</li> <li>0609_evening_allmet_smear: use measured boundary conditions for meteorology and aerosol size distribution</li> <li>precursor_0609_evening: precursor with&nbsp;modelled boundary conditions for meteorology</li> <li>precursor_0609_evening_allmet: precursor with&nbsp;measured boundary conditions for meteorology</li> </ul> </li> <li>12 December morning: <ul> <li>1207_morning: use modelled boundary conditions for meteorology and air pollutants</li> <li>1207_morning_allmet_smear: use measured boundary conditions for meteorology and aerosol size distribution&nbsp;</li> <li>precursor_1207_morning: precursor with&nbsp;modelled boundary conditions for meteorology</li> <li>precursor_1207_morning_allmet: precursor with&nbsp;measured boundary conditions for meteorology</li> </ul> </li> </ul> <p>&nbsp;</p> <p>Datasets are given separately for the root (no suffix), parent (suffix _N02) and child (_N03) domain. The content is following:</p> <ul> <li>input_monitoring_output_usercode <ul> <li>Input data <ul> <li>&lt;run_identifier&gt;_chemistry: emission data for gases</li> <li>&lt;run_identifier&gt;_dynamic: initialisation and forcing data for meteorological variables and air pollutants</li> <li>&lt;run_identifier&gt;_p3d: parameter file for model steering</li> <li>&lt;run_identifier&gt;_salsa: emission data for aerosol particles</li> <li>&lt;run_identifier&gt;_static: topography information</li> </ul> </li> <li>Simulation performance information <ul> <li>&lt;run_identifier&gt;_cpu: information on the CPU time consumed</li> <li>&lt;run_identifier&gt;_header: information about the selected model parameters</li> <li>&lt;run_identifier&gt;_rc: time step control output</li> </ul> </li> <li>Output data <ul> <li>&lt;run_identifier&gt;_av_masked_N03_M01.nc: temporally averaged wind speed data close to the ground</li> <li>&lt;run_identifier&gt;_av_masked_N03_M04.nc: temporally averaged aerosol particle concentration data close to the ground</li> <li>&lt;run_identifier&gt;_av_masked_N03_M06.nc: temporally averaged aerosol particle concentration data in a vertical column next to the air quality monitoring station on M&auml;kel&auml;nkatu</li> <li>&lt;run_identifier&gt;_av_masked_N03_M07.nc: temporally averaged aerosol particle concentration data in a vertical column on the other side of the street from the air quality monitoring station on M&auml;kel&auml;nkatu</li> <li>&lt;run_identifier&gt;_pr.nc: temporally vertical profile data on meteorological variables</li> <li>&lt;run_identifier&gt;_ts.nc: flow statistics data</li> </ul> </li> <li>Modifications made to the source code (PALM revision, https://palm.muk.uni-hannover.de/trac/browser?rev=4416, last access: 10 Sept 2019) <ul> <li>chem_gasphase_mod.f90: chemical mechanism salsa+simple</li> <li>chem_emissions_mod.f90 (modifications indicated with &quot;MONA&quot;)</li> <li>user_module.f90 (modification listed under &quot;Current revisions&quot;)</li> <li>Makefile (modification listed under &quot;Current revisions&quot;)</li> </ul> </li> </ul> </li> </ul> <p>See the PALM model webpage (https://palm.muk.uni-hannover.de) for details.</p> <p>&nbsp;</p>

opencc-by-4.0May 2020View details →
zenodo32/100

Dynamic Load Balancing for Predictions of Storm Surge and Coastal Flooding-Model setup and source code

<p>Source code&nbsp;and model setup/inputs&nbsp;for the paper titled &quot;Dynamic Load Balancing for Predictions of Storm Surge and Coastal Flooding&quot; article.&nbsp; Simulations were conducted using a modified version of ADCIRC+DLB (ADCIRC + Dynamic Load Balancing)&nbsp;on unstructured triangular meshes.</p> <p>Contains:</p> <ol> <li>Model input files. <ol> <li>ADCIRC model input files for the ideal channel setup and Hurricane Irene simulation (*.13, *.14, *.15)</li> </ol> </li> <li>Zipped archive of the ADCIRC code (adcirc-cg-DLB.zip) used to produce the simulations for the paper.</li> <li>Step-by-step compilation&nbsp;and usage instructions for ADCIRC+DLB.&nbsp; <ol> <li>Installation.html&nbsp;</li> <li>Usage.html</li> </ol> </li> </ol>

opencc-by-4.0Jul 2020View details →
zenodo32/100

Amory et al. (2021), Geoscientific Model Development : data, model outputs and source code

<p><strong>Data and model outputs for the replication of the analysis made in:</strong><br> (see the published version of this article in Geoscientific Model Development,&nbsp;2021 - please cite this version if you use these data)<br> C. Amory, C. Kittel, L. Le Toumelin, C. Agosta,&nbsp;A. Delhasse, V. Favier,&nbsp;and X. Fettweis: Performance of MAR (v3.11) in simulating the drifting-snow climate and surface mass balance of Adelie Land, East Antarctica, Geoscientific Model Development, accepted, 2021.&nbsp;</p> <p>See README.txt for a full description of the dataset content</p> <p>Please contact me at&nbsp;amory.charles@live.fr&nbsp;if you need other half-hourly outputs or for more details on the dataset</p>

opencc-by-4.0Dec 2020View details →
zenodo32/100

Test cases (input), test scripts (output), and source code

<p>Test cases (input), test scripts (output), and source code used in the paper &quot;NLP-assisted Test Generation Toward Script-free Web Testing&quot; submitted to ICSME2021 NIER track</p>

opencc-by-4.0Jun 2021View details →
zenodo32/100

Supporting data and source code for Hnilica et al. (submitted to HESS)

<p>Data and code to reproduce the results and plots presented in Technical note: Changes of cross- and auto-dependence structures in climate projections of daily precipitation and their sensitivity to outliers (submitted to Hydrology and Earth System Sciences)</p>

opencc-by-4.0Jul 2017View details →
zenodo32/100

Refactoring Code Smells in Open Source Projects: A Hands-on Approach to Teaching Software Maintenance

<p>Code smells are suboptimal code structures that can undermine software quality and maintainability. On the one&nbsp;hand, software engineers commonly apply refactoring techniques to address these deficiencies and improve internal quality attributes. On the other hand, when performed manually and without discipline, refactoring can lead to&nbsp;code degradation. Despite its importance, refactoring and code smells are rarely explored in depth in undergraduate&nbsp;computing courses, which can be reflected in industry practices. To address this gap, this paper presents a hands-on approach to teaching code smell refactoring through contributions to Open Source Software (OSS) projects, an&nbsp;environment where developers with diverse skill levels collaborate, and maintaining code quality is particularly&nbsp;challenging. Code smells accumulate over time in such scenarios, hindering software evolution and collaboration.&nbsp;Our study in two undergraduate Software Quality and Software Maintenance courses expands on previous findings&nbsp;by incorporating an in-depth analysis of students&rsquo; learning experiences. The results indicate that: (i) students rec-&nbsp;ognized improvements in code quality after refactoring; (ii) they identified strong connections between refactoring,&nbsp;testing, and debugging; (iii) their confidence decreased when refactoring required changes across multiple files; (iv)&nbsp;code complexity posed a significant challenge to refactoring; (v) students&rsquo; choices of refactoring techniques were&nbsp;influenced by project structure and personal preferences, often combining multiple techniques to address a single&nbsp;smell; (vi) in some cases, refactoring introduced new code smells; (vii) the longest refactoring efforts were also&nbsp;the most likely to reintroduce code smells; (viii) contributing to OSS projects improved students&rsquo; programming&nbsp;skills and fostered a sense of professional growth; (ix) students faced challenges in understanding OSS contribution&nbsp;processes, particularly regarding issue resolution, adherence to contribution guidelines, and responding to maintainer feedback; (x) automated checks and review workflows varied across projects, affecting students&rsquo; ability to&nbsp;submit successful contributions; and (xi) despite these challenges, engagement with OSS enabled students to gain&nbsp;practical experience in collaborative software development. Our findings offer valuable insights for software engineering educators seeking to integrate refactoring practices into coursework while leveraging OSS contributions as&nbsp;an educational tool.</p>

opencc-by-4.0Jul 2024View details →
zenodo32/100

Inlist and Source Code Files for "Fossil Signatures of Main-sequence Convective Core Overshoot Estimated through Asteroseismic Analyses"

<p>MESA (r12778) and GYRE (version 6.0) inlist files used in the work&nbsp;described in "Fossil Signatures of Main-sequence Convective Core Overshoot Estimated through Asteroseismic Analyses".</p><p>Two subdirectories are provided in the&nbsp;archive:</p><p>1) The directory called "inlists" contains different MESA inlist files for different evolutionary period (pms=pre main sequence, ms=main sequence, and rgb=red giant branch) of the stellar model. The file named "inlist_0all" is applied to all evolutionary periods. The different inlist files for the different evolutionary states are called by putting their names in the "inlist" file, for example the include file named "inlist" evolves a MESA model from the pre main sequence until ZAMS (using the stop_near_zams = .true. option in the "inlist_1pms" file). An example gyre (version 6.0) inlist is also included in the "inlists" subdirectory.&nbsp; The Python scripts used to evaluate the matrix elements (which are used to determine the dipolar mixed-mode frequencies) discussed in this work are available at https://gitlab.com/darthoctopus/mesatricks.&nbsp;</p><p>2) Custom stopping conditions (used for stopping a model before the red giant branch) as well as a custom diffusion cutoff (see Viani et al. 2018, ApJ, 858, 28) are included in the run_star_extras.f file in the "src" subdirectory.&nbsp;</p>

opencc-by-4.0Nov 2023View details →
zenodo32/100

Source code and model outputs for 'Increasing Aerosol Direct Effect Despite Declining Global Emissions'

<p>Source code, model outputs and python scripts used for the publication of Hermant, A., Huusko, L., &amp; Mauritsen, T: Increasing Aerosol Direct Effect Despite Declining Global Emissions</p>

opencc-by-4.0Nov 2023View details →
zenodo32/100

The evaluation data and source codes of a new conceptual coupled Earth system model and the MOC box model.

<p>The dataset contains the results of a conceptual Atmosphere-Ocean-Ice-Land coupled Earth system model and a MOC box model and the evaluation data of their.</p>

opencc-by-4.0Jan 2024View details →
zenodo32/100

X-ray Fluorescence Core Scanning Dataset and Calibration Source Code

<p>This repository includes the raw output from X-ray fluorescence (XRF) core scanning, reference element concentration data, and the source code used for calibrating XRF counts. It is associated with the following publication:</p> <p>Kabiri S, Holden NM, Flood RP, Turner JN, O&rsquo;Rourke SM. X-ray Fluorescence Core Scanning for&nbsp;High-Resolution Geochemical Characterisation of Soils. Soil Systems. 2024; 8(2):56.&nbsp;https://doi.org/10.3390/soilsystems8020056.</p> <p>This research was funded by the Irish Research Council, grant no. IRCLA/2017/137.</p>

opencc-by-4.0Apr 2024View details →
zenodo32/100

Build Prediction in Continuous Integration Using Textual Analysis of Source Code and Traditional Software Metrics

<p>Continuous Integration (CI) systems integrate code changes committed by software developers, tests the results of the integration, and feed developers with information about the outcome of the integration and testing. Predicting the outcome of the integration is important since it reduces the feedback time between the CI system and the developers. This data-set comprises of historical code changes extracted from the TravisTorrent data-set&nbsp;(found in the train-lines folder) and their corresponding feature vectors (found in the train-bag-of-words folder) for Java projects. It also includes a set of files that contains historical build records and a set of traditional software metrics.</p>

opencc-by-4.0Jan 2022View details →
zenodo32/100

Dataset and model codes for "Large daytime molecular chlorine missing source at a suburban site in East China"

<p>This dataset is for measurements of trace gases (Cl<sub>2</sub>, ClNO<sub>2</sub>, N<sub>2</sub>O<sub>5</sub>, HONO, NO<sub>2</sub>, NO, O<sub>3</sub>, NH<sub>3</sub>, SO<sub>2</sub>, CO, VOCs), aerosols (Cl<sup>&minus;</sup>, NO<sub>3</sub><sup>&minus;</sup>, SO<sub>4</sub><sup>2&minus;</sup>, NH<sub>4</sub><sup>+</sup>, organics), and meteorological parameters (<em>T</em>, RH, <em>j</em><sub>NO2</sub>) at a suburban site (32.12&deg;N, 118.95&deg;E) in Nanjing, China during April 13-20, 2018.</p>

opencc-by-4.0Jan 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record