Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

116

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

116 results for “code review”

Learn how ShareScore rates datasets ↗
zenodo28/100

Replication Package for the Paper: "Code Smells Detection via Code Review: An Empirical Study"

<p>This&nbsp;repository&nbsp;contains&nbsp;the&nbsp;data&nbsp;and&nbsp;results&nbsp;from&nbsp;the&nbsp;paper&nbsp;&quot;Code&nbsp;Smells&nbsp;Detection&nbsp;via&nbsp;Code&nbsp;Review:&nbsp;An&nbsp;Empirical&nbsp;Study&quot;&nbsp;submitted&nbsp;to&nbsp;ESEM&nbsp;2020.</p> <p>&nbsp;</p> <p><strong>1. data&nbsp;folder</strong></p> <p>The&nbsp;data&nbsp;folder&nbsp;contains&nbsp;the&nbsp;retrieved&nbsp;269&nbsp;reviews&nbsp;that&nbsp;discuss&nbsp;code&nbsp;smells.&nbsp;Each&nbsp;review&nbsp;includes&nbsp;four&nbsp;parts:&nbsp;Code&nbsp;Change&nbsp;URL,&nbsp;Code&nbsp;Smell&nbsp;Term,&nbsp;Code&nbsp;Smell&nbsp;Discussion,&nbsp;and&nbsp;Source&nbsp;Code&nbsp;URL.</p> <p>&nbsp;</p> <p><strong>2. scripts&nbsp;floder</strong></p> <p>The&nbsp;scripts&nbsp;folder&nbsp;contains&nbsp;the&nbsp;Python&nbsp;script&nbsp;that&nbsp;was&nbsp;used&nbsp;to&nbsp;search&nbsp;for&nbsp;code&nbsp;smell&nbsp;terms&nbsp;and&nbsp;the&nbsp;list&nbsp;of&nbsp;code&nbsp;smell&nbsp;terms.</p> <ul> <li><em>smell-term/general_smell_terms.txt</em>&nbsp;contains&nbsp;general&nbsp;code&nbsp;smell&nbsp;terms,&nbsp;such&nbsp;as&nbsp;&quot;code&nbsp;smell&quot;.</li> <li><em>smell-term/specific_smell_terms.txt</em>&nbsp;contains&nbsp;specific&nbsp;code&nbsp;smell&nbsp;terms,&nbsp;such&nbsp;as&nbsp;&quot;dead&nbsp;code&quot;.</li> <li><em>smell-term/misspelling_terms_of_smell.txt</em>&nbsp;contains&nbsp;the&nbsp;misspelling&nbsp;terms&nbsp;of&nbsp;&#39;smell&#39;,&nbsp;such&nbsp;as&nbsp;&quot;ssell&quot;.</li> <li><em>get_changes.py</em>&nbsp;is&nbsp;used&nbsp;for&nbsp;getting&nbsp;code&nbsp;changes&nbsp;from&nbsp;OpenStack.</li> <li><em>get_comments.py</em>&nbsp;is&nbsp;used&nbsp;for&nbsp;getting&nbsp;review&nbsp;comments&nbsp;for&nbsp;each&nbsp;code&nbsp;change.</li> <li><em>smell_search.py</em>&nbsp;is&nbsp;used&nbsp;for&nbsp;searching&nbsp;review&nbsp;comments&nbsp;that&nbsp;contain&nbsp;code&nbsp;smell&nbsp;terms.</li> </ul> <p>&nbsp;</p> <p><strong>3. project&nbsp;folder</strong></p> <p>The&nbsp;project&nbsp;folder&nbsp;contains&nbsp;the&nbsp;MAXQDA&nbsp;project&nbsp;files.&nbsp;The&nbsp;files&nbsp;can&nbsp;be&nbsp;opened&nbsp;by&nbsp;MAXQDA&nbsp;12&nbsp;or&nbsp;higher&nbsp;versions,&nbsp;which&nbsp;are&nbsp;available&nbsp;at&nbsp;https://www.maxqda.com/&nbsp;for&nbsp;download.&nbsp;You&nbsp;may&nbsp;also&nbsp;use&nbsp;the&nbsp;free&nbsp;14-day&nbsp;trial&nbsp;version&nbsp;of&nbsp;MAXQDA&nbsp;2018,&nbsp;which&nbsp;is&nbsp;available&nbsp;at&nbsp;https://www.maxqda.com/trial&nbsp;for&nbsp;download.</p> <ul> <li><em>Data&nbsp;Labeling&nbsp;&amp;&nbsp;Encoding&nbsp;for&nbsp;RQ2.mx12</em>&nbsp;is&nbsp;the&nbsp;results&nbsp;of&nbsp;data&nbsp;labeling&nbsp;and&nbsp;encoding&nbsp;for&nbsp;RQ2,&nbsp;which&nbsp;were&nbsp;analyzed&nbsp;by&nbsp;the&nbsp;MAXQDA&nbsp;tool.</li> <li><em>Data&nbsp;Labeling&nbsp;&amp;&nbsp;Encoding&nbsp;for&nbsp;RQ3.mx12</em>&nbsp;is&nbsp;the&nbsp;results&nbsp;of&nbsp;data&nbsp;labeling&nbsp;and&nbsp;encoding&nbsp;for&nbsp;RQ3,&nbsp;which&nbsp;were&nbsp;analyzed&nbsp;by&nbsp;the&nbsp;MAXQDA&nbsp;tool.</li> </ul>

opencc-by-4.0May 2020View details →
zenodo28/100

EEG datasets for healthcare: a scoping review - Data and Code

<p>This repository contains the data extracted for the scoping review "EEG datasets for healthcare: a scoping review" and the code used in the analysis.</p><p>&nbsp;</p>

restrictedcc-by-4.0Nov 2023View details →
zenodo28/100

Data and code for peer review - MEE-24-11-799

Open the record for dataset details and reuse information.

opencc-by-4.0Dec 2024View details →
zenodo28/100

On the acceptance by code reviewers of candidate security patches suggested by Automated Program Repair tools - Dataset

<p>Dataset of the empirical experiment presented in the paper&nbsp;On the acceptance by code reviewers of candidate security patches suggested by Automated Program Repair tools. The dataset includes the participants&#39; responses regarding their background, and the responses of the tasks from the experiment.&nbsp;</p>

openJun 2023View details →
zenodo28/100

Effective Teaching through Code Reviews: Patterns and Anti-Patterns

Open the record for dataset details and reuse information.

opencc-by-4.0May 2024View details →
zenodo28/100

Accessibility of Low-Code Approaches: a Systematic Literature Review

Open the record for dataset details and reuse information.

opencc-by-4.0Jul 2024View details →
zenodo28/100

Table-Matrix, containing all the codes generated from the inductive analysis process of the 78 articles that made up the final sample. (Not only Opportunity, but also Uncertainty: A systematic review of how entrepreneurship literature appropriates both constructs.)

Open the record for dataset details and reuse information.

opencc-by-4.0Jul 2024View details →
zenodo28/100

Sistematic review on the code smell effect (2000-2017)

<p>Dataset of the paper &quot;A systematic review on the code smell effect&quot;.</p>

opencc-by-4.0Mar 2018View details →
zenodo28/100

Challenges in analysis of code review comments. Various BERTopic parameters and impact on coherence.

Open the record for dataset details and reuse information.

opencc-by-4.0Aug 2024View details →
zenodo28/100

Advancing Automated Code Review Comment Generation Using Large Language Models

Open the record for dataset details and reuse information.

opencc-by-4.0Nov 2024View details →
zenodo28/100

Towards Automating Code Review Activities - JSS Happy Hour Video

<p>Towards Automating Code Review Activities - JSS Happy Hour Video</p>

opencc-by-4.0Oct 2021View details →
zenodo28/100

Data and Material for the master's thesis: Cost of code review goals and code review strategies

<p>Data and Material for the master&#39;s thesis: Cost of review goals and review strategies</p>

opencc-by-4.0Nov 2022View details →
zenodo28/100

Fire code review sketch dataset

<p>Automatic assessments of building plans are uncommon in the early design stages, especially when schematic sketches are in raster format. Existing design evaluation tools, such as fire code reviewers, which are typically used in the late design stage, primarily evaluate vector format images that contain complete building information. These tools use conditional shape-embedding techniques to analyze the vector images. However, there are limitations to identifying and evaluating drawings through vector-shape relationships. Our research aimed to develop tools that can automatically assess schematic sketches in raster format to overcome the limitations of existing tools. We integrated a conditional shape-embedding tool, named Shape Machine, to assess vector images, with machine learning techniques, namely a Generative Adversarial Network (GAN), to assess raster sketches. This integration enables the evaluation of fire evacuation sketches in the early stages of the design process, thereby improving design efficiency and reducing costs.&nbsp;Moreover, in the future, this integration could allow the evaluation of designs in multiple image formats.</p>

opencc-by-4.0Mar 2023View details →
zenodo28/100

Help Me to Understand this Commit! - A Vision for Contextualized Code Reviews

<p>Literature review results: snowball sampling results and categorization of papers.</p>

opencc-by-4.0May 2023View details →
zenodo28/100

Integrating Visual Aids to Enhance the Code Reviewer Selection Process Replication Package

<p>Modern Code Review (MCR) is an integral part of a software development strategy that accelerates product quality by identifying defects, code smells, and other harmful practices. However, assigning appropriate reviewers to evaluate changed code during the review process remains challenging. While automated tools for reviewer assignments have limited impact in practice, the process often relies on manual investigation of project histories to retrieve knowledge of team members and their activities. Therefore, in this study, we present an approach to automatically assemble developers&#39; information and visualize it meaningfully, which helps to choose appropriate reviewers. First, we propose three metrics that measure developers&#39; collaboration, reviewers&#39; expertise, and reviewers&#39; workload and visualize them through networks. Second, we perform a case study of three popular open-source projects, where we compute and visualize each developers&#39; information according to the proposed metrics. Finally, we conducted two online surveys to assess the developers&#39; perceptions of the proposed visual benefits. The results show that the proposed method can assist in identifying relevant reviewers and be immensely helpful to new developers. Additionally, survey respondents expressed reliance on the efficacy of the visual aids in workload balancing and reducing review time.&nbsp;</p>

opencc-by-4.0Jul 2023View details →
dryad28/100

Code review regression analysis of open source GitHub projects

Open the record for dataset details and reuse information.

publicAug 2017View details →
zenodo24/100

Dataset of "Primers or Reminders? The Effects of Existing Review Comments on Code Review"

<p>Dataset of &quot;Primers or Reminders? The Effects of Existing Review Comments on Code Review&quot;.</p> <p>See README.md for more information.&nbsp;</p>

openother-openJan 2020View details →
zenodo24/100

Dataset of the paper "Information Needs in Contemporary Code Review"

<p>Dataset of the paper &quot;Information Needs in Contemporary Code Review&quot;,&nbsp;Proceedings of the ACM on Human-Computer Interaction&nbsp;2, CSCW, Article 135 (November 2018).</p> <p>Read README.md for information on the content.</p>

openother-openMar 2020View details →
zenodo24/100

Model code and data for "Mitigation of the double ITCZ syndrome in BCC-CSM2-MR through improving parameterizations of boundary-layer turbulence and shallow convection" by Lu et al., submitted to Geoscientific Model Development, https://doi.org/10.5194/gmd-2020-40, in review, 2020.

<p>Description of the files:</p> <p>&ldquo;BCC_CSM2_MR.code.tar&rdquo; contains the codes and run scripts for the medium-resolution Beijing Climate Center Climate System Model version 2 (BCC-CSM2-MR). Detailed description of the model refers to the paper &ldquo;The Beijing Climate Center Climate System Model (BCC-CSM): the main progress from CMIP5 to CMIP6&rdquo; by Wu et al., Geosci. Model Dev., 12, 1573&ndash;1600, https://doi.org/10.5194/gmd-12-1573-2019, 2019.</p> <p>&ldquo;BCC_CSM2_MR.inputdata.tar&rdquo; contains the input data needed to run the model.</p> <p>&ldquo;REF_amip.rar&rdquo; contains the output data from the REF_amip experiment.</p> <p>&ldquo;NEW_amip.rar&rdquo; contains the output data from the NEW_amip experiment.</p> <p>&ldquo;REF_cmip.rar&rdquo; contains the output data from the REF_cmip experiment.</p> <p>&ldquo;NEW_cmip.rar&rdquo; contains the output data from the NEW_cmip experiment.</p> <p>&ldquo;UWMT_amip.rar&rdquo; contains the output data from the UWMT_amip experiment.</p> <p>&ldquo;mHack_amip.rar&rdquo; contains the output data from the mHack_amip experiment.</p>

opencc-by-4.0Jul 2020View details →
zenodo24/100

Identifying prevalent quality issues in code changes by analyzing reviewers' feedback

<p>The dataset contains python scripts (for data extraction, preprocessing) and dataset for the paper "Identifying prevalent quality issues in code changes by analyzing reviewers' feedback" Submitted to ENASE 2024.</p>

opencc-by-4.0Dec 2023View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record