Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
57
datasets available to search
ShareScore release 0.9.0
Dataset results
57 results for “Program Repair”
The Dataset of ICSE'20 paper titled "On the Efficiency of Test Suite based Program Repair"
<p>The artifact contains the results and some logs after executing 16 open-source automated program repair systems (i.e., jGenProg, jKali, jMutRepair, Cardumen, AJRA, GenProg-A, Kali-A, RSRepair-A, Nopol, DynaMoth, SimFix, kPAR, AVATAR, FixMiner, TBar, and ACS). <strong>Results</strong> (i.e., Patches) consist of the patches generated by executing 16 open-source automated program repair systems on Defects4J bugs with two different fault localization settings.<strong> Logs </strong>(i.e., Execution logs) record the execution logs of 10 open-source automated program repair systems as the remaining 6 automated program repair systems do not provide execution logs.</p>
The repository of the SANER'21 paper titled "On the Impact of Flaky Tests in Automated Program Repair"
<p>This repository contains the artefact of the paper “On the Impact of Flaky Tests in Automated Program Repair” under review by SANER2021.</p> <ul> <li>results: <ul> <li><strong>RQ1:</strong> This table contains all the statistical data of flaky tests we find out from Defects4J benchmark where each line illustrates the commit date, commit id, java file name, as well as all the flaky methods in this test file. Note that the commit id is empty if there is no update about the java file in that line since last version. <ul> <li>flaky_methods.xlsx</li> </ul> </li> <li><strong>RQ2:</strong> Here are the execution results of 10 times running of flaky tests under 4 different environment. Detailed execution log of each running are classified by project and named in the format <strong><em>test_log_[project_name].txt</em></strong>,the summarized data of each project are named in the format <strong><em>Flaky_test_[project_name].xlsx</em></strong>. The <strong><em>Flaky_test_all.xlsx</em></strong> contain the summarized data for all projects. <ul> <li>jdk1.7 + ubuntu16.04</li> <li>jdk1.7 + ubuntu18.04</li> <li>jdk1.8 + ubuntu16.04</li> <li>jdk1.8 + ubuntu18.04</li> </ul> </li> <li><strong>RQ3:</strong> This folder contains execution results of the fault localization tool <a href="https://github.com/GZoltar/gzoltar">GZoltar-V1.7</a> under two jdk versions. Result of each bug is named in the format <strong><em>[project_name]_[bug_id].csv</em></strong>. <ul> <li>gzoltar_reuslt_flaky_jdk1.7</li> <li>gzoltar_reuslt_flaky_jdk1.8</li> <li>gzoltar_reuslt_noFlaky_jdk1.7</li> <li>gzoltar_reuslt_noFlaky_jdk1.8</li> </ul> </li> <li><strong>RQ4:</strong> This folder contains repair results of different APR tools. For the 10 tools from <a href="https://github.com/program-repair/RepairThemAll">RepairThemAll</a> framework, each result is named in format <strong><em>validation_result_[APR_name].txt</em></strong> which indicates whether the previous patch can be generated this time. For <a href="https://github.com/SerVal-DTF/TBar">TBar</a>, we release all the generated patches including <em>fixed</em> and <em>partially fixed</em>. <ul> <li>RepairThemAll</li> <li>TBar</li> </ul> </li> </ul> </li> </ul>
Replication package for our paper "Search-based Automated Program Repair of CPS Controllers Modeled in Simulink-Stateflow"
Open the record for dataset details and reuse information.
On the acceptance by code reviewers of candidate security patches suggested by Automated Program Repair tools - Dataset
<p>Dataset of the empirical experiment presented in the paper On the acceptance by code reviewers of candidate security patches suggested by Automated Program Repair tools. The dataset includes the participants' responses regarding their background, and the responses of the tasks from the experiment. </p>
Automated Program Repair Cardumen Results
Open the record for dataset details and reuse information.
Automated Program Repair Term Project Cardumen Results Part 2
Open the record for dataset details and reuse information.
Automated Program Repair via Conversation: Fixing 162 out of 337 Bugs for $0.42 Each using ChatGPT
Open the record for dataset details and reuse information.
Automated Program Repair in the Era of Large Pre-trained Language Models
<p>Code used for the paper along with the generated outputs</p>
Artifact of the paper: On the Effectiveness of Automated Program Repair: An Extensive Study
<p>The pre-trained Edits model.</p>
Artifact of paper: On the Effectiveness of Automated Program Repair: An Extensive Study
<p>This is the artifact repo for paper: <strong>On the Effectiveness of Automated Program Repair: An Extensive Study</strong>. This repo contains large files, e.g., patches and compilation logs for this study.</p>
Program Repair Guided by Datalog-Defined Program Analysis
<p><em>Anonymous</em></p>
Towards Efficient Program Repair with APR Tools Based on Genetic Algorithms
<p>This dataset contains experimental results and related artifacts.</p>
SOX11 and SOX4 drive the reactivation of an embryonic gene program during wound repair
GEO Series GSE120773. Mus musculus. 5 samples. Type: Genome binding/occupancy profiling by high throughput sequencing.
Cisplatin-DNA Adduct Repair of Transcribed Genes is Controlled by Two Circadian Programs in Mouse Tissues
GEO Series GSE109938. Mus musculus. 24 samples. Type: Other.
In vivo programming of adult pericytes aids axon regeneration by providing cellular bridges for SCI repair
GEO Series GSE293953. Mus musculus. 2 samples. Type: Expression profiling by high throughput sequencing.
Macrophages sense ECM mechanics and growth factor availability through cytoskeletal remodeling to regulate their tissue repair program [ATAC-seq]
GEO Series GSE238262. Mus musculus. 6 samples. Type: Genome binding/occupancy profiling by high throughput sequencing.
Systemic environment in acute myocardial infarction programs human macrophages for trauma repair
GEO Series GSE172270. Homo sapiens. 67 samples. Type: Expression profiling by high throughput sequencing.
Accelerated flexor tendon repair in superhealer mice reflects alterations in TGFB1 regulated expression programs linked to inflammatory, fibrosis and cell cycle control
GEO Series GSE175912. Mus musculus. 16 samples. Type: Expression profiling by high throughput sequencing.
MYC regulates a DNA repair gene expression program in small cell carcinoma of the ovary, hypercalcemic type
GEO Series GSE298397. Homo sapiens. 18 samples. Type: Genome binding/occupancy profiling by high throughput sequencing; Expression profiling by high throughput sequencing.
Less Training, More Repairing Please: Revisiting Automated Program Repair via Zero-shot Learning
<p>Code used for the paper along with the generated outputs</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.