Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
677
datasets available to search
ShareScore release 0.9.0
Dataset results
677 results for “Replication package”
Replication package for: Empirical Investigation of a Sufficient Statistic for Monetary Shocks
<p>This replication package codes and datasets to replicate the results in "Empirical Investigation of a Sufficient Statistic for Monetary Shocks" by Fernando Alvarez, Andrea Ferrara, Erwan Gautier, Hervé Le Bihan and Francesco Lippi when to be published at RESTUD. The readme file describes how to run the code and how to access the data underlying the results. Please note that the micro CPI and PPI data sets used to compute the price rigidity statistics are proprietary data, provided to the authors by Insee (French Statistical Office) after a bilateral agreement between Banque de France and Insee and these micro price data sets cannot be shared.</p>
Replication package for: Economic Integration and the Transmission of Democracy
<p>The package contains replication files for results of "Economic Integration and the Transmission of Democracy" by Marco Tabellini and Gicomo Magistretti to be published in the Review of Economic Studies. The README file also discusses how to obtain access to raw data and use the replication files.</p>
Replication package for: "Trading stocks builds financial confidence and compresses the gender gap"
<p>This package contains replication files for results using experimental data from "Trading stocks builds financial confidence and compresses the gender gap" by Saumitra Jha and Moses Shayo, to be published in <em>The Economic Journal</em>.</p>
Replication Package for: "Redundancy-free Analysis of Multi-revision Software Artifacts"
<p>Researchers often analyze several revisions of a software project to obtain historical data about its evolution. For example, they statically analyze the source code and monitor the evolution of certain metrics over multiple revisions. The time and resource requirements for running these analyses often make it necessary to limit the number of analyzed revisions, e.g. by only selecting major revisions or by using a coarse-grained sampling strategy, which could remove significant details of the evolution. Most existing analysis techniques are not designed for the analysis of multi-revision artifacts and they treat each revision individually. However, the actual difference between two subsequent revisions is typically very small. Thus, tools tailored for the analysis of multiple revisions should only analyze these differences, thereby preventing re-computation and storage of redundant data, improving scalability and enabling the study of a larger number of revisions. In this work, we propose the Lean Language-Independent Software Analyzer (LISA), a generic framework for representing and analyzing multi-revisioned software artifacts. It employs a redundancy-free, multi-revision representation for artifacts and avoids re-computation by only analyzing changed artifact fragments across thousands of revisions. The evaluation of our approach consists of measuring the effect of each individual technique incorporated, an in-depth study of LISA's resource requirements and a large-scale analysis over 7 million program revisions of 4,000 software projects written in four languages. We show that the time and space requirements for multi-revision analyses can be reduced by multiple orders of magnitude, when compared to traditional, sequential approaches.</p>
Replication Package for the Paper: "The Impact of Code Review on Architectural Changes"
<p>This is a replication package for the paper "The Impact of Code Review on Architectural Changes", to be published at the IEEE Transactions on Software Engineering (TSE).</p> <p> </p> <p>The TSE paper mentioned above is a journal extension of a previous conference paper published at the IEEE/ACM International Conference on Automated Software Engineering (ASE'17). Hence, this is also the replication package for our conference paper "Are Developers Aware of the Architectural Impact of Their Changes?"</p> <p> </p> <p>The replication package is composed of:</p> <p> </p> <p>1 - 103,778 structural architectures extracted from the source code of 7 open source software systems.</p> <p>2 - Churn and size metrics for each of the extracted structural architectures.</p> <p>3 - Values of structural cohesion and coupling computed for the significant architectural changes</p> <p>4 - A manual classification of the changes' intent</p> <p> </p> <p>The presented dataset can be used not only to replicate the TSE paper but also to perform new empirical studies involving software architecture, architectural changes, software changes, code review and software module clustering, to mention a few.</p> <p> </p>
Replication Package for the Paper: "Rebasing in Code Review Considered Harmful"
<p>This is the replication package for the paper: "Rebasing in Code Review Considered Harmful: A Large-scale Empirical Investigation", published at the 19th IEEE International Working Conference on Source Code Analysis and Manipulation (SCAM'19).</p> <p> </p> <p>The replication package is composed of the raw results from each of our 4 research questions. In addition, we include the results from the small empirical study we performed to evaluate the methodology we proposed to handle rebasing operations in code review data.</p> <p> </p> <p>Each of the reported results has been obtained by following the procedures and heuristics described in the paper.</p> <p> </p> <p>This publication uses the CROP dataset of code review data: <a href="https://crop-repo.github.io/">https://crop-repo.github.io/</a></p>
Replication Package: Microservice-tailored Generation of Session-based Workload Models for Representative Load Testing
<p>This is the replication package for the publication <em>Microservice-tailored Generation of Session-based Workload Models for Representative Load Testing</em>, MASCOTS 2019. It holds the experiment setup, results, and detailed analyses of the results.</p> <p>The README.md (or README.pdf) contains further descriptions and instructions.</p>
Replication package for: Moment Conditions for Dynamic Panel Logit Models with Fixed Effects
<p>These are the replication files for <span>MS# 28927-3 “Moment Conditions for Dynamic Panel Logit Models with Fixed Effects” </span></p>
Replication package for: What Drives Demand for State-Run Lotteries? Evidence and Welfare Implications
<p>Replication package:</p> <p>Lockwood, Benjamin B., Hunt Allcott, Dmitry Taubinsky, and Afras Y. Sial. 2024. "What Drives Demand for State-Run Lotteries? Evidence and Welfare Implications." Review of Economic Studies. </p> <p>This package contains all replication scripts and non-confidential data sets necessary to reproduce results in the paper. </p>
Replication Package of ``The Impacts of Managerial Autonomy on Firm Outcomes"
<p>The paper “The Impacts of Managerial Autonomy on Firm Outcomes” uses a combination of newly collected data, and proprietary data. All of the former data used in the paper have been included here. The latter is the Center for Monitoring the Indian Economy’s (CMIE) Prowess dataset (CMIE, 2021), which is commercially available. This replication package includes raw and cleaned data that I digitized for the project, as well as <span>simulated</span> data to replace the Prowess data, to be able to replicate the code. I have included the results from this simulated dataset in this document, to allow for comparison.</p> <p>Kala, Namrata. <em>The impacts of managerial autonomy on firm outcomes</em>. No. w26304. National Bureau of Economic Research, 2019.</p>
Replication Package: Anomaly detection via runtime monitoring data for structural equation modeling
<p>This replication package contains the following information:</p> <ul> <li><strong>Data extraction from literature & interviews: </strong><em>Generation Structural & Measurement Model via literature and interviews.xlsx</em> - here you can find the mapping of the extracted phrases to inductively summarise information regarding the structural and measurement models.</li> <li><strong>Dataset</strong> of runtime monitoring data extracted from TrainTicket via EvoMaster: <br> <ul> <li><em>TrainTicket faults classification.xlxs:</em> Describes the datasets and their faults, in which microservice the fault is injected for better explainability of the obtained results</li> <li><em>IndicatorDescriptionbasedonAnomalyDetectionToolsInterviews.xlsx:</em> description and mapping of selected indicators to the defined parameters from <a href="https://arxiv.org/abs/2408.07816" target="_blank" rel="noopener">previous work </a></li> <li>Unfortunately, the size of the datasets generated via EvoMaster and their injected faults are too big to upload here, thus, they will be available here: <a href="https://uibkacat-my.sharepoint.com/:f:/g/personal/monika_steidl_uibk_ac_at/EjLMt8SYWwtJtp2YuSaqavcBKJoCQ3b5H_l_OY0ifbVRCA?e=fatKyD" target="_blank" rel="noopener">Datasets with injected anomalies</a><br> <ul> <li>the error description can be found <a href="https://github.com/FudanSELab/train-ticket/wiki/Fault-Description" target="_blank" rel="noopener">here</a></li> <li>the datasets are named ts-error-<em>indicatorOfError</em>-reset.zip because the databases are getting reset so that no anomalies are introduced with wrong database entries</li> </ul> </li> </ul> </li> <li><strong>Code</strong> for handling and transforming data to extract indicators describing the whole system's and microservices' behavior from the collected runtime monitoring data collected from TrainTicket:<br> <ul> <li><a href="https://github.com/moniSt13/ConTest-Parsing" target="_blank" rel="noopener">link to the Github repository</a></li> </ul> </li> <li><strong>reports</strong> regarding the established PLS-SEM model using previously handled and transformed runtime monitoring data. Please be aware that opening the reports can leas to out of memory due to their size: <ul> <li><em>Assessment of Measurement Model: MeasurementModel_TrainTicket_erorcleaned.zip & MeasurementModel_Bootstrap_ALL_TrainTicket_errorcleaned.zip</em> </li> <li><em>Assessment of Structural Model: StructuralModel_TrainTicket_errorcleaned.zip & StructuralModel_Bootstrap_ALL_TrainTicket_errorcleaned.zip</em></li> </ul> </li> </ul> <p><br><br>---------------------------------------</p> <p><em>Future work </em>not elaborated in the associated paper due to space restrictions:</p> <ul> <li><strong>reports regarding F5 error</strong>: PLS-SEM model results without interpretation and further mediating effects between microservices included: F5_error.zip</li> </ul> <p> </p> <p> </p> <p> </p>
Replication package for: Looming Large or Seeming Small? Attitudes Towards Losses in a Representative Sample
<p>This package contains data and replication files to reproduce the results from "Looming Large or Seeming Small? Attitudes Towards Losses in a Representative Sample" by Chapman, Snowberg, Wang, and Camerer, to be published in the Review of Economic Studies. This includes 6 original survey datasets (3 general population studies, two student studies, and an expert survey), matlab code to reproduce DOSE parameter estimates, and STATA code to produce the exhibits in the paper.<br><br></p>
Replication package for Journal paper title "Supporting the identification of prevalent quality issues in code changes by analyzing reviewers' feedback"
<p>Replication package for the study "Supporting the identification of prevalent quality issues in code changes by analyzing reviewers' feedback"</p> <p> </p>
Replication package for the paper: "Code Clone Configuration as a Multi-Objective Search Problem"
<p><strong>Replication Package Context</strong></p> <p>This is the replication package for the paper "Code Clone Configuration as a Multi-Objective Search Problem". The paper was originally published in the <em>International Symposium on Empirical Software Engineering and Measurement</em> (ESEM).</p> <div> <div><strong>Search Algorithms and Datasets for MC3 Problem</strong></div> <div>This repository contains implementations of search algorithms for solving the MC3 (Multi-objective Code Clone Configuration) problem. </div> <div>Additionally, it includes various files related to Elasticsearch configurations, I/O operations, and other supporting resources. For more information about the project files and folders, you can read the README.md for general information.</div> </div>
Replication package for: "Strategic conformity or anticonformity to avoid punishment and attract reward".
<p>Replication package for: Dvorak, F., Fischbacher, U., and Schmelz, K. (2024). "Strategic conformity or anticonformity to avoid punishment and attract reward", The Economic Journal.</p>
Understanding the Implications of Changes to Build Systems (Replication Package)
<p>Online appendix for <em>"Understanding the Implications of Changes to Build</em> Systems", in the Proceedings of the International Conference on Software Engineering (ICSE), 2023. </p>
Replication Package for the Paper: Identifying Key Factors for Using Ethnography in Software Requirements Elicitation - A Systematic Literature Review
<p>This is the replication package for the paper: "Identifying Key Factors for Using Ethnography in Software Requirements Elicitation - A Systematic Literature Review"</p> <p>It contains:</p> <ul> <li>Review protocol (in spanish)</li> <li>Extracted data for each research question</li> <li>List of primary studies</li> <li>Appendix</li> </ul> <p> </p>
Replication Package for: "How do Workers Adjust to Robots? Evidence from China"
<p>Replication Package for:</p> <p> "How do Workers Adjust to Robots? Evidence from China"</p> <p>Osea Giuntella, Yi Lu, Tianyi Wang</p> <p>10.5281/zenodo.13760908</p>
Migration of Monolithic Systems to Microservices: A Systematic Mapping Study - Replication package
Open the record for dataset details and reuse information.
Replication package tied with the paper "WasteLess: An Optimal Provisioner for Self-Adaptive Second-Generation Serverless Applications"
<p>This is the replication package tied with paper "WasteLess: An Optimal Provisioner for Self-Adaptive Second-Generation Serverless Applications" submitted to the 20th International Conference on Software Engineering for Adaptive and Self-Managing Systems (SEAMS'25)</p> <p><strong>Contents<br></strong>The package contains a single zip file (WlessPackage.zip) with all the material used for conducting the paper experimentsation. Refer to the README.md file contained in the zip package for a detailed description of its content.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.