Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
677
datasets available to search
ShareScore release 0.9.0
Dataset results
677 results for “Replication package”
Data and code to replicate: Diet analysis using generalized linear models derived from foraging processes using R package mvtweedie
<p>Diet analysis integrates a wide variety of visual, chemical and biological identification of prey. Samples are often treated as compositional data, where each prey is analyzed as a continuous percentage of the total. However, analyzing compositional data results in analytical challenges, e.g., highly parameterized models or prior transformation of data. Here, we present a novel approximation involving a Tweedie generalized linear model (GLM). We first review how this approximation emerges from considering predator foraging as a thinned and marked point process (with marks representing prey species and individual prey size). This derivation can motivate future theoretical and applied developments. We then provide a practical tutorial for the Tweedie GLM using new package <i>mvtweedie</i> that extends capabilities of widely used packages in R (<i>mgcv</i> and <i>ggplot2</i>) by transforming output to calculate prey compositions. We demonstrate this approach and software using two examples. Tufted puffins (<i>Fratercula cirrhata</i>) provisioning their chicks on a colony in the northern Gulf of Alaska show decadal prey switching among sand lance and prowfish (1980-2000) and then Pacific herring and capelin (2000-2020), while wolves (<i>Canis lupus ligoni</i>) in Southeast Alaska forage on mountain goats and marmots in northern uplands and marine mammals in seaward island coastlines. </p>
Replication package for: Taking the Fed at its Word: A New Approach to Estimating Central Bank Objectives using Text Analysis
<p>This zip file contains all of the programs and data necessary to replicate the results, tables, and figures <br> in the following paper:</p> <p>Shapiro, Adam H., and Daniel J. Wilson (2021). "Taking the Fed at its Word: A New Approach to Estimating Central Bank Objectives using Text Analysis," forthcoming at Review of Economic Studies.</p>
Replication Package for: A Configurable Method for Benchmarking Scalability of Cloud-Native Applications
<p>This repository contains a replication package and experimental results for our study <em>A Configurable Method for Benchmarking Scalability of Cloud-Native Applications</em>.</p> <p>It provides benchmark execution files for repeating our experiments as well as the collected data from our experiments and Jupyter notebooks for reproducing our analysis.</p> <p>Instructions for repeating our experiments and reproducing our analysis can be found in the Readme.md file.</p>
Replication package for "Cupid's Invisible Hand"
<p>Alfred Galichon and Bernard Salanie, "Cupid’s Invisible Hand: Social Surplus and Identification in Matching Models", Review of Economic Studies</p> <p>This replication package contains the data used in the empirical application; the code needed to produce it from the raw data; the code used in estimating the matching models and simulating them; and the code that produces the plots in the paper.</p> <p> </p> <p> </p> <p> </p>
Replication Package (Virtual Machine) for article "starMC: an automata based CTL* model checker"
<p>This package is a virtual machine (VM) to execute the starMC programs and benchmark. The system setup is described in the research article "starMC: an automata based CTL* model checker", being published at PeerJ Computer Science (currently accepted). The virtual machine contains both tools and data to reproduce a CTL* model checking logic benchmark. CTL* is a formal logic for the verification of temporal properties of programs.</p> <p>The benchmark is a collection of 1018 model instances (encoded as Petri nets), derived from the Model Checking Context benchmark (https://mcc.lip6.fr/), and about 60.000 CTL, LTL and CTL* queries. It is currently the only large-scale CTL* benchmark for Petri nets publicly available.</p> <p>The starMC-benchmark.ova VM is configured to use 4 CPU cores and 16 GB of RAM. The VM was created with VirtualBox version 5, and tested on a host machine with an 8-core Xeon CPU and 32 GB of RAM. The VM uses the Ubuntu 20 OS, and username and password are both "user". The home directory contains a README file with instructions on how to run the tools, where the benchmark data (models, queries, variable orders) are found, and the script to reproduce the plots in the paper.<br> The starMC.ova contains the tools presented in the paper, for user evaluation and reproducibility.</p>
Replication Package for the Paper: Transfer Learning with Time Series Data: A Systematic Mapping Study
<p>This is a replication package for the paper "Transfer Learning with Time Series Data: A Systematic Mapping Study".</p> <p>It provides</p> <ul> <li>a documentation of the conducted electronic literature search,</li> <li>exports of the search results from each literature database,</li> <li>and an excel file on the included literature and extracted data.</li> </ul>
Master Thesis Replication Package for: Empirical Scalability Evaluation of Hopping Window Aggregation Methods in Distributed Stream Processing
<p>Master Thesis Replication Package for: Empirical Scalability Evaluation of Hopping Window Aggregation Methods in Distributed Stream Processing</p> <p>A detailed description can be found in the <em>README.md</em>.</p>
Replication package for: Subjective Models of the Macroeconomy: Evidence From Experts and Representative Samples
<p><strong>This repository contains the replication scripts and data of the analyses of the paper: </strong></p> <p>Andre, P., Pizzinelli, C., Roth, C., & Wohlfart, J. (2022). Subjective Models of the Macroeconomy: Evidence From Experts and Representative Samples. <em>Review of Economic Studies</em>.</p>
"Replication package for: {Globalization, Gender, and the Family}"
<p>The package contains all the codes necessary to reproduce the figures and tables in Keller and Utar (forthcoming). "Globalization, Gender, and the Family", Review of Economic Studies. Detailed instructions are also given about accessing the raw data.</p>
FragGen Replication Package
<p>FragGenVM (Tool and instructions )</p> <p>username: fraggen</p> <p>password: fraggenvm</p> <p> </p> <p>FragGenData (experiments for Results reported in the paper)</p> <p> </p> <p>Instructions in VM: /home/fraggen/fraggen/readme.md</p> <p>for RQ1 /home/TSE_RQ1/tse_rq1/readme.md</p> <p>Instructions also available in additional notes</p>
Social Science Theories in Software Engineering Research - Replication Package
<p>Replication package for the article "Social Science Theories in Software Engineering Research".</p> <p>See README file for additional details.</p>
Replication package for Racial Diversity and Racial Policy Preferences: The Great Migration and Civil Rights
<p>The package contains the material to replicate the paper Calderon, A., V. Fouka, and M. Tabellini. " Racial Diversity and Racial Policy Preferences: The Great Migration and Civil Rights," Review of Economic Studies</p>
Replication package for ``Peer Effects inAcademic Research: Senders and Receivers" to appear in EJ
<p>This is the replication package for the paper ``Peer Effects inAcademic Research: Senders and Receivers" to appear in the Economic Journal. The Readme file explains in detail the content and the procedure for replication.</p>
Replication Package for the Paper: "Understanding Code Snippets in Code Reviews: A Preliminary Study of the OpenStack Community"
<p>This is the replication package for the paper: "Understanding Code Snippets in Code Reviews: A Preliminary Study of the OpenStack Community", including dataset and so on (see the description below) : </p> <ul> <li> <p><strong>Data of Code Snippets in Code Review.xlsx</strong> is the dataset of our paper, which contains 10,790 review comments collected from the Nova project and Neutron project of OpenStack community. Among all the review comments, 626 review comments contain code snippets. For the rows of review comments with code snippets, we filled them with blue color as an indicator.</p> </li> <li> <p><strong>Examples for Each Purpose.xlsx</strong> contains six review comment examples for the six detailed purposes mentioned in our paper (see Section 4.2).</p> </li> <li> <p><strong>README.md</strong></p> </li> </ul>
Replication Package for the Paper: "Code Smells Detection via Modern Code Review: A Study of the OpenStack and Qt Communities"
<p>This repository contains the data and results from the paper "Code Smells Detection via Modern Code Review: A Study of the OpenStack and Qt Communities" submitted to the ICPC 2021 special issue of the Empirical Software Engineering Journal, 2021.</p> <p> </p> <p>The replication package contains the following two folders:</p> <p> </p> <p><strong>1) data folder</strong></p> <p>The data folder contains the following four folders, which is organized by research questions (RQs).</p> <ul> <li>RQ1: The RQ1 folder contains the retrieved 1,539 code reviews that discuss code smells. Each review includes four parts: Code Change URL, Code Smell, Code Smell Discussion, and Source Code URL.</li> <li>RQ2: The RQ2 folder contains the coded data for RQ2, called <em>Data Labeling & Encoding for RQ2.mx18</em>. It is the results of data labeling and encoding for RQ2, which was analyzed by the MAXQDA tool.</li> <li>RQ3 and RQ5: <ul> <li><em>Extracted data for RQ3.1.xlsx</em>: this file contains the extracted data (i.e., specific refactoring actions suggested by reviewers) for RQ3.1.</li> <li><em>Data Labeling & Encoding for RQ3 and RQ5.mx18</em>: this file contains the extracted data for RQ3 (excluding the specific refactoring actions in RQ3.1) and RQ5.</li> <li><em>Code change status for RQ5.xlsx</em>: this file contains the information of status of code changes where the developers disagreed with the reviewers and chose to ignore the identified code smells.</li> </ul> </li> <li>RQ4: The RQ4 folder contains the extracted data for RQ4, called <em>Extracted data for RQ4.xlsx</em>.</li> </ul> <p>Note: The mx18 files can be opened by MAXQDA 18 or higher versions, which are available at https://www.maxqda.com/ for download. You may also use the free 14-day trial version of MAXQDA 2018, which is available at https://www.maxqda.com/trial for download.</p> <p> </p> <p><strong>2) scripts folder</strong></p> <p>The scripts folder contains the Python scripts that were used to search for code smell terms and the list of code smell terms.</p> <ul> <li><em>keyword.txt</em> contains the keywords associated with code smells, such as "smell, duplication, and dead".</li> <li><em>get_changes.py</em> is used for getting code changes from OpenStack and Qt.</li> <li><em>get_comments.py</em> is used for getting review comments for each code change.</li> <li><em>keywords_search.py</em> is used for searching review comments that contain at least one keyword.</li> <li><em>random_select.py</em> is used for randomly selecting review comments that do not contain any keyword.</li> <li><em>keywords_improve.py</em> is used for improving the keyword-based mining approach.</li> <li><em>tools.py</em> is used for supporting the process of keywords improving.</li> </ul>
Replication package for: Economic and Social Outsiders but Political Insiders: Sweden's Populist Radical Right
<p>This package contains all the code necessary to replicate the figures and tables in Dal Bo, E., F. Finan, O. Folke, T. Persson, and J. Rickne (forthcoming). "Economic and Social Outsiders but Political Insiders: Sweden's Populist Radical Right", Review of Economic Studies. Detailed instructions are also given about how to access the underlying data. </p>
Replication package for: "Political Selection and Economic Policy"
<p>These files contain the replication package for "Political Selection and Economic Policy", published in the Economic Journal. The files contained within the package allow verifying that the codes used to produce the results reported in the paper are functional, and they also allow a partial replication of the results.</p>
Replication package of SLR of Domain-oriented Specification Techniques
<p>This is a MS Excel sheet containing the data from the initial search up to the classification in the framework, belonging to the manuscript Systematic Literature Review of Domain-oriented Specification techniques, to be published in Journal of Systems and Software. </p>
Replication package of SLR of Domain-oriented Specification Techniques
<p>This is a MS Excel sheet containing the data from the initial search up to the classification in the framework, belonging to the manuscript Systematic Literature Review of Domain-oriented Specification techniques, to be published in Journal of Systems and Software.</p>
Replication package: 40 Years of Designing Code Comprehension Experiments: A Systematic Mapping Study
<p>Replication package | 40 Years of Designing Code Comprehension Experiments: A Systematic Mapping Study</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.