Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

677

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

677 results for “Replication package”

Learn how ShareScore rates datasets ↗
dryad36/100

Data and code to replicate: Diet analysis using generalized linear models derived from foraging processes using R package mvtweedie

<p>Diet analysis integrates a wide variety of visual, chemical and biological identification of prey.  Samples are often treated as compositional data, where each prey is analyzed as a continuous percentage of the total.  However, analyzing compositional data results in analytical challenges, e.g., highly parameterized models or prior transformation of data.  Here, we present a novel approximation involving a Tweedie generalized linear model (GLM).  We first review how this approximation emerges from considering predator foraging as a thinned and marked point process (with marks representing prey species and individual prey size).  This derivation can motivate future theoretical and applied developments.  We then provide a practical tutorial for the Tweedie GLM using new package <i>mvtweedie</i> that extends capabilities of widely used packages in R (<i>mgcv</i> and <i>ggplot2</i>) by transforming output to calculate prey compositions.  We demonstrate this approach and software using two examples. Tufted puffins (<i>Fratercula cirrhata</i>) provisioning their chicks on a colony in the northern Gulf of Alaska show decadal prey switching among sand lance and prowfish (1980-2000) and then Pacific herring and capelin (2000-2020), while wolves (<i>Canis lupus ligoni</i>) in Southeast Alaska forage on mountain goats and marmots in northern uplands and marine mammals in seaward island coastlines. </p>

opencc-zeroNov 2021View details →
zenodo36/100

Replication package for: Taking the Fed at its Word: A New Approach to Estimating Central Bank Objectives using Text Analysis

<p>This zip file contains&nbsp;all of the programs and data necessary to replicate the results, tables, and figures&nbsp;<br> in the following paper:</p> <p>Shapiro, Adam H., and Daniel J. Wilson (2021). &quot;Taking the Fed at its Word: A New Approach to Estimating Central Bank Objectives using Text Analysis,&quot; forthcoming at Review of Economic Studies.</p>

opencc-by-4.0Sep 2021View details →
zenodo36/100

Replication Package for: A Configurable Method for Benchmarking Scalability of Cloud-Native Applications

<p>This repository contains a replication package and experimental results for our study <em>A Configurable Method for Benchmarking Scalability of Cloud-Native Applications</em>.</p> <p>It provides benchmark execution files for repeating our experiments as well as the collected data from our experiments and&nbsp;Jupyter notebooks for reproducing our analysis.</p> <p>Instructions for repeating our experiments and reproducing our analysis can be found in the Readme.md file.</p>

opencc-by-4.0Oct 2021View details →
zenodo36/100

Replication package for "Cupid's Invisible Hand"

<p>Alfred Galichon and Bernard Salanie, &quot;Cupid&rsquo;s Invisible Hand:&nbsp;Social Surplus and Identification in Matching Models&quot;, Review of Economic Studies</p> <p>This replication package contains&nbsp;the data used in the empirical application;&nbsp;&nbsp;the code needed to produce it from the raw data;&nbsp;&nbsp;the code used in estimating the matching models and simulating them; and&nbsp;&nbsp;the code that produces the plots in the paper.</p> <p>&nbsp;</p> <p>&nbsp;</p> <p>&nbsp;</p>

opencc-by-4.0Jul 2021View details →
zenodo36/100

Replication Package (Virtual Machine) for article "starMC: an automata based CTL* model checker"

<p>This package is a virtual machine (VM) to execute the starMC programs and benchmark. The system setup is described in the research article &quot;starMC: an automata based CTL* model checker&quot;, being published at&nbsp;PeerJ Computer Science (currently accepted). The virtual machine contains both tools and data to reproduce a CTL* model checking logic benchmark. CTL* is a formal logic for the verification of temporal properties of programs.</p> <p>The benchmark is a collection of 1018 model instances (encoded as Petri nets), derived from the Model Checking Context benchmark (https://mcc.lip6.fr/), and about 60.000 CTL, LTL and CTL* queries. It is currently the only large-scale CTL* benchmark for Petri nets publicly available.</p> <p>The starMC-benchmark.ova VM is configured to use 4&nbsp;CPU cores and 16&nbsp;GB of RAM. The VM was created with VirtualBox version 5,&nbsp;&nbsp;and tested on a host machine with an 8-core&nbsp;Xeon&nbsp;CPU and 32 GB of RAM. The VM uses the Ubuntu 20 OS, and username and password are both &quot;user&quot;.&nbsp;The home directory contains a README file with instructions on how to run the tools, where the benchmark data (models, queries, variable orders) are found, and the script to reproduce the plots in the paper.<br> The starMC.ova contains the tools presented in the paper, for user evaluation and reproducibility.</p>

opencc-by-4.0Dec 2021View details →
zenodo36/100

Replication Package for the Paper: Transfer Learning with Time Series Data: A Systematic Mapping Study

<p>This is a replication package for the paper "Transfer Learning with Time Series Data: A Systematic Mapping Study".</p> <p>It provides</p> <ul> <li>a documentation of the conducted electronic literature search,</li> <li>exports of the search results from each literature database,</li> <li>and an excel file on the included literature and extracted data.</li> </ul>

opencc-by-4.0Dec 2021View details →
zenodo36/100

Master Thesis Replication Package for: Empirical Scalability Evaluation of Hopping Window Aggregation Methods in Distributed Stream Processing

<p>Master Thesis Replication Package for: Empirical Scalability Evaluation of Hopping Window Aggregation Methods in Distributed Stream Processing</p> <p>A detailed description can be found in the&nbsp;<em>README.md</em>.</p>

opencc-by-4.0Dec 2021View details →
zenodo36/100

Replication package for: Subjective Models of the Macroeconomy: Evidence From Experts and Representative Samples

<p><strong>This repository contains the replication scripts and data of the analyses of the paper: </strong></p> <p>Andre, P., Pizzinelli, C., Roth, C., &amp; Wohlfart, J. (2022). Subjective Models of the Macroeconomy: Evidence From Experts and Representative Samples. <em>Review of Economic Studies</em>.</p>

opencc-by-4.0Nov 2021View details →
zenodo36/100

"Replication package for: {Globalization, Gender, and the Family}"

<p>The package contains all the codes necessary to reproduce the figures and tables in Keller and Utar (forthcoming). &quot;Globalization, Gender, and the Family&quot;, Review of Economic Studies. Detailed instructions are also given about accessing the raw data.</p>

opencc-by-4.0Dec 2021View details →
zenodo36/100

FragGen Replication Package

<p>FragGenVM (Tool and instructions )</p> <p>username: fraggen</p> <p>password: fraggenvm</p> <p>&nbsp;</p> <p>FragGenData (experiments&nbsp;for Results reported in the paper)</p> <p>&nbsp;</p> <p>Instructions in VM: /home/fraggen/fraggen/readme.md</p> <p>for RQ1 /home/TSE_RQ1/tse_rq1/readme.md</p> <p>Instructions also available in additional notes</p>

opencc-by-4.0Feb 2022View details →
zenodo36/100

Social Science Theories in Software Engineering Research - Replication Package

<p>Replication package for the article &quot;Social Science Theories in Software Engineering Research&quot;.</p> <p>See README file for additional details.</p>

opencc-by-4.0Feb 2022View details →
zenodo36/100

Replication package for Racial Diversity and Racial Policy Preferences: The Great Migration and Civil Rights

<p>The package&nbsp;contains the material to replicate the paper Calderon, A., V. Fouka, and M. Tabellini. &quot;&nbsp;Racial Diversity and Racial Policy Preferences: The Great Migration and Civil Rights,&quot; Review of Economic Studies</p>

opencc-by-4.0Feb 2022View details →
zenodo36/100

Replication package for ``Peer Effects inAcademic Research: Senders and Receivers" to appear in EJ

<p>This is the replication package for the paper ``Peer Effects inAcademic Research: Senders and Receivers&quot; to appear in the Economic Journal. The Readme file explains in detail the content and the procedure for replication.</p>

opencc-by-4.0Mar 2022View details →
zenodo36/100

Replication Package for the Paper: "Understanding Code Snippets in Code Reviews: A Preliminary Study of the OpenStack Community"

<p>This is the replication package for the paper: &quot;Understanding Code Snippets in Code Reviews: A Preliminary Study of the OpenStack Community&quot;, including dataset and so on (see the description below) :&nbsp;</p> <ul> <li> <p><strong>Data of Code Snippets in Code Review.xlsx</strong> is the dataset of our paper, which contains 10,790 review comments collected from the Nova project and Neutron project of OpenStack community. Among all the review comments, 626 review comments contain code snippets. For the rows of review comments with code snippets, we filled them with blue color as an indicator.</p> </li> <li> <p><strong>Examples for Each Purpose.xlsx</strong> contains six review comment examples for the six detailed purposes mentioned in our paper (see Section 4.2).</p> </li> <li> <p><strong>README.md</strong></p> </li> </ul>

opencc-by-4.0Mar 2022View details →
zenodo36/100

Replication Package for the Paper: "Code Smells Detection via Modern Code Review: A Study of the OpenStack and Qt Communities"

<p>This repository contains the data and results from the paper &quot;Code Smells Detection via Modern Code Review: A Study of the OpenStack and Qt Communities&quot; submitted to the ICPC 2021 special issue of the Empirical Software Engineering Journal, 2021.</p> <p>&nbsp;</p> <p>The replication package contains the following two folders:</p> <p>&nbsp;</p> <p><strong>1) data folder</strong></p> <p>The data folder contains the following four folders, which is organized by research questions (RQs).</p> <ul> <li>RQ1:&nbsp;The RQ1 folder contains the retrieved 1,539 code reviews that discuss code smells. Each review includes four parts: Code Change URL, Code Smell, Code Smell Discussion, and Source Code URL.</li> <li>RQ2: The RQ2 folder contains the coded data for RQ2, called <em>Data Labeling &amp; Encoding for RQ2.mx18</em>. It is the results of data labeling and encoding for RQ2, which was analyzed by the MAXQDA tool.</li> <li>RQ3 and RQ5: <ul> <li><em>Extracted data for RQ3.1.xlsx</em>: this file contains the extracted data (i.e., specific refactoring actions suggested by reviewers) &nbsp;for RQ3.1.</li> <li><em>Data Labeling &amp; Encoding for RQ3 and RQ5.mx18</em>: this file contains the extracted data for RQ3 (excluding the specific refactoring actions in RQ3.1) and RQ5.</li> <li><em>Code&nbsp;change&nbsp;status&nbsp;for&nbsp;RQ5.xlsx</em>: this&nbsp;file&nbsp;contains&nbsp;the&nbsp;information&nbsp;of&nbsp;status&nbsp;of&nbsp;code&nbsp;changes&nbsp;where&nbsp;the&nbsp;developers&nbsp;disagreed with&nbsp;the&nbsp;reviewers&nbsp;and&nbsp;chose&nbsp;to&nbsp;ignore&nbsp;the&nbsp;identified&nbsp;code&nbsp;smells.</li> </ul> </li> <li>RQ4:&nbsp;The RQ4 folder contains the extracted data for RQ4, called <em>Extracted data for RQ4.xlsx</em>.</li> </ul> <p>Note:&nbsp;The&nbsp;mx18&nbsp;files&nbsp;can&nbsp;be&nbsp;opened&nbsp;by&nbsp;MAXQDA&nbsp;18 or&nbsp;higher&nbsp;versions,&nbsp;which&nbsp;are&nbsp;available&nbsp;at&nbsp;https://www.maxqda.com/&nbsp;for&nbsp;download.&nbsp;You&nbsp;may&nbsp;also&nbsp;use&nbsp;the&nbsp;free&nbsp;14-day&nbsp;trial&nbsp;version&nbsp;of&nbsp;MAXQDA&nbsp;2018,&nbsp;which&nbsp;is&nbsp;available&nbsp;at&nbsp;https://www.maxqda.com/trial&nbsp;for&nbsp;download.</p> <p>&nbsp;</p> <p><strong>2) scripts folder</strong></p> <p>The scripts folder contains the Python scripts that were used to search for code smell terms and the list of code smell terms.</p> <ul> <li><em>keyword.txt</em>&nbsp;contains the keywords associated with code smells, such as &quot;smell, duplication, and dead&quot;.</li> <li><em>get_changes.py</em>&nbsp;is used for getting code changes from OpenStack and Qt.</li> <li><em>get_comments.py</em>&nbsp;is used for getting review comments for each code change.</li> <li><em>keywords_search.py</em>&nbsp;is used for searching review comments that contain at least one keyword.</li> <li><em>random_select.py</em>&nbsp;is used for randomly selecting review comments that do not contain any keyword.</li> <li><em>keywords_improve.py</em>&nbsp;is used for improving the keyword-based mining approach.</li> <li><em>tools.py</em>&nbsp;is used for supporting the process of keywords improving.</li> </ul>

opencc-by-4.0Mar 2022View details →
zenodo36/100

Replication package for: Economic and Social Outsiders but Political Insiders: Sweden's Populist Radical Right

<p>This&nbsp;package contains all the code necessary to replicate the figures and tables in Dal Bo, E., F. Finan, O. Folke, T. Persson, and J. Rickne (forthcoming). &quot;Economic and Social Outsiders but Political Insiders: Sweden&#39;s Populist Radical Right&quot;, Review of Economic Studies. Detailed instructions are also given about how to access the underlying data.&nbsp;</p>

opencc-by-4.0May 2022View details →
zenodo36/100

Replication package for: "Political Selection and Economic Policy"

<p>These files contain the replication package for &quot;Political Selection and Economic Policy&quot;, published in the Economic Journal. The files contained within the package allow verifying that the codes used to produce the results reported in the paper are functional, and they also allow a partial replication of the results.</p>

opencc-by-4.0May 2022View details →
zenodo36/100

Replication package of SLR of Domain-oriented Specification Techniques

<p>This is a MS&nbsp;Excel sheet containing&nbsp;the data from the initial search up to the classification in the framework, belonging to the manuscript Systematic Literature Review of Domain-oriented Specification techniques, to be published in Journal of Systems and Software.&nbsp;</p>

opencc-by-4.0Jun 2022View details →
zenodo36/100

Replication package of SLR of Domain-oriented Specification Techniques

<p>This is a MS&nbsp;Excel sheet containing&nbsp;the data from the initial search up to the classification in the framework, belonging to the manuscript Systematic Literature Review of Domain-oriented Specification techniques, to be published in Journal of Systems and Software.</p>

opencc-by-4.0Jun 2022View details →
zenodo36/100

Replication package: 40 Years of Designing Code Comprehension Experiments: A Systematic Mapping Study

<p>Replication package | 40 Years of Designing Code Comprehension Experiments: A Systematic Mapping Study</p>

openother-openJun 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record