Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
72
datasets available to search
ShareScore release 0.9.0
Dataset results
72 results for “Stack Overflow”
Replication package for the paper: "Using Stack Overflow to Assess Technical Debt Identification on Software Projects".
<p>This is the replication package for the paper "Using Stack Overflow to Assess Technical Debt Identification on Software Projects". The paper was published on the Technical Research Track of the Bralizian Symposium on Software Engineering. </p> <p> </p> <p>The replication package is composed of 3 files: 1) data.zip, 2) tables.zip, and 3) clarifications.zip</p> <p> </p> <p>In data.zip, we provide the outputs of our methodological procedure to select the discussions from StackOverflow to include in our empirical analysis. </p> <p>In tables.zip, we provide additional and detailed data regarding all tables included in the paper/</p> <p>In clarifications.zip, we provide additional clarification regarding concepts we did not fully discuss in the manuscript. </p> <p> </p> <p>For future references in this paper, please contact the main author Eliakim Gama, or one of the co-authors.</p>
INVESTIGANDO A RELAÇÃO ENTRE INDICADORES DE DÍVIDA TÉCNICA E CARACTERÍSTICAS DE QUALIDADE DE SOFTWARE EM DISCUSSÕES DO STACK OVERFLOW
<p>This is the replication package of the dissertation " INVESTIGANDO A RELAÇÃO ENTRE INDICADORES DE DÍVIDA TÉCNICA E CARACTERÍSTICAS DE QUALIDADE DE SOFTWARE EM DISCUSSÕES DO STACK OVERFLOW". Dissertation presented to the Academic Master's Course in Computer Science of the Graduate Program in Computer Science at the Science and Technology Center of the State University of Ceará, as a partial requirement for obtaining a Master's degree in Computer Science. Area of Concentration: Computer Science</p>
Does Location Influence Coding Practices? A Cross-Regional Study on Stack Overflow Code Quality – Replication Package
<p>Developers routinely integrate Stack Overflow code snippets into their codebases. However, the quality of snippets embedded in users’ answers remain elusive, and existing evaluations of code quality tend to be language or context-specific. Moreover, literature have found that contribution patterns vary depending on geographical locales, creating an unexplained rift between code quality, user location, and latent contextual regional factors. </p> <p>The proposed study evaluates the quality of SQL, JavaScript, Python, Ruby, and Java snippets across reliability, readability, performance, and security dimensions, benchmarking findings across states in the USA and investigating how different diversity indicators correlate against code quality violations. The study culminates in a series of inductive content analyses that qualitatively supplement prior quality dimensions.</p> <p>This replication package is provided for those interested in further examining our research methodology.</p>
Using Stack Overflow to Assess Technical Debt Identification on Software Projects SBES2020 - Eliakim Gama
<p>Vídeo para backup da apresentação SBES 2020.</p>
Retrieving API knowledge from Tutorials and Stack Overflow based on Natural Language Queries
<p>The replication package of PLAN</p>
Retrieving API knowledge from Tutorials and Stack Overflow based on Natural Language Queries
<p>The replication package of PLAN</p>
Dat4API.ABSA: A Dataset of API Reviews from Stack Overflow for Aspect-Based Sentiment Analysis
<p>This is the dataset created from Stack Overflow discussions with manually labeled<br> Aspect-API-Sentiment information, used in the paper 'Dat4API.ABSA: A Dataset of API Reviews from Stack Overflow for Aspect-Based Sentiment Analysis'.</p>
Knowledge Curation in a Developer Community: A Study of Stack Overflow and Mailing Lists
<p>This file was closed as a result of some issues with the data. Please visit the following link to have access to the updated version of the information. </p> <p>https://zenodo.org/record/47484</p> <p>This PostgreSQL dump file contains the data from Stack Overflow r-tag and the R-help mailing list. The data is framed between 2008 and 2013. It is part of "Knowledge Curation in a Developer Community: A Study of Stack Overflow and Mailing Lists" by Carlos Gómez Teshima (2015 MSc. Thesis), and "How the R Community Creates and Curates Knowledge" by Alexey Zagalsky, Carlos Gómez Teshima, Daniel M. German, Margaret-Anne Storey, Germán Poo-Caamaño (MSR 2016 paper).</p>
Research Artefact: A Large-Scale Study of React-Related Questions in Stack Overflow: Frequency, Popularity, and Types
<p>This is a research artefact for the paper: A Large-Scale Study of React-Related Questions in Stack Overflow: Frequency, Popularity, and Types. This artefact is a repository consisting of the collected dataset including 447,542 React-related SO question posts and 384 representative samples that were obtained randomly. This artefact aims to enable researchers to replicate our dataset of the paper and reuse the dataset for further research.</p>
Research Artefact: A Large-Scale Study of React-Related Questions in Stack Overflow: Frequency, Popularity, and Types
<p>This is a research artefact for the paper: A Large-Scale Study of React-Related Questions in Stack Overflow: Frequency, Popularity, and Types. This artefact is a repository consisting of the collected dataset including 447,542 React-related SO question posts and 384 representative samples that were obtained randomly. This artefact aims to enable researchers to replicate our dataset of the paper and reuse the dataset for further research.</p>
Dataset of AI-generated code and human-written code created by Stack Overflow questions
Open the record for dataset details and reuse information.
Supplemental materials for studying GPU programming with Stack Overflow posts
<p>It includes the data used in our study of GPU programming and the complete results obtained from the study.</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.