Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

52

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

52 results for “Programming Languages”

Learn how ShareScore rates datasets ↗
ClinicalTrials.gov36/100

English as a Second Language Health Literacy Program

ClinicalTrials.gov study NCT04125680. IPD Sharing: Not stated. Countries: 1. Publications: 1.

restrictedIPD-UNDECIDEDFeb 2026View details →
dryad36/100

Diversification and change in the R programming language

Open the record for dataset details and reuse information.

publicMar 2023View details →
zenodo32/100

Domain-specific Language for Data-driven Design Time Analyses and Result Mappings for Logic Programs - Data Set

<p>This data set contains additional information to the master thesis of Sebastian Hahner. This includes the prototypical implementation of the domain-specific language (DSL) as well as sample scenarios and mapping results.</p>

opencc-by-4.0Aug 2020View details →
zenodo32/100

Comparing Action-Oriented Language in the Assessment of EFL Writing: An Action Research for Combining the First-Year Instruction of the National Curriculum and the International Baccalaureate (IB) Diploma Program in a Finnish High School

<p>The data here were used to complete my master&#39;s thesis. Data have been anonymized and can thus be used freely. The thesis can be downloaded from the following URL:</p> <p>https://helda.helsinki.fi/handle/10138/327351</p>

opencc-by-4.0Jan 2021View details →
zenodo32/100

Artifact of Program Selection from Large Language Models

<p>Artifact of <em>Program Selection from Large Language Models</em>, including documentation, source code, and experimental data.</p>

opencc-by-4.0Dec 2023View details →
zenodo32/100

ASYMPTOTIC PERFORMANCES OF POPULAR PROGRAMMING LANGUAGES FOR POPULAR SORTING ALGORITHMS

Open the record for dataset details and reuse information.

opencc-by-4.0Jan 2023View details →
zenodo32/100

Leveraging Search-Based and Pre-Trained Code Language Models for Automated Program Repair

<p>This page serves as supplementary material for the article: <strong>Leveraging Search-Based and Pre-Trained Code Language Models for Automated Program Repair</strong>. Here, we provide the ARJACLM code utilized in the study, enabling other researchers to replicate the experiments and further develop the tool.&nbsp;</p>

opencc-by-4.0Nov 2024View details →
zenodo32/100

"Temporal Network Dataset of OSS Programming Language Ecosystems" reproducibility package

<p>Reproducibility package for the paper submission&nbsp;&quot;Temporal Network Dataset of OSS Programming Language Ecosystems&quot;&nbsp; to the Mining Software Repositories 2022 conference. Contains the dataset, extracted metrics, and all scripts used for the construction and the analysis done in the paper.</p> <p>&nbsp;</p> <p>Anonymised for the double-blind review process.</p>

opencc-by-4.0Jan 2022View details →
zenodo32/100

Source Code Classifications: Code Dataset of Programming Languages

Open the record for dataset details and reuse information.

opencc-by-4.0May 2024View details →
zenodo32/100

Dataset related to: An Empirical Study on Low Code Programming using Traditional vs Large Language Model Support

<p>This repository contains data and prompts related to our research. The files included are:</p> <p><strong>prompt.py</strong>: This script contains the original prompts used in our study.</p> <p><strong>LLM_lowcode.mx20</strong>: This file includes posts and annotation data related to LLM-based low-code platforms.</p> <p><strong>Traditional_lowcode.mx20</strong>: This file includes posts and annotation data related to traditional low-code platforms.</p> <p>The .mx20 files can be opened using the MAXQDA software, which can be downloaded from the official website. MAXQDA offers a 14-day free trial.</p>

opencc-by-4.0Jan 2024View details →
zenodo32/100

Evaluation Data For "Feature-Sensitive Coverage for Conformance Testing of Programming Language Implementations"

<p>It contains the evaluation data for PLDI 2023 paper:&nbsp;&quot;Feature-Sensitive Coverage for Conformance Testing of Programming Language Implementations&quot;</p>

opencc-by-4.0Mar 2023View details →
ClinicalTrials.gov32/100

Demonstrating the Efficacy of a Spanish-Language Program for Latino Dementia Caregivers

ClinicalTrials.gov study NCT06747637. IPD Sharing: NO. Countries: 1. Publications: 1.

closedIPD-NOFeb 2026View details →
ClinicalTrials.gov32/100

Pilot Study to Test the Feasibility and the Efficacy of the German Language Adapted PRO-SELF© Plus Pain Control Program

ClinicalTrials.gov study NCT00920504. IPD Sharing: Not stated. Countries: 1. Publications: 2.

restrictedIPD-UNDECIDEDFeb 2026View details →
ClinicalTrials.gov32/100

Implementation of Incredible Years for Autism and Language Delay Program in Spain

ClinicalTrials.gov study NCT04358484. IPD Sharing: NO. Countries: 1. Publications: 1.

closedIPD-NOFeb 2026View details →
zenodo28/100

Leveraging Natural Language for Program Search and Abstraction Learning Regex

<p>Program synthesis dataset containing text editing tasks and language annotations (synthetic and human annotated) for the&nbsp;Leveraging Natural Language for Program Search and Abstraction Learning (currently under review at NeurIPS 2020). Will be deanonymized upon review.</p>

opencc-by-4.0Jun 2020View details →
zenodo28/100

Dataset of the Automatic Unit Test Generation for Programming Assignments Using Large Language Models

Open the record for dataset details and reuse information.

opencc-by-4.0Sep 2024View details →
zenodo28/100

THE IMPORTANCE OF PROGRAMMING LANGUAGES IN THE FIELD OF ARTIFICIAL INTELLIGENCE

Open the record for dataset details and reuse information.

opencc-by-4.0Oct 2024View details →
zenodo28/100

Automated Program Repair in the Era of Large Pre-trained Language Models

<p>Code used for the paper along with the generated outputs</p>

opencc-by-4.0May 2023View details →
zenodo24/100

Leveraging Natural Language for Program Search and Abstraction Learning Graphics

<p>Program synthesis dataset containing graphics programs tasks and language annotations (synthetic and human annotated) for the&nbsp;Leveraging Natural Language for Program Search and Abstraction Learning (currently under review at NeurIPS 2020). Will be deanonymized upon review.</p>

opencc-by-4.0Jun 2020View details →
zenodo24/100

Collaborative Live Coding With Glicol Music Programming Language

<p>Web Audio Conference 2021 - Online - July 5-7</p> <p><a href="https://webaudioconf2021.com/performance-c-1">https://webaudioconf2021.com/performance-c-1</a></p> <p>Qichao Lan (University Of Oslo); Alexander Refsum Jensenius (University Of Oslo) Abstract: Glicol is a graph-oriented live coding language developed with Rust, WebAssembly and AudioWorklet. This language can run in its web-based IDE that supports collaborative coding. In this performance, we invite the participants from the Glicol workshop to join the performance virtually. Each performer will be given the opportunity to write at least one musical loop, and the first author will be in charge of when to execute the code. The music style will be improvised experimental/ambient techno. Intro music by Caucenus - soundcloud.com/caucenus</p>

opencc-by-4.0Jul 2021View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record