Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

166

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

166 results for “dialects”

Learn how ShareScore rates datasets ↗
ClinicalTrials.gov24/100

Evaluation of a Modified Dialectical Behavior Therapy Program

ClinicalTrials.gov study NCT01635556. IPD Sharing: NO. Countries: 1. Publications: 0.

closedIPD-NOFeb 2026View details →
ClinicalTrials.gov24/100

Pain Acceptance Training in Patients Experiencing Emotional Distress and Somatic Symptoms: Examination of Dialectical Thinking as a Mediating Factor

ClinicalTrials.gov study NCT07067619. IPD Sharing: NO. Countries: 1. Publications: 0.

closedIPD-NOFeb 2026View details →
zenodo20/100

STT4SG-350: A Speech Corpus for All Swiss German Dialect Regions

<p>We present STT4SG-350 (Speech-to-Text for Swiss German), a corpus of Swiss German speech, annotated with Standard German text at the sentence level. The data is collected using a web app in which the speakers are shown Standard German sentences, which they translate to Swiss German and record. We make the corpus publicly available. It contains 343 hours of speech from all dialect regions and is the largest public speech corpus for Swiss German to date. Application areas include automatic speech recognition (ASR), text-to-speech, dialect identification, and speaker recognition. Dialect information, age group, and gender of the 316 speakers are provided. Genders are equally represented and the corpus includes speakers of all ages. Roughly the same amount of speech is provided per dialect region, which makes the corpus ideally suited for experiments with speech technology for different dialects. We provide training, validation, and test splits of the data. The test set consists of the same spoken sentences for each dialect region and allows a fair evaluation of the quality of speech technologies in different dialects. We train an ASR model on the training set and achieve an average BLEU score of 74.7 on the test set. The model beats the best published BLEU scores on 2 other Swiss German ASR test sets, demonstrating the quality of the corpus.</p>

restrictedOct 2023View details →
zenodo20/100

A Linguistic Study on Rhyming in the Beijing Dialect. Supplementary Material

<p>A Linguistic Study on Rhyming in the Beijing Dialect. Supplementary Material</p>

opencc-by-4.0Nov 2019View details →
ClinicalTrials.gov20/100

Adapted Dialectical Behaviour Therapy for Adolescents With Deliberate Self-Harm: A Pre-post Observational Study

ClinicalTrials.gov study NCT02988037. IPD Sharing: NO. Countries: 0. Publications: 0.

closedIPD-NOFeb 2026View details →
zenodo16/100

At the intersection between acoustic phonetics and information structure. An empirical investigation on the phonetic realization of vowel length in three Ligurian dialects

<p>This repository contains the data gathered within the research project &ldquo;At the intersection between acoustic phonetics and information structure. An empirical investigation on the phonetic realization of vowel length in three Ligurian dialects&rdquo;, funded by the Swiss National Science Foundation (University of Zurich, 01.06.2018-31.08.2022; Project Nr. 100015_178932)</p> <p>The materials include:</p> <p>- Production data. Recordings of each speaker for each experiment (.wav).</p> <p>- Perception data. Audio stimuli (.wav), stimuli conditions (.xlsx), PsychoPy test files (.psyexp).</p> <p>- Metadata. Documentation about the experiments and the speakers involved.</p> <p>- A scientific report with a short description of each experiment as well as a summary of the research work carried out within the project.</p> <p>For more information, see the ReadMe .txt files in the folder.</p>

restrictedAug 2022View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record