Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

6

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

6 results for “language contacts”

Learn how ShareScore rates datasets ↗
zenodo48/100

Data on the typology and stability of evidentiality in language contact situations

<p>This material contains the dataset from the&nbsp;<a href="https://version.helsinki.fi/gramadapt/evidentiality/" target="_blank" rel="noopener">gitlab repository</a> of the following MA thesis. Please cite the thesis when using the data.</p> <p>Hyv&ouml;nen, Anu. 2024. <em>Typology and stability of evidentiality in language contact situations</em>. MA thesis, University of Helsinki. Openly available at <a href="https://helda.helsinki.fi/items/2ee41e80-0a04-4af4-8a90-a83598447b0e">https://helda.helsinki.fi/items/2ee41e80-0a04-4af4-8a90-a83598447b0e</a>.</p>

opencc-by-4.0Apr 2024View details →
zenodo36/100

Data belonging to the article "Syllabification in Language Contact between French and Vietnamese"

<p>This dataset belongs to an article which will appear in the ICAAL 8 conference proceedings and which is dedicated to the syllabification in two language contact settings between Vietnamese and French. The paper discusses gemination and related phenomena. First, it contributes to the theoretical discussion on borrowing processes in loanwords and presents a model that includes all extant views and also accounts for other language contact situations. &nbsp;In our data analysis, we compare loan data to experimental data, and find similar processes in both scenarios. Syllable boundaries are shifted, consonants are geminated and consonant slots are doubled in certain syllabic environments. The emergence of gemination is especially noteworthy as it is only a marginal phenomenon in French and has not been found to exist in Vietnamese. Results provide evidence for our hypothesis that optimal Vietnamese syllables are closed, as these syllables have the highest frequency in the native lexicon. This goes beyond, but also corresponds to the idea that Vietnamese has a dimoraic minimum, and this paper provides more evidence for it.</p>

opencc-by-4.0Aug 2022View details →
zenodo36/100

CLDF dataset derived from Wang's "Language Contact and Language Comparison" from 2004

<p>Cite the source of the dataset as:</p> <blockquote> <p>Wang, Feng (2004): Language contact and language comparison. The case of Bai. PhD thesis. Hong Kong: City University of Hong Kong.</p> </blockquote>

opencc-by-4.0Jul 2021View details →
zenodo32/100

Data belonging to the thesis intitled "Prosody in Language Contact. French and Vietnamese"

<p>This dataset belongs to our doctoral thesis intitled &quot;Prosody in Language Contact. French and Vietnamese&quot;. The dataset contains recordings of oral data as well as tables of written data, transcriptions and annotations. It also contains metadata about oral and written data. Detailed information about this data and how to understand it can be taken from our thesis.</p> <p>In the thesis, we are dealing with prosodic language contact between the languages Vietnamese and French. Our research is devoted to different language contact situations as well as to both directions of language contact. The starting point of our experimental research is the observation of French loanwords in Vietnamese. In this context, we examine prosodic adaptation patterns that speakers have undertaken to adapt the loanwords to Vietnamese. The focus is on the repair of structures that are illicit in Vietnamese: Consonants in certain positions in the syllable are replaced or deleted, consonant clusters are dissolved by epenthesis or by deletion of one of the two consonants, syllable boundaries are shifted, consonants or consonant slots in certain syllabic structures are doubled, and tones are assigned to syllables according to certain patterns.</p> <p>The aim of our experimental investigations then is to find out whether similar or different patterns occur in an instantaneous situation of language contact. Monolingual speakers of Vietnamese are exposed to French stimuli and asked to reproduce them in three different conditions. We have additionally conducted the same experiment with learners of French whose first language is Vietnamese. The experimental data show many similar patterns to the loanword data. However, the data from monolingual speakers in particular display much more variability. Finally, we reverse the direction of language contact: Native speakers of French are asked to reproduce Vietnamese stimuli. In this case, the question is whether certain patterns can be reversed in the opposite direction, which can partly be observed.</p> <p>With the help of the present work, we can gain a deep understanding of systematicity and variability in borrowing and second language acquisition. We conclude that from a phonological perspective, there are many similarities between the two fields of prosodic language contact, and we suggest, for the phenomena under consideration, to understand some aspects in their complexity in a gradual rather than a categorical way.</p>

opencc-by-4.0Nov 2021View details →
zenodo32/100

Voices from Cabinda: Audio files of the examples in the paper "Você + 2SG in Cabindan Portuguese (Angola): Two Effects of Language Contact" (LaborHistórico)

<p>Audio files of the examples presented throughout the paper "Voc&ecirc; + 2SG in Cabindan Portuguese (Angola): Two Effects of Language Contact" (accepted for publication in LaborHist&oacute;rico). The audio excerpts are taken from my corpus of interviews in Cabinda, which will be made available soon. This fieldwork, conducted in compliance with ethical standards, was made possible through the assistance of the NGO <em>Ajuda de Desenvolvimento de Povo para Povo</em>&nbsp;and with the authorization of the&nbsp;<em>Secretaria Provincial da Cultura </em>in Cabinda. It received funding from the University of Augsburg (<em>F&ouml;rderung des wissenschaftlichen Nachwuchses</em>).</p>

opencc-by-4.0Jul 2024View details →
zenodo32/100

Voices from Cabinda: Audio files of the examples in the chapter of "Afro-Iberian Languages: Contact and Sociohistory" (LangSciPress)

<p>Audio files (and a few videos) corresponding to the examples presented throughout the chapter on Cabindan Portuguese, included in the volume <em>Afro-Iberian Languages: Contact and Sociohistory</em> (Lamberti/Agostinho, eds.; Language Science Press: https://langsci-press.org/catalog/book/392). The chapter provides a panoramic overview of the main structural features of Cabindan Portuguese, preceded by a historical section on the spread of Portuguese in this Central African region, as well as a sociolinguistic section.</p>

opencc-by-4.0Oct 2024View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record