Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
54
datasets available to search
ShareScore release 0.9.0
Dataset results
54 results for “Phonology”
A Finite State Tranducer that models Chinese Historical Phonology
<p>This is a finite state transducer that tries to model Chinese historical phonology from Old Chinese as reconstructed by Baxter and Sagart in <em>Old Chinese: a New Reconstruction</em> (Oxford, 2014) to Middle Chinese as presented in the system of Baxter in <em>A Handbook of Old Chinese Phonology </em>(Mouton, 1992).</p>
Source Code Accompanying the Paper "More on network approaches in Historical Chinese Phonology (音韻學)"
<p>First version of the source code and data accompanying the paper "More on Network Approaches in Historical Chinese Phonology".</p> <p>This paper is available here:</p> <ul> <li>List, Johann-Mattis (2018): <strong>More on network approaches in Historical Chinese Phonology (音韻學)</strong>. Paper prepared for the <em>LFK Society Young Scholars Symposium</em>. Taibei: Li Fang-Kuei Society ofr Chinese Linguistics. URL: <a href="https://hal.archives-ouvertes.fr/hal-01706927">https://hal.archives-ouvertes.fr/hal-01706927</a>.</li> </ul> <pre><code>@InProceedings{List2018a, author = {List, Johann-Mattis}, title = {{More on Network Approaches in Historical Chinese Phonology (音韻學)}}, booktitle = {{LFK Society Young Scholars Symposium}}, year = {2018}, publisher = {Li Fang-Kuei Society for Chinese Linguistics}, pdf = {https://hal.archives-ouvertes.fr/hal-01706927/file/main.pdf}, url = {https://hal.archives-ouvertes.fr/hal-01706927}, address = {Taipei}, hal_id = {hal-01706927}, } </code></pre> <p>See the README.md for mor information.</p> <ul> <li> </li> </ul>
Audio files referred to in Roettger, Timo B. (2017). Tonal Placement in Tashlhiyt - How an intonation system accommodates to adverse phonological environments.
<p>Audio files referred to in Roettger, Timo B. (2017). Tonal Placement in Tashlhiyt - How an intonation system accommodates to adverse phonological environments. Berlin: Language Science Press (DOI 10.5281/zenodo.814472)</p>
CLDF dataset derived from Lee's "Phonological Features of Caijia" from 2023
<p>Cite the source of the dataset as:</p> <blockquote> <p>Lee, Man Hei (2023): Phonological features of Caijia that are notable from a diachronic perspective. Journal of Historical Linguistics. DOI: https://doi.org/10.1075/jhl.21025.lee</p> </blockquote>
CLDF dataset derived from Mann's "Phonological Reconstruction of Proto Northern Burmic" from 1998
<p>Cite the source of the dataset as:</p> <blockquote> <p>Mann, Noel W. 1998. A phonological reconstruction of Proto Northern Burmic. (PhD Thesis).</p> </blockquote>
CLDF dataset derived from Hóu's "Phonological Database of Chinese Dialects" from 2004
<p>Cite the source of the dataset as:</p> <blockquote> <p>Hóu, J. (2004): Xiàndài Hànyǔ fāngyán yīnkù 现代汉语方言音库 [Phonological database of Chinese dialects]. Shànghǎi: Shànghǎi Jiàoyù.</p> </blockquote>
CLDF dataset derived from Sūn's "Tibeto-Burman Phonology and Lexicon" from 1991
<p>Cite the source of the dataset as:</p> <blockquote> <p>Sūn, Hóngkāi 孙宏开 (1991): Zangmianyu yuyin he cihui 藏缅语音和词汇 [Tibeto-Burman phonology and lexicon]. Beijing: Chinese Social Sciences Press.</p> </blockquote>
CLDF dataset derived from Lundgren's "Phonological Reconstruction of Proto-Omagua-Kokama-Tupinambá" from 2020
<p>Cite the source of the dataset as:</p> <blockquote> <p>Lundgren, Olof (2020): A phonological reconstruction of Proto-Omagua-Kokama-Tupinambá. Master's thesis. Lund: Lund University.</p> </blockquote>
Data for: Unmasking the Effects of Orthography, Semantics, and Phonology on 2AFC Visual Word Perceptual Identification
<p>This data was used in analyses for "Unmasking the Effects of Orthography, Semantics, and Phonology on 2AFC Visual Word Perceptual Identification".</p>
Western Thrace Turkish: Phonology - Phonetic Features, Morphology and Syntax
<p>These snippets present linguistic features of Western Thrace Turkish, a Balkan Turkish dialect spoken in Northeastern Greece. This lecture is part of the lecture series: <em>Glottothèque: Languages of the Anatolia, Caucasus, Iran, Mesopotamia; grammatical snippets online </em>(electronic resource). Bamberg, Cambridge, Göttingen, Moskow, Nicosia, Paris: LACIM network, at https://spw.uni-goettingen.de/projects/lacim/, edited by Christiane Bulut, Anaïd Donabédian-Demopoulos, Geoffrey Haig, Geoffrey Khan, Pollet Samvelian, Stavros Skopeteas, Nina Sumbatova.</p>
West-Central Thailand Pwo Karen Phonology
<p>West-Central Thailand Pwo Karen consonant, vowel, and tone recordings, along with examples of complex monosyllables, sesquisyllables, and disyllables. All of these words can be found in the word lists in the appendix of Phillips (1996).</p> <p>In the file names, consonant sounds are symbolized as follows. Note that glottal is represented by q and the voiced velar approximant is represented by g.</p> <p>p, t, c, k, q</p> <p>ph, th, ch, kh</p> <p>b, d</p> <p>m, n, ng</p> <p>s, sh, x, h</p> <p>w, r, j, g</p> <p>l</p> <p>Vowel sounds are symbolized as follows:</p> <p>Oral vowels:</p> <p>i, ue, u</p> <p>e, oe, o</p> <p>a</p> <p>ai, aue</p> <p>Nasalized vowels:</p> <p>ing, ueng, ung</p> <p>eng, oeng, ong</p> <p>oang</p> <p>aing</p> <p>Tones are symbolized as follows:</p> <p>H - high rise</p> <p>M - mid rise</p> <p>L - Low falling</p> <p>HF - High falling</p> <p>HQ - High glottalized</p> <p>LQ - Low glottalized</p> <p> </p>
Summary data to support Hall (2019), '(e) in Normandy: The sociolinguistics, phonology and phonetics of the "Loi de Position"'
<p>Article abstract:</p> <p>This article uses the pronunciation of stressed Intonational Phrase-final /ε/ and /e/ in two communities in Normandy, France, to illustrate the convergence of two sociolinguistic processes on the same phonological result: increasing application of the <em>Loi de Position</em>. In both communities (one rural and further from Paris, one urban and closer to Paris), there is now no consistent community-wide phonetic distinction between the two phonemes in that environment. It is suggested that the <em>Loi de Position</em> is already widely applied in the rural site, but speakers are still conscious of the formal norm whereby it is not applied; for the urban site, apparent-time changes for this variable reflect changes in Parisian speech. The theoretical implications of the study concerning speakers’ organisation of their vowel-space, and concerning the increasing application of the <em>Loi de Position</em> in the French of France, are examined. These conclusions are reached by per-speaker analysis of F1 and F2 separately from each other (rare in French linguistics). As a measure of community cohesion, the article introduces to linguistics the coefficient of variation (more common in biology and medicine).</p>
Duhumbi Phonology: Origin of non-native and marginal phonemes
<p>These files present the origins of the Duhumbi non-native and marginal consonant and vowel phonemes a supplementary material to section 2.2.1, 2.2.2 and 4.10.1 and 4.10.2 of the Duhumbi grammar.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Duhumbi Phonology - Approximant rhymes
<p>These data present the arguments for the analysis of the Duhumbi approximant rhymes /oj ~ uj, ej ~ aj, aw ~ ow/ rather than distinctive diphthong phonemes. The zip files contain sound files for illustration.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Duhumbi Phonology - Coda clusters
<p>This data set provides the sound files and analysis that show the coda consonant clusters in Duhumbi.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Duhumbi Phonology - Onset clusters
<p>This set presents the overview of the Duhumbi onset clusters and their origins, including the decision made for the phonological description of the language, supplementing the information in section 2.5.2 of the Duhumbi grammar and including several sound files.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Duhumbi Phonology - Pitch and Tone
<p>Duhumbi is classified as a non-tonal language. Unlike several Central and East Bodish languages, in Duhumbi, distinctive tone cannot be conclusively attested even on the most common consonants affected by tone, namely the nasals /n, m, ng, ny/ and the approximants /l, r, w, y/. Nonetheless, Duhumbi clearly shows the first signs of the development of distinctive tone, or at least a distinctive pitch contrast. There are five main phonotactic conditions in which pitch distinctions can be observed, of which some form of glottalisation is the main one. The distinctive glottal constricted vowels /a, e, o, u/ in open syllables all have a rising pitch. The same holds for pre-glotallised initial vowels. Also, a high-falling pitch contour can be shown to be triggered by glottal reinforcement of syllable-final plosives. In addition, a high-falling versus low-level contour pitch contrast has been observed between lexemes with unaspirated unvoiced, aspirated unvoiced and unaspirated voiced plosive onsets. Finally, of the few attested contrastive minimal pairs for tone, the high onset/high-falling pitch lexemes of these minimal pairs can be shown to have Bodish cognates with a high onset, and the pitch contrast may thus be presumed to be borrowed as phonological feature of the entire lexeme.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Duhumbi Phonology - Syllabic and Word Stress
<p>In Duhumbi, stress in general is non-distinctive, prosodic, and relatively unpronounced. In both disyllabic and polysyllabic words, stress falls on the first syllable. This also holds for polymorphemic lexemes, such as inflected words with suffixes. In glossary items in the Duhumbi lexicon, stress is indicated by a stress mark [ˈ] before the stressed syllable, whenever it is not predictable. Phrase stress is generally initial and falling towards the end of the phrase, combined with dependant-head word order in which subject/object precede the verb and we find postpositions rather than prepositions, although nouns always precede adjectives, but adverbs generally precede verbs. Word order is relatively flexible. </p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Duhumbi Phonology - Iambic to Trochaic Rhythm?
<p>There are indications for a change from iambic rhythm in the purported linguistic ancestor of Duhumbi, Proto-Western Kho-Bwa, to the trochaic rhythm of present-day Duhumbi.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
Duhumbi Phonology - Glottal Stop and Glottal Constriction
<p>The phonologic position of the glottal stop /ʔ/ in Duhumbi is elusive. The glottal stop and glottal</p> <p>constriction are clearly phonetic features among all speakers of all varieties of the language.</p> <p>However, there is no convincing evidence that it occurs as a distinctive phoneme. Rather, the glottal</p> <p>stop co-occurs as a glottal closure on onset vowels /a, e, i, o/, as a sub-phonemic feature combined</p> <p>with creaky voice and rising pitch on certain open vowels /a, e, o, u/, or as allophone of the plosives,</p> <p>mainly the velar plosive /k/, in coda position in certain phonotactic environments.</p> <p>This file presents an overview of the position of the glottal stop and glottal constriction in Duhumbi and presents some of the recordings that form the basis of it.</p> <p>This material is made freely available to everyone for informative or scientific purposes as long as the source (this DOI) / the collectors are properly credited. Please note that use of the material for commercial purposes <em><strong>of any kind</strong>, which includes conversion into commercial audio-visual media (documentaries etc.), storage and dissemination through sites that require registration & payment for access, or sites that rely on advertisement (including YouTube) </em>is <strong>not</strong> permitted without <strong>specific written consent</strong> from the speakers and their community, obtained through the collectors of the material. By downloading our material, you agree to these restrictions.</p> <p>This data set falls under the Attribution-NonCommercial-ShareAlike (CC BY-NC-SA) license. This license lets you remix, tweak, and build upon this work non-commercially, as long as you credit us and license your new creations under the identical terms. License Deed on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/">https://creativecommons.org/licenses/by-nc-sa/4.0/</a>. Legal Code on <a href="https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode">https://creativecommons.org/licenses/by-nc-sa/4.0/legalcode</a>.</p> <p>Tim Bodt: bodttim (at) gmail (dot) com</p>
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.