Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

16

datasets available to search

ShareScore release 0.9.0

Reset

Dataset results

16 results for “PROSODY”

Learn how ShareScore rates datasets ↗
zenodo40/100

Princeton Prosody Archive Dataset Generated from T. V. F. Brogan's Original Bibliography

<p><strong>Overview</strong></p> <p>The <a href="https://prosody.princeton.edu/">Princeton Prosody Archive</a>&nbsp;(PPA) is a full-text searchable database of thousands of digitized books in English published between 1570 and 1923. The Archive collects historical documents and highlights discourses about the study of language, the study of poetry, and where and how these intersect and diverge. Currently, the PPA is comprised of public domain texts held by the <a href="https://www.hathitrust.org/">HathiTrust Digital Library</a>. It is a <a href="https://cdh.princeton.edu/projects/princeton-prosody-archive/">Sponsored Project</a> of <a href="https://cdh.princeton.edu/">the Center for Digital Humanities at Princeton</a>.</p> <p><strong>Dataset</strong></p> <p>The spreadsheets in this dataset were generated using T. V. F. Brogan&#39;s&nbsp;1981 bibliography&nbsp;<em><a href="http://oregonstate.edu/versif/resources/evrg/sepfiles.html">English Versification, 1570-1980: A Reference Guide With a Global Appendix</a></em>, which provides the foundation for the Archive&#39;s holdings. Because the PPA is an ongoing project, we&nbsp;are making the full Brogan dataset available to scholars here, as well as two curated versions that pose particular data problems for us: first, a list of HathiTrust-held excerpts cited in Brogan (HathiTrust does not index periodicals or journals to allow for excerpting, so we have yet to integrate these into the PPA); and second, a list of public domain items that are not held by HathiTrust with their digital and/or analog locations and hyperlinks where available&nbsp;(our platform does not yet support integrating non-HathiTrust material into the PPA).&nbsp;</p> <p><strong>Data fields</strong></p> <ul> <li><strong>ID:</strong> unique alphanumerical ID assigned by&nbsp;Brogan</li> <li><strong>PPA Checked:</strong> indicates whether or not the work is in the PPA</li> <li><strong>Hathi?</strong> indicates whether or not the work is in the HathiTrust Digital Library&nbsp;</li> <li><strong>YEAR:</strong> year of the work&#39;s publication</li> <li><strong>AUTHOR:</strong> author of the work</li> <li><strong>TITLE:</strong> title of the work</li> <li><strong>PUBLISHER: </strong>publisher of the work</li> <li><strong>EXCERPT:</strong> indicates whether or not the work is an excerpt inside of a larger work (such as an article, essay, book chapter, etc)</li> <li><strong>JOURNAL:</strong> indicates whether or not the work is contained within a journal</li> <li><strong>ENUM:</strong> indicates multivolume works</li> <li><strong>VOLUME IDs: </strong>unique alphanumerical IDs for digitized copies of the work assigned by HathiTrust&nbsp;</li> <li><strong>NOTES: </strong>discursive notes field for data collectors</li> </ul> <p><em>additional fields&nbsp;</em><em>for Not-in-HT_Brogan-List_06.24.19 only:</em></p> <ul> <li><strong>In PUL?</strong> indicates whether or not the Princeton University Library has a copy of the work</li> <li><strong>PUL LOCATION:</strong> call number for works held by the Princeton University Library or its shared collections (ReCAP)</li> <li><strong>ALTERNATE LOCATION: </strong>analog and/or digital locations of works not held by PUL</li> </ul>

opencc-by-4.0Jun 2019View details →
zenodo40/100

ADEPT: A Dataset for Evaluating Prosody Transfer

<p>The ADEPT dataset consists of prosodically-varied natural speech samples for&nbsp;evaluating prosody transfer in english text-to-speech&nbsp;models.&nbsp;The samples include global variations reflecting emotion and interpersonal attitude, and local variations reflecting topical emphasis, propositional attitude, syntactic phrasing and marked tonicity.</p> <p>Txt and wav files are organised according to the folder structure {speech_class}/{subcategory_or_interpretation}/{filename}, where filename&nbsp;follows the naming convention {speaker}_{utterance_id}. Speakers comprise &#39;ad00&#39; (female voice) and &#39;ad01&#39; (male voice). For classes with multiple&nbsp;interpretations, we provide the interpretations used in&nbsp;the disambiguation tasks in&nbsp;&#39;adept_prompts.json&#39;.</p> <p>The corpus only includes prosodic variations that listeners are able to distinguish with reasonable accuracy, and we report these figures as a benchmark against which text-to-speech prosody transfer can be compared. More details can be found in our pre-print about the dataset (https://arxiv.org/abs/2106.08321).</p>

opencc-by-4.0Jun 2021View details →
zenodo36/100

Intonational Speech Prosody Encoding in Human Auditory Cortex

<p>This dataset contains data and results associated with the manuscript, "Intonational Speech Prosody Encoding in Human Auditory Cortex", as well as code used to analyze the data and generate the figures of the manuscript. </p> <p>intonatang-2017.7.17.tar.gz contains the entire project, including neural data, stimulus sound files, analysis code, and documentation. The code and documentation can also be viewed on Github at https://github.com/ChangLabUcsf/intonatang.</p> <p>We additionally included each block of neural data and a zipped file containing the experimental, acoustic stimuli as separate files in this dataset. The data comprise three experiment types, "Speech", "Non-speech control", and "Non-speech missing f0 control". The neural data files are named with a subject identification number and a block number.</p> <p>Speech:</p> <ol> <li>EC113_B13</li> <li>EC113_B20</li> <li>EC113_B21</li> <li>EC118_B3</li> <li>EC118_B7</li> <li>EC118_B13</li> <li>EC122_B30</li> <li>EC122_B40</li> <li>EC122_B43</li> <li>EC122_B53</li> <li>EC123_B4</li> <li>EC123_B5</li> <li>EC123_B10</li> <li>EC125_B13</li> <li>EC125_B1044</li> <li>EC129_B10</li> <li>EC129_B16</li> <li>EC129_B37</li> <li>EC131_B47</li> <li>EC131_B48</li> <li>EC137_B7</li> <li>EC137_B10</li> <li>EC142_B36</li> <li>EC142_B37</li> <li>EC143_B9</li> <li>EC143_B11</li> <li>EC143_B13</li> </ol> <p>Non-speech control:</p> <ol> <li>EC122_B33</li> <li>EC122_B45</li> <li>EC123_B11</li> <li>EC123_B16</li> <li>EC125_B30</li> <li>EC129_B40</li> <li>EC129_B42</li> <li>EC131_B54</li> <li>EC131_B59</li> </ol> <p>Non-speech missing f0 control:</p> <ol> <li>EC137_B9</li> <li>EC137_B11</li> <li>EC142_B38</li> <li>EC142_B40</li> <li>EC143_B10</li> <li>EC143_B12</li> <li>EC143_B14</li> </ol> <p>These .mat files contain the following variables: </p> <ul> <li> badTimeSegments - (n_badTimeSegments x 2) <ul> <li>This variable contains manually marked time segments containing epileptiform, electrical, or movement artifacts. Each row indicates one bad time segment, with the start time and end time in seconds.</li> </ul> </li> <li>bcs - (n_bcs) <ul> <li>This array contains manually marked bad channels. These channels from the ECoG grid either had continuous epileptiform activity or signal indistinguishable from noise. The channels are indexed from 0.</li> </ul> </li> <li>ECXXX_BXX_hg_100Hz - (n_chans x n_timepoints) <ul> <li>This variable contains the mean high-gamma analytic amplitude signal for each channel, sampled at 100Hz. The mean is taken across 8 bands between 70-150Hz. The variable name contains the subject number, ECXXX, and block number BXX. </li> </ul> </li> <li>ECXXX_BXX_log_hg_100Hz - (n_chans x n_timepoints) <ul> <li>This variable contains the mean of the natural logarithm of the high-gamma analytic amplitude signal for each channel. The log is taken for each of the 8 bands between 70-150Hz and then averaged.</li> </ul> </li> <li>experiment <ul> <li>This variable holds the experiment type and is either "Speech", "Non-speech control", or "Non-speech missing f0 control".</li> </ul> </li> <li>sentence_numbers - (n_trials) <ul> <li>The integers in this array are the sentence number condition for each trial in this block. The sentence number conditions depend on the experiment type.</li> </ul> <ol> <li>For the "Speech" experiment, the four sentences indicated by 1, 2, 3, and 4 are "Humans value genuine behavior", "Movies demand minimal energy", "Lawyers give a relevant opinion", and "Reindeer are a visual animal".</li> <li>For the "Non-speech control" experiment, the sentence number conditions indicate which sentence from the main experiment the amplitude contour for the control stimuli came from. A sentence number of 5 means that the amplitude contour was flat. </li> <li>For the "Non-speech missing f0 control", the sentence number holds information about the composition of the stimulus (which harmonics were present), whether noise was added, and how much the pitch range was stretched.  <ul> <li>0: 4h + 5h + 6h, no noise, stretch = 1</li> <li>1: f0 + 2h + 3h, no noise, stretch = 1</li> <li>2: 4h + 5h + 6h, noise, stretch = 1</li> <li>3: 4h + 5h + 6h, noise, stretch = 0.5</li> <li>4: 4h + 5h + 6h, noise, stretch = 2</li> </ul> </li> </ol> </li> <li>sentence_types - (n_trials) <ul> <li>The sentence type is the intonation contour condition. Across all experiment types, a sentence type of 1 is Neutral, 2 is Question, 3 is Emphasis 1, and 4 is Emphasis 3.</li> </ul> </li> <li>speakers - (n_trial) <ul> <li>The speaker conditions depend on the experiment. <ul> <li>The speaker condition for the "Speech" experiment is an integer between 1 and 3. 1 is the low-formant, low-pitch male speaker. 2 is the high-formant, high-pitch female speaker. 3 is the low-formant, high-pitch female speaker. The absolute pitch values of speakers 2 and 3 match, while the formant values of speaker 1 and 3 match.</li> <li>The speaker condition for both of the two non-speech experiments are either 1 or 2. 1 means low absolute pitch (male) and 2 means high absolute pitch (female).</li> </ul> </li> </ul> </li> <li>stims - (n_trials) <ul> <li>This array holds the stimulus name that was played for each trial. The names refer to the wav files in the tokens, tokens_nonspeech, and tokens_missing_f0 folders, for the "Speech", "Non-speech control", and "Non-speech missing f0 control" experiments, respectively.</li> </ul> </li> <li>times - (n_trials) <ul> <li>This array contains the onset times of each trial in seconds.</li> </ul> </li> </ul>

opencc-by-sa-4.0Jul 2017View details →
zenodo32/100

Appendix 5 Semantic prosody of ser+PP

<p>Appendix 5 of the paper "<span>Between source language constructions and target language expectations. </span>An analysis of passive constructions in translated and non-translated Spanish", published in <em>Review of Cognitive Linguistics.</em></p> <p>It contains an additional analysis of the semantic prosody of the verbs used with Spanish ser+PP.&nbsp;&nbsp;</p>

opencc-by-4.0Apr 2024View details →
zenodo32/100

Prosody [IO Islamic 1077]

<ul> <li><strong>Prosody.</strong></li> <li><strong>This manuscript is now IO Islamic 1077&nbsp;</strong><strong>in the India Office collections.</strong></li> <li><strong>[metadata:</strong><a href="https://de.wikipedia.org/wiki/Otto_Loth">&nbsp;<strong>Otto Loth,&nbsp;</strong></a><strong><em><a href="http://doi.org/10.5281/zenodo.3923636">A Catalogue of the Arabic Manuscripts in the Library of the India Office</a></em>, (volume 1), no. 845&nbsp;here with further notations and hyperlinks]</strong>.</li> </ul> <p>PROSODY.</p> <p><a href="https://archive.org/details/catalogueofarabi01greauoft/page/244/mode/2up?view=theater"><strong>845</strong></a>.</p> <p>1077. Size 7 in. by 4<sup>1/2</sup> in.; foll. 75. Seventeen lines in a page.</p> <p>هذا الكتاب المسمى بالكافى فى علم العروض و القوافى فى شرح القصيدة الساوية التى نظمها الامام صدر الدين محمد الساوى رحمه الله تع آمين.</p> <p>A Commentary on Ṣadr al-d&icirc;n Muḥammad <em>S&acirc;w&icirc;&rsquo;</em>s Ḳaṣ&icirc;dah on Metre and Rhyme. This is a commentary by قال and اقول . The author, who is not mentioned, is, according to <a href="https://en.wikipedia.org/wiki/Kashf_al-Zunun">Ḥ. Kh. </a>iv. 204 (<em>v</em>. عروض الساوى), &lsquo;UBAIDALLAH B. &lsquo;ABD AL-K&Acirc;FI b. &lsquo;Abd al-maj&icirc;d &lsquo;Ubaid&icirc;, and this is his second and shorter commentary. Cf. Ḥ. Kh. v. 21, 296; and Catal. Mus. Brit. 202, <em>b</em>.</p> <p>Plainly written by two hands. Completed by &lsquo;Abd al-&rsquo;az&icirc;z b. Ḥusain Nahrw&acirc;l&icirc;. Collated with the original copy, which belonged to &lsquo;Abd al-malik b. Abu&rsquo;l-barak&acirc;t البنبانى , by Ism&acirc;&rsquo;&icirc;l b. Aḥmad Ja&rsquo;far Ḥusain&icirc;, in Rab&icirc;&rsquo; I., 1017.</p> <p>A table of the metres and their varieties is on the title-page.</p> <p>[Gaikwar.]</p> <p>&nbsp;</p>

opencc-by-4.0Nov 2021View details →
zenodo32/100

Data belonging to the thesis intitled "Prosody in Language Contact. French and Vietnamese"

<p>This dataset belongs to our doctoral thesis intitled &quot;Prosody in Language Contact. French and Vietnamese&quot;. The dataset contains recordings of oral data as well as tables of written data, transcriptions and annotations. It also contains metadata about oral and written data. Detailed information about this data and how to understand it can be taken from our thesis.</p> <p>In the thesis, we are dealing with prosodic language contact between the languages Vietnamese and French. Our research is devoted to different language contact situations as well as to both directions of language contact. The starting point of our experimental research is the observation of French loanwords in Vietnamese. In this context, we examine prosodic adaptation patterns that speakers have undertaken to adapt the loanwords to Vietnamese. The focus is on the repair of structures that are illicit in Vietnamese: Consonants in certain positions in the syllable are replaced or deleted, consonant clusters are dissolved by epenthesis or by deletion of one of the two consonants, syllable boundaries are shifted, consonants or consonant slots in certain syllabic structures are doubled, and tones are assigned to syllables according to certain patterns.</p> <p>The aim of our experimental investigations then is to find out whether similar or different patterns occur in an instantaneous situation of language contact. Monolingual speakers of Vietnamese are exposed to French stimuli and asked to reproduce them in three different conditions. We have additionally conducted the same experiment with learners of French whose first language is Vietnamese. The experimental data show many similar patterns to the loanword data. However, the data from monolingual speakers in particular display much more variability. Finally, we reverse the direction of language contact: Native speakers of French are asked to reproduce Vietnamese stimuli. In this case, the question is whether certain patterns can be reversed in the opposite direction, which can partly be observed.</p> <p>With the help of the present work, we can gain a deep understanding of systematicity and variability in borrowing and second language acquisition. We conclude that from a phonological perspective, there are many similarities between the two fields of prosodic language contact, and we suggest, for the phenomena under consideration, to understand some aspects in their complexity in a gradual rather than a categorical way.</p>

opencc-by-4.0Nov 2021View details →
zenodo32/100

A Systematic Review and Bayesian Meta-Analysis of Acoustic Measures of Prosody in Parkinson's Disease

<p>The folder contains the dataset used for this study and a file defining the titles of the dataset.</p> <div> <div> <div>&nbsp;</div> <div> <div> <div>&nbsp;</div> <div> <p>&nbsp;</p> <p>&nbsp;</p> </div> </div> </div> </div> </div>

opencc-by-4.0Apr 2024View details →
ClinicalTrials.gov32/100

Affective Prosody Recognition and Auditory Intervention for Children With Autism Spectrum Disorders (ASD)

ClinicalTrials.gov study NCT04427150. IPD Sharing: YES. Countries: 1. Publications: 33.

controlledIPD-YESFeb 2026View details →
ClinicalTrials.gov32/100

Adaptation of a Rehabilitation Program for Prosody and Its Application on Egyptian Hearing Impaired Children

ClinicalTrials.gov study NCT04691830. IPD Sharing: NO. Countries: 1. Publications: 1.

closedIPD-NOFeb 2026View details →
ClinicalTrials.gov32/100

Synchrony and Reciprocity of Body Movements and Prosody Between Psychotherapist and Patient

ClinicalTrials.gov study NCT06463951. IPD Sharing: NO. Countries: 1. Publications: 1.

closedIPD-NOFeb 2026View details →
zenodo28/100

Prosody annotated recited poetry in Finnish

<p>This dataset contains the audio for four poems recited in Finnish in the wav folder and their corresponding transcription with prosodic annotation in the text folder. The audio files contain a single verse each and their order has been shuffled.&nbsp;</p> <p>Please, cite the following paper, if you use the data:</p> <p>H&auml;m&auml;l&auml;inen, M., &amp; Rueter, J.&nbsp;(2020).&nbsp;<a href="https://tuhat.helsinki.fi/ws/portalfiles/portal/159764495/runonlausunta.pdf">Runonlausunnan prosodia ja sen mallintaminen koneellisesti puhesynteesill&auml;</a>. In&nbsp;<em>Материалы Международного образовательного салона&nbsp;</em>(pp. 5-17)</p>

opencc-by-nc-4.0Dec 2020View details →
ClinicalTrials.gov28/100

Prosody Assessment After Right Hemisphere Stroke

ClinicalTrials.gov study NCT05874011. IPD Sharing: YES. Countries: 1. Publications: 0.

controlledIPD-YESFeb 2026View details →
ClinicalTrials.gov28/100

Emotional Prosody Treatment in Parkinson's

ClinicalTrials.gov study NCT01956266. IPD Sharing: YES. Countries: 1. Publications: 0.

controlledIPD-YESFeb 2026View details →
ClinicalTrials.gov24/100

Music Therapy for Speech and Prosody in Autistic Children (MTSPAC)

ClinicalTrials.gov study NCT06110884. IPD Sharing: NO. Countries: 1. Publications: 0.

closedIPD-NOFeb 2026View details →
ClinicalTrials.gov24/100

Emotional Prosody Recognition and Decision Making Inf fMRI and Vulnerability to Suicide

ClinicalTrials.gov study NCT02901769. IPD Sharing: Not stated. Countries: 1. Publications: 0.

restrictedIPD-UNDECIDEDFeb 2026View details →
zenodo12/100

When prosody meets syntax: The processing of the syntax-prosody interface in children with developmental dyslexia and developmental language disorder.

<p>Dataset containing the raw scores obtained at the various experimental and standardized tests by&nbsp;the three groups of children described in the paper &quot;When prosody meets syntax: The processing of the syntax-prosody interface in children with developmental dyslexia and developmental language disorder.&nbsp;<em>Lingua</em>, <a href="https://doi.org/10.1016/j.lingua.2019.03.008">https://doi.org/10.1016/j.lingua.2019.03.008</a></p>

restrictedSep 2020View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record