Find research datasets worth reusing
Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.
60
datasets available to search
ShareScore release 0.7.1
Dataset results
60 results for “speech perception”
AVbook, a high-frame-rate corpus of narrative audiovisual speech for investigating multimodal speech perception
<p><strong>Please cite</strong><br> Varano E, Guilleminot P, Reichenbach T. <em>AVbook, a high-frame-rate corpus of narrative audiovisual speech for investigating multimodal speech perception</em>. J Acoust Soc Am. 2023 May 1;153(5):3130. doi: 10.1121/10.0019460. PMID: 37249407.<br> <br> Seeing a speaker's face can help substantially in understanding them, in particular in challenging listening conditions. Research into the neurobiological mechanisms behind the audiovisual integration has recently begun to employ continuous natural speech. However, these efforts are impeded by a lack of high-quality audiovisual recordings of a speaker narrating a longer text. Here we seek to close this gap by developing AVbook, an audiovisual speech corpus designed for cognitive neuroscience studies and audiovisual speech recognition. The corpus consists of 3.6 hours of audiovisual recordings of two speakers, one male and one female, reading 59 passages from a narrative English text. The recordings were acquired at a high frame rate of 119.88 frames per second. The corpus includes a sets of multiple-choice questions to test attention to the different passages. We verified the efficacy of these questions in a pilot study. A short written summary is also provided for each recording. To enable audiovisual synchronization when presenting the stimuli, four videos of an electronic clapperboard were recorded with the corpus. The corpus is available for download to support research into the neurobiology of audiovisual speech processing as well as the development of computer algorithms for audiovisual speech recognition.</p>
Repeatedly experiencing the McGurk effect induces long-lasting changes in auditory speech perception
<p>In the McGurk effect, presentation of incongruent auditory and visual speech evokes a fusion percept different than either component modality. We show that repeatedly experiencing the McGurk effect for 14 days induces a change in auditory-only speech perception: the auditory component of the McGurk stimulus begins to evoke the fusion percept, even when presented on its own without accompanying visual speech. This perceptual change, termed fusion-induced recalibration (FIR), was talker-specific and syllable-specific and persisted for a year or more in some participants without any additional McGurk exposure. Participants who did not experience the McGurk effect did not experience FIR, showing that recalibration was driven by multisensory prediction error. A causal inference model of speech perception incorporating multisensory cue conflict accurately predicted individual differences in FIR. Just as the McGurk effect demonstrates that visual speech can alter the perception of auditory speech, FIR shows that these alterations can persist for months or years. The ability to induce seemingly permanent changes in auditory speech perception will be useful for studying plasticity in brain networks for language and may provide new strategies for improving language learning.</p>
Repeatedly experiencing the McGurk effect induces long-lasting changes in auditory speech perception
Open the record for dataset details and reuse information.
The role of isochrony in speech perception in noise - Dataset
<p>This dataset contains speech stimuli and listener data reported on in Aubanel & Schwartz (2020), DOI: <a href="http://dx.doi.org/10.1038/s41598-020-76594-1">10.1038/s41598-020-76594-1</a>. </p> <p><strong>French data</strong></p> <ul> <li>French sentences are taken from the Fharvard corpus (Aubanel et al., 2020, DOI: <a href="https://dx.doi.org/10.1016/j.specom.2020.07.004">10.1016/j.specom.2020.07.004</a>)</li> <li>Speech material and sentence recordings are available at: <a href="https://dx.doi.org/10.5281/zenodo.1462854">10.5281/zenodo.1462854</a></li> <li><strong>fr_stimuli.zip</strong> contains the stimuli presented to the listeners</li> <li><strong>fr_responses.csv</strong> contains the responses typed by listeners</li> </ul> <p><strong>English data</strong></p> <ul> <li>English sentences are taken from the Harvard corpus (Rothauser et al. 1969)</li> <li>Speech material and sentence recordings are taken from the MAVA corpus, available at: <a href="https://dx.doi.org/10.4227/139/59a4c21a896a3">10.4227/139/59a4c21a896a3</a></li> <li><strong>en_stimuli.zip</strong> contains the stimuli presented to the listeners</li> <li><strong>en_responses.csv</strong> contains the responses typed by listeners</li> </ul> <p> </p>
Data from: Neural correlates of multisensory enhancement in audiovisual narrative speech perception: a fMRI investigation
<p>This fMRI study investigated the effect of seeing articulatory movements of a speaker while listening to a naturalistic narrative stimulus. It had the goal to identify regions of the language network showing multisensory enhancement under synchronous audiovisual conditions. We expected this enhancement to emerge in regions known to underlie the integration of auditory and visual information such as the posterior superior temporal gyrus as well as parts of the broader language network, including the semantic system. To this end we presented 53 participants with a continuous narration of a story in auditory alone, visual alone, and both synchronous and asynchronous audiovisual speech conditions while recording brain activity using BOLD fMRI. We found multisensory enhancement in an extensive network of regions underlying multisensory integration and parts of the semantic network as well as extralinguistic regions not usually associated with multisensory integration, namely the primary visual cortex and the bilateral amygdala. Analysis also revealed involvement of thalamic brain regions along the visual and auditory pathways more commonly associated with early sensory processing. We conclude that under natural listening conditions, multisensory enhancement not only involves sites of multisensory integration but many regions of the wider semantic network and includes regions associated with extralinguistic sensory, perceptual and cognitive processing.</p>
Speech, timbre, and pitch perception in cochlear implant users after chronic use with flat panel CT-based frequency reallocations
Open the record for dataset details and reuse information.
Examining relationship between auditory brainstem responses, cognitive ability, and speech-in-noise perception among young adults with normal hearing thresholds
Open the record for dataset details and reuse information.
Data from: Neural correlates of multisensory enhancement in audiovisual narrative speech perception: a fMRI investigation
Open the record for dataset details and reuse information.
The influence of infant-directed speech on 12-month-olds' intersensory perception of fluent speech
<p>The present study examined whether infant-directed (ID) speech facilitates intersensory matching of audio–visual fluent speech in 12-month-old infants. German-learning infants’ audio–visual matching ability of German and French fluent speech was assessed by using a variant of the intermodal matching procedure, with auditory and visual speech information presented sequentially. In Experiment 1, the sentences were spoken in an adult-directed (AD) manner. Results showed that 12-month-old infants did not exhibit a matching performance for the native, nor for the non-native language. However, Experiment 2 revealed that when ID speech stimuli were used, infants did perceive the relation between auditory and visual speech attributes, but only in response to their native language. Thus, the findings suggest that ID speech might have an influence on the intersensory perception of fluent speech and shed further light on multisensory perceptual narrowing.</p>
Data from: Parallel processing in speech perception with local and global representations of linguistic context
<p>Speech processing is highly incremental. It is widely accepted that human listeners continuously use the linguistic context to anticipate upcoming concepts, words, and phonemes. However, previous evidence supports two seemingly contradictory models of how a predictive context is integrated with the bottom-up sensory input: Classic psycholinguistic paradigms suggest a two-stage process, in which acoustic input initially leads to local, context-independent representations, which are then quickly integrated with contextual constraints. This contrasts with the view that the brain constructs a single coherent, unified interpretation of the input, which fully integrates available information across representational hierarchies, and thus uses contextual constraints to modulate even the earliest sensory representations. To distinguish these hypotheses, we tested magnetoencephalography responses to continuous narrative speech for signatures of local and unified predictive models. Results provide evidence that listeners employ both types of models in parallel. Two local context models uniquely predict some part of early neural responses, one based on sublexical phoneme sequences, and one based on the phonemes in the current word alone; at the same time, even early responses to phonemes also reflect a unified model that incorporates sentence-level constraints to predict upcoming phonemes. Neural source localization places the anatomical origins of the different predictive models in nonidentical parts of the superior temporal lobes bilaterally, with the right hemisphere showing a relative preference for more local models. These results suggest that speech processing recruits both local and unified predictive models in parallel, reconciling previous disparate findings. Parallel models might make the perceptual system more robust, facilitate processing of unexpected inputs, and serve a function in language acquisition.</p>
Speech Perception and High Cognitive Demand
ClinicalTrials.gov study NCT04997577. IPD Sharing: YES. Countries: 1. Publications: 2.
Modulating Speech Perception With Current Stimulation
ClinicalTrials.gov study NCT05446350. IPD Sharing: NO. Countries: 1. Publications: 3.
Left and Right Hemisphere Contributions to Speech Perception
ClinicalTrials.gov study NCT04989309. IPD Sharing: YES. Countries: 1. Publications: 2.
Improving Perception of Speech in Noise in Children With Communication Disorders
ClinicalTrials.gov study NCT04473729. IPD Sharing: NO. Countries: 1. Publications: 3.
Study of Music and Speech Perception in New Cochlear Implanted Subjects Using or Not a Tonotopy Based Fitting
ClinicalTrials.gov study NCT04922619. IPD Sharing: NO. Countries: 1. Publications: 1.
Longitudinal Assessment of Early Premature Infant Skills for Audiovisual Speech Perception
ClinicalTrials.gov study NCT07245693. IPD Sharing: NO. Countries: 1. Publications: 13.
Hearing Aids Use in Elderly: Efficacy in Speech Perception and in Health-related Quality of Life
ClinicalTrials.gov study NCT04333043. IPD Sharing: UNDECIDED. Countries: 1. Publications: 2.
Subconscious Effects of Intraoperative Speech: Health Professionals' Perceptions
ClinicalTrials.gov study NCT07396727. IPD Sharing: NO. Countries: 1. Publications: 4.
The impact of cognitive ability on multitalker speech perception in neurodivergent individuals
Open the record for dataset details and reuse information.
Data from: Parallel processing in speech perception with local and global representations of linguistic context
Open the record for dataset details and reuse information.
ScienceDex guides
Understand access before you commit
These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.
Allen Brain Atlas
Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.
Annotated Behaviour and Observability Dataset (ABODe)
ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.
DANDI Archive for NWB datasets
DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.
International Brain Laboratory public data
The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.
OpenNeuro
OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.