Skip to main content
Powered by ShareScore

Find research datasets worth reusing

Search datasets from major research repositories and use ShareScore to quickly assess how well each record supports discovery, access, and reuse.

5

datasets available to search

ShareScore release 0.7.1

Reset

Dataset results

5 results for “synthetic speech”

Learn how ShareScore rates datasets ↗
zenodo40/100

TIMIT-TTS: a Text-to-Speech Dataset for Synthetic Speech Detection

<p>With the rapid development of deep learning techniques, the generation and counterfeiting of multimedia material are becoming increasingly straightforward to perform. At the same time, sharing fake content on the web has become so simple that malicious users can create unpleasant situations with minimal effort. Also, forged media are getting more and more complex, with manipulated videos (e.g., deepfakes where both the visual and audio contents can be counterfeited) that are taking the scene over still images.<br> The multimedia forensic community has addressed the possible threats that this situation could imply by developing detectors that verify the authenticity of multimedia objects. However, the vast majority of these tools only analyze one modality at a time.<br> This was not a problem as long as still images were considered the most widely edited media, but now, since manipulated videos are becoming customary, performing monomodal analyses could be reductive. Nonetheless, there is a lack in the literature regarding multimodal detectors (systems that consider both audio and video components). This is due to the difficulty of developing them but also to the scarsity of datasets containing forged multimodal data to train and test the designed algorithms.</p> <p>In this paper we focus on the generation of an audio-visual deepfake dataset.<br> First, we present a general pipeline for synthesizing speech deepfake content from a given real or fake video, facilitating the creation of counterfeit multimodal material. The proposed method uses Text-to-Speech (TTS) and Dynamic Time Warping (DTW) techniques to achieve realistic speech tracks. Then, we use the pipeline to generate and release TIMIT-TTS, a synthetic speech dataset containing the most cutting-edge methods in the TTS field. This can be used as a standalone audio dataset, or combined with DeepfakeTIMIT and VidTIMIT video datasets to perform multimodal research. Finally, we present numerous experiments to benchmark the proposed dataset in both monomodal (i.e., audio) and multimodal (i.e., audio and video) conditions.<br> This highlights the need for multimodal forensic detectors and more multimodal deepfake data.</p> <ul> <li>For the initial version of TIMIT-TTS&nbsp;<strong>v1.0</strong> <ul> <li>Arxiv: https://arxiv.org/abs/2209.08000</li> <li>TIMIT-TTS Database v1.0: https://zenodo.org/record/6560159</li> </ul> </li> </ul>

opencc-by-4.0Sep 2022View details →
zenodo36/100

Synthetic Göttingen Sentence Test material created with a text-to-speech system

<p>The speech material of the German G&ouml;ttingen Sentence Test [1] was synthesized using a commercial text-to-speech system (Acapela Cloud Service). More details will be found in [2].</p> <p>Files:</p> <p>goesa_synth_female.zip<br> contains all 200 sentences with a synthetic female voice and the corresponding speech adjusted noise, which was generated by superimposing the speech material 30 times according to [3].</p> <p>goesa_synth_male.zip<br> contains all 200 sentences with a synthetic male voice and the corresponding speech adjusted noise, which was generated by superimposing the speech material 30 times according to [3].</p> <p>&nbsp;</p>

opencc-by-nc-4.0May 2022View details →
zenodo36/100

ODSS: An Open Dataset of Synthetic Speech

<p>ODSS is a multilingual, multispeaker dataset of synthetic and natural speech, designed to foster research and benchmarking of novel studies on synthetic speech detection.&nbsp;</p> <p>ODSS comprises audio utterances generated&nbsp;from text&nbsp;by state-of-the-art synthesis methods, paired with their corresponding natural counterparts. The synthetic audio data includes several languages, with an equal representation of genders.</p> <p>Natural and synthetic speech audio files within ODSS are released under the CC-BY-SA 4.0 license:&nbsp;Usage, extension and redistribution by the research community are strongly encouraged.</p>

openSep 2023View details →
zenodo24/100

Synthetic Speech Dataset

<p>This contains the speech data synthesized for the following paper:</p> <p>&quot;Using Speech Synthesis to Train End-to-End Spoken Language Understanding Models&quot; by Loren Lugosch, Brett Meyer, Derek Nowrouzezahrai, and Mirco Ravanelli</p>

opencc-by-4.0Oct 2019View details →
zenodo24/100

Synthetic phrase-based test material created with a text-to-speech system

<p>New speech material consisting of phrases was synthesized using a commercial text-to-speech system (Acapela Cloud Service). More details will be found in [1].</p> <p>Files:</p> <p>syntheticPhrases.zip<br>contains all 772 phrases with a synthetic female voice and the corresponding speech adjusted noise, which was generated by superimposing the speech material 30 times according to [2].</p>

opencc-by-nc-4.0Oct 2023View details →

ScienceDex guides

Understand access before you commit

These curated guides explain access requirements, typical timelines, costs, and reuse considerations for widely used research datasets.

Compare curated datasets

Allen Brain Atlas

Allen Brain Atlas is an Allen Institute collection of brain map atlases, datasets, APIs, and analysis tools covering mouse, human, and non-human primate brain resources.

allen-brain-atlas
neuroscienceopenDocumentation, web resources, and API references are available online.
Last verified 2026-04-30Open record

Annotated Behaviour and Observability Dataset (ABODe)

ABODe is a University of Edinburgh DataShare dataset for behavior classification in group-housed mice using home-cage video, identities, bounding boxes, ground-plate positions, and annotator labels.

abode-home-cage
behavioral-neuroscienceopenThe DataShare record exposes download links for annotations, documentation, license text, and the zipped per-snippet data directory.
Last verified 2026-04-30Open record

DANDI Archive for NWB datasets

DANDI is a BRAIN Initiative archive for publishing and sharing neurophysiology data, including electrophysiology, optophysiology, and behavioral data packaged as NWB and related standards.

dandi-nwb
electrophysiologyopenPublished Dandiset metadata and archive endpoints are available through the production DANDI API.
Last verified 2026-04-30Open record

International Brain Laboratory public data

The International Brain Laboratory public data releases expose standardized mouse decision-making experiments, including Neuropixels recordings, widefield calcium imaging, behavior, and session metadata accessed through the ONE API.

ibl
behavioral-neuroscienceopenPublic sessions can be searched and loaded from the IBL public data server through ONE.
Last verified 2026-04-29Open record

OpenNeuro

OpenNeuro is a free, open platform for sharing neuroimaging datasets, with public search, dataset pages, and download paths for web, S3, DataLad, and the OpenNeuro CLI.

openneuro
neuroscienceopenPublished datasets are available on demand over the internet.
Last verified 2026-04-29Open record