TAVGBench_1m
Installation Download this repo to a local folder, and unzip these .zip files under the TAVGBench 1m/data/. Then,…
464 results
Installation Download this repo to a local folder, and unzip these .zip files under the TAVGBench 1m/data/. Then,…
Voice "Cloning" is Style Transfer — Audio Dataset
MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations Humans rely on multisensory integration to perceive…
MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
Marco-LongSpeech Dataset Marco-LongSpeech is a multi-task long speech understanding dataset containing 8 different speech understanding tasks…
PIAST Dataset This repo is for downloading transcribed MIDI & and text data of the PIAST Dataset. The…
Music Arena Dataset This is the official dataset from Music Arena, an open platform for evaluating text-to-music (TTM)…
X-Voice Training Dataset Overview The X-Voice training dataset is a large-scale multilingual speech corpus curated for high-performance speech…
Russian Podcasts (unlabeled)
Deep Dialogue (Orpheus TTS)
Multitask-National-Speech-Corpus (MNSC v1) is derived from IMDA's NSC Corpus. MNSC is a multitask speech understanding dataset derived and…