Datasets
3,216 results
finetuned-hindi-punjabi-denoised
Multilingual Speaker Diarization Dataset This dataset contains synthetic multilingual speaker diarization data with Hindi, English, and Punjabi audio…
open-asr-leaderboard
ESB Test Sets: Parquet & Sorted This dataset takes the open-asr-leaderboard/datasets-test-only data and sorts each split by audio…
FreeSound.org LAION-640k Dataset
FreeSound.org LAION-640k Dataset
JavisInst-Omni
JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation HomePage Paper GitHub TL;DR We introduce JavisGPT, a…
AVQA JSONL (Audio Multiple-Choice QA)
AVQA JSONL (Audio Multiple-Choice QA)
NatureLM-audio-training
NatureLM-audio-training
synthetic-asr-zh
Synthetic ASR data — zh Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream…
malaysian-youtube
Malaysian Youtube Malaysian and Singaporean youtube channels, total up to 60k audio files with total 18.7k hours. URLs…
AIR-Bench-Dataset
AIR-Bench Arxiv: is the AIR-Bench dataset download page.AIR-Bench encompasses two dimensions: foundation and chat benchmarks. The former consists…
TAVGBench_1m
Installation Download this repo to a local folder, and unzip these .zip files under the TAVGBench 1m/data/. Then,…
Voice “Cloning” is Style Transfer — Audio Dataset
Voice "Cloning" is Style Transfer — Audio Dataset