Skip to content
Advertisement

Audio

514 results

AudioMultimodalText

opendata-iisys-hui

HUI-Audio-Corpus-German Dataset Overview The HUI-Audio-Corpus-German is a high-quality Text-To-Speech (TTS) dataset developed by researchers at the…

10K–100K·MIT·Parquet
AudioMultimodalText

voicebench

License The dataset is available under the Apache 2.0 license. Citation If you use the VoiceBench dataset in…

10K–100K·Apache-2.0·Parquet
AudioMultimodalText

UrbanSound8K

UrbanSound8K This is an audio classification dataset for Sound Event Classification. Classes = 10 , Split = Ten-Fold…

10K–100K·MIT·CSV
AudioMultimodalText

mls_sidon

MLS-Sidon Overview This dataset is a cleansed version of Multilingual LibriSpeech (MLS) with Sidon speech restoration mode for…

10M–100M·CC-BY·WebDataset
AudioMultimodalText

floras

FLORAS FLORAS is a 50-language benchmark For LOng-form Recognition And Summarization of spoken language. The goal of FLORAS…

10K–100K·CC-BY·Parquet
AudioMultimodalText

urbansound8K

(card and dataset copied from This dataset contains 8732 labeled sound excerpts (<=4s) of urban sounds from 10…

1K–10K·CC-BY-NC·Parquet
AudioMultimodalText

cineaudiosynth

CineAudioSynth Synthetic cinematic audio for source separation. 453 scenes, ~22.8 h, 48 kHz / 16-bit / stereo WAV.…

1K–10K·CC-BY-NC-SA·Audio (folder)
AudioMultimodalText

mu-bench

Dataset Card for μ-Bench (Leaderboard Code) μ-Bench is a multilingual transcription benchmark built from real customer-service phone conversations.…

1K–10K·CC-BY-NC·Audio (folder)
Advertisement