esc50
The dataset is available under the terms of the Creative Commons Attribution Non-Commercial license. K. J. Piczak. ESC:…
2,147 results
The dataset is available under the terms of the Creative Commons Attribution Non-Commercial license. K. J. Piczak. ESC:…
Russian voices for train AI. ♀ Male and ♂ Female Male voices - 497 pcs. Female voices -…
EuroSpeech 24 kHz Dataset Dataset Description EuroSpeech is a large-scale multilingual speech corpus containing high-quality aligned parliamentary…
Multilingual MFA-Aligned Speech Dataset
Reazon Speech v2 DENOISED Same Speaker Pair prj-beatrice/reazon-speech-v2-denoised-titanet-embeddings のコサイン類似度の総和がなるべく大きくなるように…
AppTek Call-Center Dialogues
Tarteel AI - EveryAyah Dataset
AudioMarathon AudioMarathon is a long-context audio benchmark for evaluating multimodal LLMs on speech, music, environmental audio, and meetings.…
Afrivoice Swahili Agriculture Subset
AudioSet AudioSet 1 is a large-scale dataset comprising approximately 2 million 10-second YouTube audio clips, categorised into 527…
ASMR-Archive-Processed-SFW
AVSpeech Video + Audio A restructured subset of the AVSpeech dataset with separated media streams and derived identifiers.…
Dataset Card for Meow-10K Meow-10K is a high-fidelity, synchronized quad-modal dataset comprising 10,000 feline samples. It is the…
Vietnamese Restaurant Order Speech
Dataset Card for LibriTTS-R LibriTTS-R 1 is a sound quality improved version of the LibriTTS corpus ( which…