Audio2Tool — Spoken Tool-Calling Benchmark
Audio2Tool — Spoken Tool-Calling Benchmark
2,175 results
Audio2Tool — Spoken Tool-Calling Benchmark
Suno AI Music Dataset — Multi-Genre Curated
Dataset Card for LibriTTS LibriTTS is a multi-speaker English corpus of approximately 585 hours of read English speech…
PROCESS-2: Speech Dataset for Early Cognitive Impairment Detection
FalAR FalAR is a large-scale, speaker-annotated European Portuguese speech corpus built from recordings of parliamentary sessions of the…
The dataset is available under the terms of the Creative Commons Attribution Non-Commercial license. K. J. Piczak. ESC:…
Russian voices for train AI. ♀ Male and ♂ Female Male voices - 497 pcs. Female voices -…
EuroSpeech 24 kHz Dataset Dataset Description EuroSpeech is a large-scale multilingual speech corpus containing high-quality aligned parliamentary…
Multilingual MFA-Aligned Speech Dataset
Reazon Speech v2 DENOISED Same Speaker Pair prj-beatrice/reazon-speech-v2-denoised-titanet-embeddings のコサイン類似度の総和がなるべく大きくなるように…
AppTek Call-Center Dialogues
Tarteel AI - EveryAyah Dataset
AudioMarathon AudioMarathon is a long-context audio benchmark for evaluating multimodal LLMs on speech, music, environmental audio, and meetings.…