Skip to content
Advertisement

Text

2,175 results

AudioMultimodalText

libritts

Dataset Card for LibriTTS LibriTTS is a multi-speaker English corpus of approximately 585 hours of read English speech…

100K–1M·CC-BY·Parquet
AudioMultimodalText

FalAR

FalAR FalAR is a large-scale, speaker-annotated European Portuguese speech corpus built from recordings of parliamentary sessions of the…

100K–1M·CC-BY·Parquet
AudioMultimodalText

esc50

The dataset is available under the terms of the Creative Commons Attribution Non-Commercial license. K. J. Piczak. ESC:…

1K–10K·Parquet
AudioMultimodalText

EuroSpeech-24kHz

EuroSpeech 24 kHz Dataset Dataset Description EuroSpeech is a large-scale multilingual speech corpus containing high-quality aligned parliamentary…

10M–100M·MIT·Parquet
AudioMultimodalText

AudioMarathon

AudioMarathon AudioMarathon is a long-context audio benchmark for evaluating multimodal LLMs on speech, music, environmental audio, and meetings.…

<1K·Custom / Research-only·CSV
Advertisement