Skip to content
Advertisement

Multimodal

2,368 results

AudioMultimodalText

AISHELL-4

AISHELL-4 Identifier: SLR111 Summary: A Free Mandarin Multi-channel Meeting Speech Corpus, provided by Beijing Shell Shell Technology Co.,Ltd…

Apache-2.0
AudioMultimodalText

BigAudioDataset

AstraMindAI/BigAudioDataset Dataset Description AstraMindAI/BigAudioDataset is a large-scale, multilingual dataset designed for a wide range of audio…

1M–10M·Apache-2.0·Arrow
AudioMultimodalText

commonvoice22_sidon

CV22-Sidon Overview This dataset hosts a release of Mozilla Common Voice 22 restored with the Sidon speech restoration…

10M–100M·CC0·WebDataset
AudioMultimodalText

openwhisper

Streaming ASR Dataset This dataset is designed for training real-time (streaming) ASR models, with a focus on handling…

100K–1M·MIT·Parquet
Advertisement