Skip to content
Advertisement

Multilingual

575 results

AudioMultimodalText

dowis

Do What I Say (DOWIS): A Spoken Prompt Dataset for Instruction-Following NEW DOWIS now also contains spoken and…

1K–10K·CC-BY·Parquet
MultimodalTabularText

mls-annotated

Dataset Card for Annotations of non English MLS This dataset consists in annotations of a the Non English…

1M–10M·CC-BY·Parquet
AudioMultimodalText

OpenDialog

OpenDialog OpenDialog is a 6.8k hours spoken dialogue dataset, introduced in the paper ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation…

100K–1M·CC-BY-NC·WebDataset
MultimodalTabularText

UniST

UniST This dataset contains UniST codec-token training data exported from local metadata and codec results. We train UniSS…

10M–100M·CC-BY-NC
AudioMultimodalTabular

YodasSpeakerPool

Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique…

1K–10K·CC-BY
Advertisement