Skip to content
Advertisement

Datasets

3,216 results

ImageMultimodalText

emova-sft-speech-231k

EMOVA-SFT-Speech-231K 🤗 EMOVA-Models 🤗 EMOVA-Datasets 🤗 EMOVA-Demo 📄 Paper 🌐 Project-Page 💻 Github 💻 EMOVA-Speech-Tokenizer-Github Overview…

100K–1M·Apache-2.0·Parquet
AudioMultimodalText

Habibi

A systematic and standardized benchmark for the multi-dialect Arabic zero-shot TTS task Paper: "Habibi: Laying the Open-Source Foundation…

10K–100K·Apache-2.0·Parquet
Audio

Urdu-Munch-Lina

Urdu-Munch-Lina Processed version of zuhri025/Urdu-Munch with LinaCodec encoding. Dataset Structure This dataset contains 2 batches of audio data…

10K–100K·MIT
Audio

AISHELL-3

AISHELL-3 is a large-scale and high-fidelity multi-speaker Mandarin speech corpus published by Beijing Shell Shell Technology Co.,Ltd. It…

10K–100K·Apache-2.0·Audio (folder)
Advertisement