Skip to content
Advertisement
AudioMultimodalTabularText

YodasSpeakerPool

Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique…

Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique speakers (3.4K Chinese, 4.2K English). Derived from the Emilia-YODAS corpus, each sample is annotated by Gemini 2.5 Flash for its vocal characteristics and audio quality. Dataset Features Audio Specs: 4–15 second WAV samples of clean speech. Rich Metadata: Includes ASR… See the full description on the dataset page:

Source: Hugging Face Hub (HuHaiYang/YodasSpeakerPool). Metadata imported from the dataset’s Hub tags.

Advertisement