Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique…
Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique speakers (3.4K Chinese, 4.2K English). Derived from the Emilia-YODAS corpus, each sample is annotated by Gemini 2.5 Flash for its vocal characteristics and audio quality. Dataset Features Audio Specs: 4–15 second WAV samples of clean speech. Rich Metadata: Includes ASR… See the full description on the dataset page:
Source: Hugging Face Hub (HuHaiYang/YodasSpeakerPool). Metadata imported from the dataset’s Hub tags.