Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Vaja-Thai (วาจา) — Combined Thai TTS Dataset A unified, quality-filtered Thai speech dataset combining multiple sources for Text-to-Speech (TTS)…
Vaja-Thai (วาจา) — Combined Thai TTS Dataset A unified, quality-filtered Thai speech dataset combining multiple sources for Text-to-Speech (TTS) research. All audio is resampled to 24 kHz WAV format. Dataset Summary Metric Value Total samples 289,916 Total hours 554.6h Sampling rate 24,000 Hz Format WAV 16-bit PCM Language Thai (ภาษาไทย) Sources Source Samples Hours License Description tsync2 1,823 3.7h CC-BY-NC-SA-3.0 NECTEC… See the full description on the dataset page:
Source: Hugging Face Hub (dubbing-ai/vaja-thai). Metadata imported from the dataset’s Hub tags.