Skip to content
Advertisement
AudioMultimodalText

vaja-thai

Vaja-Thai (วาจา) — Combined Thai TTS Dataset A unified, quality-filtered Thai speech dataset combining multiple sources for Text-to-Speech (TTS)…

Vaja-Thai (วาจา) — Combined Thai TTS Dataset A unified, quality-filtered Thai speech dataset combining multiple sources for Text-to-Speech (TTS) research. All audio is resampled to 24 kHz WAV format. Dataset Summary Metric Value Total samples 289,916 Total hours 554.6h Sampling rate 24,000 Hz Format WAV 16-bit PCM Language Thai (ภาษาไทย) Sources Source Samples Hours License Description tsync2 1,823 3.7h CC-BY-NC-SA-3.0 NECTEC… See the full description on the dataset page:

Source: Hugging Face Hub (dubbing-ai/vaja-thai). Metadata imported from the dataset’s Hub tags.

Advertisement