Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Dataset Summary We present a speech corpus for Classical Arabic Text-to-Speech (ClArTTS) to support the development of end-to-end TTS systems for Arabic. The speech is extracted from a LibriVox audiobook, which is then processed, segmented, and manually transcribed and annotated. The final ClArTTS corpus contains about 12 hours of speech from a single male speaker sampled at 40100 kHz. Dataset Structure A typical data point comprises the name of the audio file, called… See the full description on the dataset page:
Source: Hugging Face Hub (MBZUAI/ClArTTS). Metadata imported from the dataset’s Hub tags.