Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Mostafa Mahmoud Arabic Speech Dataset Dataset Summary The Mostafa Mahmoud Arabic Speech Dataset is a large-scale Arabic speech corpus containing approximately 187 hours of speech recordings and corresponding transcripts derived from publicly available lectures, interviews, television appearances, and talks by Dr. Mostafa Mahmoud. The dataset was created to support Arabic speech technology research and development, including: Automatic Speech Recognition (ASR)… See the full description on the dataset page:
Source: Hugging Face Hub (oddadmix/arabic-audio-collection-mostafa-mahmoud). Metadata imported from the dataset’s Hub tags.