Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Deep Confessions Podcast Arabic Speech Dataset
Deep Confessions Podcast Arabic Speech Dataset Dataset Summary The Deep Confessions Podcast Arabic Speech Dataset is a large-scale, first-of-its-kind Arabic speech corpus containing approximately 175 hours of speech recordings and corresponding transcripts. What distinguishes this dataset as a pioneering resource in Arabic language technology is its comprehensive inclusion of rich non-verbal transcriptions. Alongside the spoken Arabic text, the transcripts… See the full description on the dataset page:
Source: Hugging Face Hub (oddadmix/arabic-audio-collection-tunisian-deep-confessions). Metadata imported from the dataset’s Hub tags.