Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
AfVoices-Translated (Bambara-English) This is a Bambara speech translation dataset, which is built on the African Next Voices (AfVoices) Bambara ASR corpus. It provides English translations for the human-corrected subset of the original collection, creating a parallel corpus for Bambara-English machine translation and speech-to-text tasks. Methodology We machine-translated the human-validated transcriptions from AfVoices using the Oolel-translator repository. Inference… See the full description on the dataset page:
Source: Hugging Face Hub (soynade-research/Bambara-Speech-Translation-Data). Metadata imported from the dataset’s Hub tags.