Skip to content
Advertisement
AudioMultimodalText

Robots Backtranslated

Robots Backtranslated

AfVoices-Translated (Bambara-English) This is a Bambara speech translation dataset, which is built on the African Next Voices (AfVoices) Bambara ASR corpus. It provides English translations for the human-corrected subset of the original collection, creating a parallel corpus for Bambara-English machine translation and speech-to-text tasks. Methodology We machine-translated the human-validated transcriptions from AfVoices using the Oolel-translator repository. Inference… See the full description on the dataset page:

Source: Hugging Face Hub (soynade-research/Bambara-Speech-Translation-Data). Metadata imported from the dataset’s Hub tags.

Advertisement