Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Dataset Card for Annotations of non English MLS This dataset consists in annotations of a the Non English subset of the Multilingual LibriSpeech…
Dataset Card for Annotations of non English MLS This dataset consists in annotations of a the Non English subset of the Multilingual LibriSpeech (MLS) dataset. MLS dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages – English, German, Dutch, Spanish, French, Italian, Portuguese, Polish. It includes about 44.5K hours of English and a total of about 6K hours for other languages.… See the full description on the dataset page:
Source: Hugging Face Hub (PHBJT/mls-annotated). Metadata imported from the dataset’s Hub tags.