Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
A Multilingual Speech Dataset for SLU and Beyond
Speech-MASSIVE Dataset Description Speech-MASSIVE is a multilingual Spoken Language Understanding (SLU) dataset comprising the speech counterpart for a portion of the MASSIVE textual corpus. Speech-MASSIVE covers 12 languages (Arabic, German, Spanish, French, Hungarian, Korean, Dutch, Polish, European Portuguese, Russian, Turkish, and Vietnamese) from different families and inherits from MASSIVE the annotations for the intent prediction and slot-filling tasks. MASSIVE… See the full description on the dataset page:
Source: Hugging Face Hub (FBK-MT/Speech-MASSIVE). Metadata imported from the dataset’s Hub tags.