Skip to content
Advertisement
AudioMultimodalText

A Multilingual Speech Dataset for SLU and Beyond

A Multilingual Speech Dataset for SLU and Beyond

Speech-MASSIVE Dataset Description Speech-MASSIVE is a multilingual Spoken Language Understanding (SLU) dataset comprising the speech counterpart for a portion of the MASSIVE textual corpus. Speech-MASSIVE covers 12 languages (Arabic, German, Spanish, French, Hungarian, Korean, Dutch, Polish, European Portuguese, Russian, Turkish, and Vietnamese) from different families and inherits from MASSIVE the annotations for the intent prediction and slot-filling tasks. MASSIVE… See the full description on the dataset page:

Source: Hugging Face Hub (FBK-MT/Speech-MASSIVE). Metadata imported from the dataset’s Hub tags.

Advertisement