Skip to content
Advertisement
AudioMultimodalText

nonverbalspeech38k

🎉 🎉 🎉 NonVerbalSpeech-38K: A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding The official repository for…

🎉 🎉 🎉 NonVerbalSpeech-38K: A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding The official repository for NonVerbalSpeech-38K (NVS-38K) dataset. ( News Demo Page ) The NVS-38K dataset is constructed from in-the-wild audio sources, such as movies, cartoons, and audiobooks (see Section: Source Distribution of NVS-38K). It contains a total of 38,718 samples spanning approximately 131 hours, annotated with 10 non-verbal categories (see Section: Special… See the full description on the dataset page:

Source: Hugging Face Hub (nonverbalspeech/nonverbalspeech38k). Metadata imported from the dataset’s Hub tags.

Advertisement