Skip to content
Advertisement
AudioMultimodalTabularText

Simple Voice Questions

Simple Voice Questions

Simple Voice Questions Simple Voice Questions (SVQ) is a set of short audio questions recorded in 26 locales across 17 languages under multiple audio conditions. It serves as a core evaluation componenet for Massive Sound Embedding Benchmark (MSEB). Technical Specifications Feature Details Locales 26 Languages 17 Total Speakers ~700 (Capped at 250 recordings per speaker) Audio Conditions Clean, Background Speech, Media, Traffic Noise Gender… See the full description on the dataset page:

Source: Hugging Face Hub (google/svq). Metadata imported from the dataset’s Hub tags.

Advertisement