Skip to content
Advertisement
Audio

Unsupervised Peoples Speech

Unsupervised Peoples Speech

Dataset Card for Unsupervised Peoples Speech Dataset Description Dataset Summary The Unsupervised Peoples Speech Dataset is a compilation of audiofiles extracted from Archive.org that is licensed for academic and commercial usage under CC-BY and CC-BY-SA licenses. It includes more than one million hours of audio with a diverse set of speakers. Point of Contact: MLCommons Datasets Discord Dataset Structure This dataset is a collection of audio… See the full description on the dataset page: peoples speech.

Source: Hugging Face Hub (MLCommons/unsupervised_peoples_speech). Metadata imported from the dataset’s Hub tags.

Advertisement