Skip to content
Advertisement

Parquet

1,501 results

AudioMultimodalText

libritts_r

Dataset Card for LibriTTS-R LibriTTS-R 1 is a sound quality improved version of the LibriTTS corpus ( which…

100K–1M·CC-BY·Parquet
AudioMultimodalText

Galgame-VisualNovel-Reupload

Galgame VisualNovel Reupload This repository is a reupload of the visual novel dataset OOPPEENN/56697375616C4E6F76656C5F44617461736574. The goal of…

1M–10M·Custom / Research-only·Parquet
AudioMultimodalText

legco-speech

香港立法會會議語音數據集 本數據集係由香港立法會會議製成嘅大規模語音數據集。原始錄音總時長 22,196 個鐘,切分語音後總時長 20,471 個鐘。數據集分兩個子集,raw同segmented,分別為原始錄音同VAD識別切分後嘅語音。 數據集製作流程 先去香港特別行政區立法會…

1M–10M·CC0·Parquet
AudioMultimodalText

opendata-iisys-hui

HUI-Audio-Corpus-German Dataset Overview The HUI-Audio-Corpus-German is a high-quality Text-To-Speech (TTS) dataset developed by researchers at the…

10K–100K·MIT·Parquet
AudioMultimodalText

voicebench

License The dataset is available under the Apache 2.0 license. Citation If you use the VoiceBench dataset in…

10K–100K·Apache-2.0·Parquet
AudioMultimodalText

floras

FLORAS FLORAS is a 50-language benchmark For LOng-form Recognition And Summarization of spoken language. The goal of FLORAS…

10K–100K·CC-BY·Parquet
AudioMultimodalText

urbansound8K

(card and dataset copied from This dataset contains 8732 labeled sound excerpts (<=4s) of urban sounds from 10…

1K–10K·CC-BY-NC·Parquet
Audio

JamendoMaxCaps

JamendoMaxCaps Dataset JamendoMaxCaps is a large-scale dataset of over 362,000 instrumental tracks sourced from the Jamendo platform. It…

100K–1M·CC-BY-SA·Parquet
Advertisement