libritts_r
Dataset Card for LibriTTS-R LibriTTS-R 1 is a sound quality improved version of the LibriTTS corpus ( which…
1,501 results
Dataset Card for LibriTTS-R LibriTTS-R 1 is a sound quality improved version of the LibriTTS corpus ( which…
Dataset Card for Filtred and CML-TTS This dataset is a filtred version of a CML-TTS 1 . CML-TTS…
FreeSound.org LAION-640k Dataset (16 KHz)
Galgame VisualNovel Reupload This repository is a reupload of the visual novel dataset OOPPEENN/56697375616C4E6F76656C5F44617461736574. The goal of…
香港立法會會議語音數據集 本數據集係由香港立法會會議製成嘅大規模語音數據集。原始錄音總時長 22,196 個鐘,切分語音後總時長 20,471 個鐘。數據集分兩個子集,raw同segmented,分別為原始錄音同VAD識別切分後嘅語音。 數據集製作流程 先去香港特別行政區立法會…
HUI-Audio-Corpus-German Dataset Overview The HUI-Audio-Corpus-German is a high-quality Text-To-Speech (TTS) dataset developed by researchers at the…
common-voice-asr-clean Filtered ASR dataset. Samples with <3 words, repetitive tokens, or chat token leaks removed.
قاعدة بيانات المعلم القرآنية هذه ال dataset هي جزء من مشروع الملم الرقرآني: quran-muaalem وهي تهدف لكشف أخاطاء…
Speech Recognition Alignment Dataset
License The dataset is available under the Apache 2.0 license. Citation If you use the VoiceBench dataset in…
Daily-Omni (QA repackaged for lmms-eval)
FLORAS FLORAS is a 50-language benchmark For LOng-form Recognition And Summarization of spoken language. The goal of FLORAS…
(card and dataset copied from This dataset contains 8732 labeled sound excerpts (<=4s) of urban sounds from 10…
Multilingual MFA-Aligned Speech Dataset
JamendoMaxCaps Dataset JamendoMaxCaps is a large-scale dataset of over 362,000 instrumental tracks sourced from the Jamendo platform. It…