Recomendation of OOPPEENN’s 56697375616C4E6F76656C5F44617461736574 I recommend OOPPEENN/56697375616C4E6F76656C5F44617461736574 for Japanese voice corpus, which is: Similar speech domain to this one (Japanese anime-style speech from Japanese Visual Novel), but Huge amounts of audio compared to this dataset (600 hours for this, 10,000 hours for Galgame Dataset!) This dataset contains about 50 games, and Galgame Dataset contains more than 500 games! Contains true transcripts of each… See the full description on the dataset page:
Source: Hugging Face Hub (litagin/moe-speech). Metadata imported from the dataset’s Hub tags.