Skip to content
Advertisement
AudioMultimodalText

Irodori-Ja-Spk2-10k

SynDataLab/Irodori-Ja-Spk2-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…

SynDataLab/Irodori-Ja-Spk2-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker family — see also Spk1, Spk2, Spk3, Spk4 repos. Speaker (Spk2): 30s female, narrator-style natural — 30代女性、ナレーター風の自然な声. How this speaker was made The voice identity for Spk2 was created in two stages: Stage 1 — voice anchor design. Using Irodori-TTS-500M-v2-VoiceDesign, several candidate audio samples were synthesized… See the full description on the dataset page:

Source: Hugging Face Hub (SynDataLab-JA-Refs/Irodori-Ja-Spk2-10k). Metadata imported from the dataset’s Hub tags.

Advertisement