Skip to content
Advertisement
AudioMultimodalText

Irodori-Ja-Spk1-10k

SynDataLab/Irodori-Ja-Spk1-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…

SynDataLab/Irodori-Ja-Spk1-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker family — see also Spk1, Spk2, Spk3, Spk4 repos. Speaker (Spk1): 30s male, calm conversational — 30代男性、落ち着いた自然な会話調. How this speaker was made The voice identity for Spk1 was created in two stages: Stage 1 — voice anchor design. Using Irodori-TTS-500M-v2-VoiceDesign, several candidate audio samples were synthesized from… See the full description on the dataset page:

Source: Hugging Face Hub (SynDataLab-JA-Refs/Irodori-Ja-Spk1-10k). Metadata imported from the dataset’s Hub tags.

Advertisement