Skip to content
Advertisement
Text

CapSpeech

CapSpeech DataSet used for the paper: CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech Please refer to CapSpeech repo…

CapSpeech DataSet used for the paper: CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech Please refer to CapSpeech repo for more details. Overview 🔥 CapSpeech is a new benchmark designed for style-captioned TTS (CapTTS) tasks, including style-captioned text-to-speech synthesis with sound effects (CapTTS-SE), accent-captioned TTS (AccCapTTS), emotion-captioned TTS (EmoCapTTS) and text-to-speech synthesis for chat agent (AgentTTS). CapSpeech… See the full description on the dataset page:

Source: Hugging Face Hub (OpenSound/CapSpeech). Metadata imported from the dataset’s Hub tags.

Advertisement