Skip to content
Advertisement
AudioMultimodalText

InstructTTSEval

InstructTTSEval InstructTTSEval is a comprehensive benchmark designed to evaluate Text-to-Speech (TTS) systems' ability to follow complex…

InstructTTSEval InstructTTSEval is a comprehensive benchmark designed to evaluate Text-to-Speech (TTS) systems’ ability to follow complex natural-language style instructions. The dataset provides a hierarchical evaluation framework with three progressively challenging tasks that test both low-level acoustic control and high-level style generalization capabilities. Github Repository: Paper: InstructTTSEval: Benchmarking Complex… See the full description on the dataset page:

Source: Hugging Face Hub (CaasiHUANG/InstructTTSEval). Metadata imported from the dataset’s Hub tags.

Advertisement