Skip to content
Advertisement
AudioMultimodalText

wmt-human-all-TTS

WMT Human + TTS Audio WMT human evaluation data (zouharvi/wmt-human-all) extended with TTS-synthesised source audio, covering 49 language pairs. Used…

WMT Human + TTS Audio WMT human evaluation data (zouharvi/wmt-human-all) extended with TTS-synthesised source audio, covering 49 language pairs. Used as training data for SpeechCOMET. Part of the SpeechCOMET model family Paper: Why We Need Speech to Evaluate Speech Translation (Züfle et al., 2026) Code: github.com/MaikeZuefle/speechCOMET Dataset Each row contains a source sentence, a machine translation hypothesis, a human quality score, and TTS-synthesised source… See the full description on the dataset page:

Source: Hugging Face Hub (maikezu/wmt-human-all-TTS). Metadata imported from the dataset’s Hub tags.

Advertisement