Skip to content
Advertisement
AudioMultimodalText

C3T

C3T: Cross-modal Capabilities Conservation Test Dataset Description C3T (Cross-modal Capabilities Conservation Test) is a benchmark for assessing the…

C3T: Cross-modal Capabilities Conservation Test Dataset Description C3T (Cross-modal Capabilities Conservation Test) is a benchmark for assessing the performance of speech-aware language models. The benchmark utilizes textual tasks synthesized with a voice cloning text-to-speech model to verify if language understanding capabilities are preserved when the model is accessed via speech input. C3T quantifies the fairness of the model for different categories of speakers and… See the full description on the dataset page:

Source: Hugging Face Hub (amu-cai/C3T). Metadata imported from the dataset’s Hub tags.

Advertisement