Skip to content
Advertisement
AudioMultimodalText

talkbank_4_stt

talkbank 4 stt

Dataset Card Dataset Description This dataset is a benchmark based on the TalkBank 1 corpus—a large multilingual repository of conversational speech that captures real-world, unstructured interactions. We use CA-Bank 2 , which focuses on phone conversations between adults, which include natural speech phenomena such as laughter, pauses, and interjections. To ensure the dataset is highly accurate and suitable for benchmarking conversational ASR systems, we employ… See the full description on the dataset page: 4 stt.

Source: Hugging Face Hub (diabolocom/talkbank_4_stt). Metadata imported from the dataset’s Hub tags.

Advertisement