Skip to content
Advertisement
AudioMultimodalText

FalAR

FalAR FalAR is a large-scale, speaker-annotated European Portuguese speech corpus built from recordings of parliamentary sessions of the Portuguese…

FalAR FalAR is a large-scale, speaker-annotated European Portuguese speech corpus built from recordings of parliamentary sessions of the Portuguese Parliament. The dataset contains aligned speech segments, reference transcripts, automatic transcripts, and speaker metadata. This release is intended to support research in automatic speech recognition (ASR), speaker-aware speech processing, and related studies on parliamentary speech in European Portuguese. Highlights… See the full description on the dataset page:

Source: Hugging Face Hub (inesc-id/FalAR). Metadata imported from the dataset’s Hub tags.

Advertisement