Skip to content
Advertisement

Parliament Parliament is an OpenFormosa Traditional Chinese speech dataset derived from the zh tw split of disco-eth/WorldSpeech. It contains audio clips and human transcripts from Taiwan Legislative Yuan IVOD parliamentary proceedings. This release keeps only rows that passed the Taiwan-OmniData / FineWeb2-style text filtering pipeline. Audio is preserved from the upstream dataset and cast as a Hugging Face Audio(sampling rate=24000) feature. Dataset Summary… See the full description on the dataset page:

Source: Hugging Face Hub (OpenFormosa/parliament). Metadata imported from the dataset’s Hub tags.

Advertisement