Skip to content
Advertisement
AudioMultimodalText

Persian Farsi Speech Dataset

Persian Farsi Speech Dataset

Persian (Farsi) TTS Dataset 🗂️ Dataset Description This dataset is a Persian (Farsi) text-to-speech (TTS) corpus built by concatenating and denoising multiple existing Farsi datasets.It is intended for training and evaluation of speech synthesis (TTS) models in Persian. Since the basic datasets were contaminated with unintelligible audio, I used dnsmos to keep only clean audio (mos ovr = 3.0, same value as for the Emilia dataset). The dataset contains two main columns:… See the full description on the dataset page:

Source: Hugging Face Hub (Thomcles/Persian-Farsi-Speech). Metadata imported from the dataset’s Hub tags.

Advertisement