Skip to content
Advertisement
Text

SIFT-50M

SIFT-50M

Dataset Card for SIFT-50M SIFT-50M (Speech Instruction Fine-Tuning) is a 50-million-example dataset designed for instruction fine-tuning and pre-training of speech-text large language models (LLMs). It is built from publicly available speech corpora containing a total of 14K hours of speech and leverages LLMs and off-the-shelf expert models. The dataset spans five languages, covering diverse aspects of speech understanding and controllable speech generation instructions. SIFT-50M… See the full description on the dataset page:

Source: Hugging Face Hub (amazon-agi/SIFT-50M). Metadata imported from the dataset’s Hub tags.

Advertisement