Skip to content
Advertisement
ImageMultimodalText

LLaVA-OneVision-1.5-Instruct-Data

LLaVA-OneVision-1.5 Instruction Data Paper Code 📌 Introduction This dataset, LLaVA-OneVision-1.5-Instruct, was collected and integrated during the…

LLaVA-OneVision-1.5 Instruction Data Paper Code 📌 Introduction This dataset, LLaVA-OneVision-1.5-Instruct, was collected and integrated during the development of LLaVA-OneVision-1.5. LLaVA-OneVision-1.5 is a novel family of Large Multimodal Models (LMMs) that achieve state-of-the-art performance with significantly reduced computational and financial costs. This meticulously curated 22M instruction dataset (LLaVA-OneVision-1.5-Instruct) is part of a… See the full description on the dataset page:

Source: Hugging Face Hub (mvp-lab/LLaVA-OneVision-1.5-Instruct-Data). Metadata imported from the dataset’s Hub tags.

Advertisement