Skip to content
Advertisement

1M–10M

329 results

ImageMultimodalText

ImgCode-8.6M

MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning Repo: Paper: Introduction We introduce MathCoder-VL, a series…

1M–10M·Apache-2.0·Parquet
ImageMultimodalText

ChartVerse-SFT-1.8M

ChartVerse-SFT-1800K is an extended large-scale chart reasoning dataset with Chain-of-Thought (CoT) annotations, developed as part of the…

1M–10M·Apache-2.0·Parquet
ImageMultimodalText

EO-Data1.5M

🤖 EO-Data-1.5M A Large-Scale Interleaved Vision-Text-Action Dataset for Embodied AI The first large-scale interleaved embodied dataset emphasizing…

1M–10M·Apache-2.0·Parquet
ImageMultimodalText

ULVR_v2_clean

ULVR v2 clean Universal Latent Visual Reasoning training data, cleaned. 8 categories (subsets); each has train + validation…

1M–10M·Apache-2.0·Parquet
ImageMultimodalText

REVERSE

REVERSE Dataset Dataset for REVERSE (Reinforcing Evidence Verification and Search for Agentic Image Geolocation). This dataset supports training…

1M–10M·CC-BY·Parquet
ImageMultimodalText

Cauldron-JA

Dataset Card for The Cauldron-JA Dataset description The Cauldron-JA is a Vision Language Model dataset that translates 'The…

1M–10M·CC-BY·Parquet
AudioMultimodalText

cantonese-youtube-tts

Cantonese Audio TTS Dataset This dataset contains alvanlii/cantonese-radio, alvanlii/cantonese-youtube, plus a dataset of equal size. It is catered…

1M–10M·Parquet
Advertisement