Skip to content
Advertisement

Multimodal

2,368 results

MultimodalTabularText

fake_news

TODO: Add YAML tags here. Copy-paste the tags obtained with the online tagging app: annotations creators: - no-annotation…

10K–100K·Parquet
ImageMultimodalTabular

filtered-wit

Filtered WIT, an Image-Text Dataset. A reliable Dataset to run Image-Text models. You can find WIT, Wikipedia Image…

1M–10M·Parquet
MultimodalTabularText

wildguardmix

Dataset Card for WildGuardMix Disclaimer: The data includes examples that might be disturbing, harmful or upsetting. It includes…

10K–100K·ODC-BY·Parquet
MultimodalTabularText

SurvHTE-Bench

SurvHTE-Bench: A Benchmark for Heterogeneous Treatment Effect Estimation in Survival Analysis Paper: ICLR 2026 — SurvHTE-Bench: A Benchmark…

1M–10M·CC-BY·Parquet
MultimodalTabularText

ScaleEdit-12M

ScaleEdit-12M: Scaling Open-Source Image Editing Data Generation via Multi-Agent Framework       📌 Overview The largest open-source…

10M–100M·CC-BY-NC-SA·Parquet
MultimodalTabularText

AlgoTune

Website Paper Code How good are language models at coming up with new algorithms? To try to answer…

<1K·MIT·JSON
MultimodalTabularText

toxic-chat

Update 01/31/2024 We update the OpenAI Moderation API results for ToxicChat (0124) based on their updated moderation model…

10K–100K·CC-BY-NC·CSV
3D / Point CloudImageMultimodal

CADS-dataset

CADS: A Comprehensive Anatomical Dataset and Segmentation for Whole-Body Anatomy in Computed Tomography Overview CADS is a robust,…

10K–100K·Custom / Research-only·CSV
MultimodalTabularText

PopQA

Dataset Card for PopQA Dataset Summary PopQA is a large-scale open-domain question answering (QA) dataset, consisting of 14k…

10K–100K·CSV
Advertisement