Skip to content
Advertisement

Datasets

911 results

MultimodalTabularText

TVBench

Lost in Time: A New Temporal Benchmark for Video LLMs Daniel Cores , Michael Dorkenwald , Manuel Mucientes,…

1K–10K·CC-BY·JSON
MultimodalTabularText

MVTamperBenchStart

MVTamperBench Dataset Overview MVTamperBenchStart is a robust benchmark designed to evaluate Vision-Language Models (VLMs) against adversarial video…

10K–100K·MIT·JSON
MultimodalTabularText

CG-Bench

CG-Bench Project Website: Repository: (includes running code) Summary We introduce CG-Bench, a groundbreaking benchmark for clue-grounded question…

10K–100K·MIT·JSON
ImageMultimodalTabular

SLAKE

Dataset Info: SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering ISBI 2021 oral Project Page: click…

10K–100K·CC-BY·JSON
MultimodalTabularText

mls-annotated

Dataset Card for Annotations of non English MLS This dataset consists in annotations of a the Non English…

1M–10M·CC-BY·Parquet
MultimodalTabularText

UniST

UniST This dataset contains UniST codec-token training data exported from local metadata and codec results. We train UniSS…

10M–100M·CC-BY-NC
AudioMultimodalTabular

YodasSpeakerPool

Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique…

1K–10K·CC-BY
Advertisement