Skip to content
Advertisement

Multimodal

2,368 results

MultimodalTabularText

osm-polygon-selection

osm-polygon-selection dataset A curated set of OpenStreetMap polygons from 310 geographic units — sovereign countries plus sub-country regions…

10M–100M·ODbL·Parquet
MultimodalTabularText

Mana-TTS

ManaTTS-Persian-Speech-Dataset ManaTTS is the largest publicly available single-speaker Persian corpus, comprising over 114 hours of high-quality…

10K–100K·CC0·Parquet
MultimodalTabularText

medicine-tasks

Adapting LLMs to Domains via Continual Pre-Training (ICLR 2024) This repo contains the evaluation datasets for our paper…

1K–10K·JSON
ImageMultimodalTabular

Core-S2L1C

Core-S2L1C Contains a global coverage of Sentinel-2 (Level 1C) patches, each of size 1,068 x 1,068 pixels. Source…

1M–10M·CC-BY-SA·Parquet
MultimodalTabularText

reward-bench-2

Code Leaderboard Results Paper RewardBench 2 Evaluation Dataset Card The RewardBench 2 evaluation dataset is the new version…

1K–10K·ODC-BY·Parquet
MultimodalTabularText

MVTamperBench

MVTamperBench Dataset Overview MVTamperBench is a robust benchmark designed to evaluate Vision-Language Models (VLMs) against adversarial video…

10K–100K·MIT
MultimodalTabularText

dojo_forex_kline

Languages: 简体中文 · English dojo forex kline — FX Daily Bars Overview Daily OHLC and amplitude for major…

1K–10K·Apache-2.0·Parquet
MultimodalTabularText

mumospee_emilia

Dataset Summary This dataset is a modified version of the Emilia corpus, converted into parquet format to facilitate…

10M–100M·CC-BY-NC·Parquet
MultimodalTabularText

AIDev

AIDev: Studying AI Coding Agents on GitHub (The Rise of AI Teammates in Software Engineering 3.0) Papers: The…

1M–10M·CC-BY·Parquet
Advertisement