Skip to content
Advertisement

Question Answering

190 results

MultimodalTabularText

MapEval-API

MapEval-API MapEval-API is created using MapQaTor. Usage from datasets import load dataset Load dataset ds = load dataset("MapEval/MapEval-API",…

<1K·Apache-2.0·JSON
ImageMultimodalText

CHOICE

CHOICE: Benchmarking The Remote Sensing Capabilities of Large Vision-Language Models Abstract: The rapid advancement of Large Vision-Language Models…

1K–10K·MIT·Images (folder)
Image

EarthVLSet

EarthVL: A Progressive Earth Vision-Language Understanding and Generation Framework by Junjue Wang, Yanfei Zhong, Zihang Chen, Zhuo Zheng,…

>1B·CC-BY-NC-ND
MultimodalTabularText

M4LE

Introduction M4LE is a Multi-ability, Multi-range, Multi-task, bilingual benchmark for long-context evaluation. We categorize long-context…

10K–100K·MIT·JSON
Text

CodeMixBench

ℹ️Dataset Card for CodeMixBench EMNLP'25 CodeMixBench: Evaluating Code-Mixing Capabilities of LLMs Across 18 Languages Code-mixing is a linguistic…

10K–100K·Apache-2.0·CSV
Text

vietnamese_health_dataset

Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…

100K–1M·MIT·Parquet
MultimodalTabularText

planetarium

Dataset Card for Planetarium🪐 Planetarium🪐 is a dataset and benchmark for assessing LLMs in translating natural language descriptions…

100K–1M·CC-BY·Parquet
MultimodalTabularText

JMedBench

Maintainers Junfeng Jiang@Aizawa Lab: jiangjf (at) is.s.u-tokyo.ac.jp Jiahao Huang@Aizawa Lab: jiahao-huang (at) g.ecc.u-tokyo.ac.jp If you find any…

100K–1M·JSON
Text

WenYanWen_English_Parallel

Dataset Card for WenYanWen English Parallel Dataset Summary The WenYanWen English Parallel dataset is a multilingual parallel corpus…

1M–10M·MIT·Parquet
MultimodalTabularText

bhasha-sft

Bhasha SFT Bhasha SFT is a massive collection of multiple open sourced Supervised Fine-Tuning datasets for training Multilingual…

10M–100M·Mixed·Parquet
Advertisement