10K–100K
748 results
NL2SH-ALFA
Dataset Card for NL2SH-ALFA This dataset is a collection of natural language (English) instructions and corresponding Bash commands…
ChEBI-20-MM
ChEBI-20-MM Dataset Overview The ChEBI-20-MM is an extensive and multi-modal benchmark developed from the ChEBI-20 dataset. It is…
Flickr30k
Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…
MultiMed-ST
MultiMed-ST: Large-scale Many-to-many Multilingual Medical Speech Translation 📘 EMNLP 2025 Khai Le-Duc , Tuyen Tran , Bach Phan…
IndonesianNMT
This dataset is used on the paper "Replicable Benchmarking of Neural Machine Translation (NMT) on Low-Resource Local Languages…
WMT20 – MultiLingual Quality Estimation (MLQE) Task2
WMT20 - MultiLingual Quality Estimation (MLQE) Task2
TinyStories-Multilingual
Novelist: TinyStories Multilingual Edition Dataset Summary The TinyStories Multilingual Edition is a high-fidelity synthetic dataset of short,…
multiclass-sentiment-analysis-dataset
multiclass-sentiment-analysis-dataset
Alexandria Multudialectal Arabic Conversational Dataset for Machine Translation
Alexandria Multudialectal Arabic Conversational Dataset for Machine Translation
BEAF
BEAF: Before-After Changes for Hallucination Evaluation BEAF is a benchmark for evaluating object hallucination in vision-language models using…
GQA-ru
GQA-ru This is a translated version of original GQA dataset and stored in format supported for lmms-eval pipeline.…