Skip to content
Advertisement

Tabular

1,117 results

MultimodalTabularText

M4LE

Introduction M4LE is a Multi-ability, Multi-range, Multi-task, bilingual benchmark for long-context evaluation. We categorize long-context…

10K–100K·MIT·JSON
MultimodalTabularText

TED_talks

!NOTE Dataset origin: Context TED is devoted to spreading powerful ideas in just about any topic. These datasets…

10K–100K·CSV
MultimodalTabularText

planetarium

Dataset Card for Planetarium🪐 Planetarium🪐 is a dataset and benchmark for assessing LLMs in translating natural language descriptions…

100K–1M·CC-BY·Parquet
MultimodalTabularText

undl_zh2en_aligned

联合国数字图书馆的段落级中-英对齐平行语料 用我口胡的方法弄出来的平行语料,统计数据和拿argostranslate直接又跑了一份bleu score的结果已经丢论文里了,论文在写了在写了。应该拿这份去练机翻模型没问题,数据源是人写的。 bleu score…

10M–100M·MIT·Parquet
MultimodalTabularText

panlex-meanings

Dataset Card for panlex-meanings This is a dataset of words in several thousand languages, extracted from Dataset Details…

10M–100M·CC0·CSV
ImageMultimodalTabular

ChEBI-20-MM

ChEBI-20-MM Dataset Overview The ChEBI-20-MM is an extensive and multi-modal benchmark developed from the ChEBI-20 dataset. It is…

10K–100K·MIT·CSV
Advertisement