Text
2,175 results
CLUE: Chinese Language Understanding Evaluation benchmark
CLUE: Chinese Language Understanding Evaluation benchmark
bias_in_bios
Bias in Bios Bias in Bios was created by (De-Artega et al., 2019) and published under the MIT…
KMMLU
KMMLU (Korean-MMLU) We propose KMMLU, a new Korean benchmark with 35,030 expert-level multiple-choice questions across 45 subjects ranging…
smollm-chunked
FAISS Indices and Chunked Datasets for SmolLM and SmolLM2 corpora This repository contains part of the FAISS indices…
GPT-NL_Public_Corpus
Dataset Card GPT-NL Public Corpus The GPT-NL Public Corpus is the largest permissively licensed Dutch-language resource available for…
ragbench
RAGBench Dataset Overview RAGBEnch is a large-scale RAG benchmark dataset of 100k RAG examples. It covers five unique…
Core-S2L2A
Core-S2L2A Contains a global coverage of Sentinel-2 (Level 2A) patches, each of size 1,068 x 1,068 pixels. Source…
fake_news
TODO: Add YAML tags here. Copy-paste the tags obtained with the online tagging app: annotations creators: - no-annotation…
filtered-wit
Filtered WIT, an Image-Text Dataset. A reliable Dataset to run Image-Text models. You can find WIT, Wikipedia Image…
ML-ArXiv-Papers
This dataset contains the subset of ArXiv papers with the "cs.LG" tag to indicate the paper is about…
indoor-safety-hazard-detection-and-work-zone-monitoring
Indoor Safety Hazard Detection & Work-Zone Monitoring Generated by datapack-import.ts This dataset mirrors public data-pack render outputs from…
dataset-the-stack-v2-dedup-sub
The Stack v2 Subset with File Contents (Python, Java, JavaScript, C, C++) TempestTeam/dataset-the-stack-v2-dedup-sub Dataset Summary This dataset is…
PersonaMem v2, Implicit Persona, LLM Personalization
PersonaMem v2, Implicit Persona, LLM Personalization
wildguardmix
Dataset Card for WildGuardMix Disclaimer: The data includes examples that might be disturbing, harmful or upsetting. It includes…
coding agent traces – security audits
coding agent traces - security audits