Skip to content
Advertisement

Open

2,922 results

Audio

VoxEval

VoxEval Github repository for paper: VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models Also check…

CC-BY
Audio

Marco_Longspeech

Marco-LongSpeech Dataset Marco-LongSpeech is a multi-task long speech understanding dataset containing 8 different speech understanding tasks…

10K–100K·Apache-2.0
Text

open-web-math

Keiran Paster , Marco Dos Santos , Zhangir Azerbayev, Jimmy Ba GitHub ArXiv PDF OpenWebMath is a dataset…

1M–10M·Parquet
Text

wmdp

Dataset Card for WMDP The Weapons of Mass Destruction Proxy (WMDP) benchmark is a dataset of multiple-choice questions…

1K–10K·MIT·Parquet
Text

Fineweb-Edu-Chinese-V2.1

Chinese Fineweb Edu Dataset V2.1 中文 English OpenCSG Community 👾github wechat Twitter 📖Technical Report The Chinese Fineweb Edu…

100M–1B·Apache-2.0·Parquet
AudioMultimodalText

PIAST

PIAST Dataset This repo is for downloading transcribed MIDI & and text data of the PIAST Dataset. The…

MIT
Text

WideSearch

WideSearch: Benchmarking Agentic Broad Info-Seeking Dataset Summary WideSearch is a benchmark designed to evaluate the capabilities of Large…

<1K·Custom / Research-only·JSON
MultimodalTextVideo

FLARE

FLARE: Full-Modality Long-Video Audiovisual Retrieval Benchmark with User-Simulated Queries 🤗 About This Benchmark This repository hosts the full…

100K–1M·CC-BY·JSON
Text

stsbenchmark-sts

STSBenchmark An MTEB dataset Massive Text Embedding Benchmark Semantic Textual Similarity Benchmark (STSbenchmark) dataset. Task category t2t Domains…

1K–10K·Custom / Research-only·JSON
Text

Argimi-Ardian-Finance-10k-text

The ArGiMI Ardian datasets : Text-only version The ArGiMi project is committed to open-source principles and data sharing.…

1M–10M·CC-BY·WebDataset
Text

BLiMP

BLiMP

10K–100K·CC-BY·Parquet
Text

hh-rlhf

Dataset Card for HH-RLHF Dataset Summary This repository provides access to two different kinds of data: Human preference…

100K–1M·MIT·JSON
Advertisement