<1K
439 results
MIT_environmental_impulse_responses
MIT Environmental Impulse Response Dataset The audio recordings in this dataset are originally created by the Computational Audition…
french-tts-conversational-dataset
French Conversational TTS Dataset Dataset Description This dataset contains high-fidelity French text-to-speech audio clips generated using Mistral's…
Reachy Mini Emotions Library
Reachy Mini Emotions Library
AIR-Bench-Dataset
AIR-Bench Arxiv: is the AIR-Bench dataset download page.AIR-Bench encompasses two dimensions: foundation and chat benchmarks. The former consists…
WideSearch
WideSearch: Benchmarking Agentic Broad Info-Seeking Dataset Summary WideSearch is a benchmark designed to evaluate the capabilities of Large…
browsecomp-plus
BrowseComp-Plus BrowseComp-Plus is a new benchmark for Deep-Research system, isolating the effect of the retriever and the LLM…
gaming-500-hours
Gaming Dataset (gaming-1) — 494.7 Hours Native PC/console gameplay screen-recordings, organized by game. Each workflow is one play…
LongBench-v2
LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks 🌐 Project Page: 💻 Github Repo: 📚…
SWE-bench_Pro
Dataset Summary SWE-Bench Pro is a challenging, enterprise-level dataset for testing agent ability on long-horizon software engineering tasks.…
SWE-bench_Verified
Dataset Summary SWE-bench Verified is a subset of 500 samples from the SWE-bench test set, which have been…
coraal
Corpus of Regional African American Language (CORAAL) Dataset link: Hugging Face Hub preparation scripts: coraal Citation Kendall, Tyler…
SWE-bench_Lite
Dataset Summary SWE-bench Lite is subset of SWE-bench, a dataset that tests systems’ ability to solve GitHub issues…