Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Dataset Card for μ-Bench (Leaderboard Code) μ-Bench is a multilingual transcription benchmark built from real customer-service phone conversations.…
Dataset Card for μ-Bench (Leaderboard Code) μ-Bench is a multilingual transcription benchmark built from real customer-service phone conversations. Dataset Details Dataset Description Most public ASR benchmarks are either English-only or built from read speech in quiet studios. μ-Bench fills that gap: real phone-call audio, five languages, and metrics that go beyond Word Error Rate to distinguish meaning-changing errors from surface-level ones. The calls… See the full description on the dataset page:
Source: Hugging Face Hub (sierra-research/mu-bench). Metadata imported from the dataset’s Hub tags.