Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Enhanced Audiosnippets Long 2.8M Enhanced version of mitermix/audiosnippets long 2 8M with speech enhancement, emotion annotations, speaker…
Enhanced Audiosnippets Long 2.8M Enhanced version of mitermix/audiosnippets long 2 8M with speech enhancement, emotion annotations, speaker embeddings, and comprehensive metadata analysis. Dataset Summary Metric Value Total samples 2,633,037 Total audio hours 4,932 h Duration range 3.0s – 1124.3s Mean duration 6.7s Audio format WAV, 48kHz mono Tar files 1,410 Processing Pipeline Each audio sample was processed through: Speech… See the full description on the dataset page:
Source: Hugging Face Hub (ai-music4you3/enhanced-audiosnippets-long-2-8M). Metadata imported from the dataset’s Hub tags.