Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
OmniMMI Paper: OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts Code Dataset Description we introduce OmniMMI,…
OmniMMI Paper: OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts Code Dataset Description we introduce OmniMMI, a comprehensive multi-modal interaction benchmark tailored for OmniLLMs in streaming video contexts. OmniMMI encompasses over 1,121 interactive videos and 2,290 questions, addressing two critical yet underexplored challenges in existing video benchmarks: streaming video understanding and proactive reasoning, across six… See the full description on the dataset page:
Source: Hugging Face Hub (bigai-nlco/OmniMMI). Metadata imported from the dataset’s Hub tags.