Video
706 results
JavisInst-Omni
JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation HomePage Paper GitHub TL;DR We introduce JavisGPT, a…
TAVGBench_1m
Installation Download this repo to a local folder, and unzip these .zip files under the TAVGBench 1m/data/. Then,…
FLARE
FLARE: Full-Modality Long-Video Audiovisual Retrieval Benchmark with User-Simulated Queries 🤗 About This Benchmark This repository hosts the full…
gaming-500-hours
Gaming Dataset (gaming-1) — 494.7 Hours Native PC/console gameplay screen-recordings, organized by game. Each workflow is one play…
HIW-500: Humanoids In-the-Wild Dataset (LeRobot)
HIW-500: Humanoids In-the-Wild Dataset (LeRobot)
LLaVA-Video-178K
Dataset Card for LLaVA-Video-178K Uses This dataset is used for the training of the LLaVA-Video model. We only…
InsViE
InsViE-1M: Effective Instruction-based Video Editing with Elaborate Dataset Construction Citation If you find this work helpful, please consider…
ProLongVid_data
Dataset Card for ProLongVid-data Uses This dataset is used for the training of the ProLongVid model. We only…
L2D
TL;DR of L2D, the world's largest self-driving dataset! Read more about L2D on the official Huggingface blog: LeRobot…
shofo-tiktok-general-small
Shofo TikTok General (Small) Overview Shofo TikTok General (Small) is a dataset containing 50,000 TikTok videos with comprehensive…
agibot_alpha_v30
This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v3.0", "robot type": "AgiBot A2D", "total…
GOKU-2M
Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing GOKU-2M is a large-scale, unified instruction-based…