Skip to content
Advertisement

MIT

552 results

Video

hawk

Hawk: Learning to Understand Open-World Video Anomalies (paper). Copyright Statement This statement serves to clarify that the copyright…

<1K·MIT
MultimodalTabularText

full-modality-data

Full Modality Dataset Statistics Video Statistics Total Videos: 28,472 Total Duration: 1422.33 hours Average Duration: 179.84 seconds Median…

1M–10M·MIT·Parquet
MultimodalSensor / Time-seriesTabular

pusht

This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v2.0", "robot type": "unknown", "total episodes":…

10K–100K·MIT·Parquet
Video

siftformer2-data

SiftFormer2 Data K400 / SSv2 video classification 학습용 데이터. 구성 siftformer2-data/ ├── k400/ │ ├── train manifest.jsonl │…

10K–100K·MIT
MultimodalTextVideo

deform360

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models Project Page Paper GitHub Repository Deform360 is a…

MIT
Video

behavior_224_rgb

This is the compressed version of the original BEHAVIOR dataset It contains only RGB videos compressed to 224x224…

MIT
Video

FreeTacMan

📦 FreeTacman Robot-free Visuo-Tactile Data Collection System for Contact-rich Manipulation ICRA 2026 🎯 Overview This dataset supports the…

MIT
Video

American-Sign-Language-Dataset

American Sign Language (ASL) Dataset Description:This dataset contains 108,618 videos representing 2,208 ASL words, with each word having…

MIT
Video

EgoLife

Data cleaning, stay tuned! Please refer to first for general info. Checkout the paper EgoLife ( for more…

10K–100K·MIT
AudioMultimodalText

EuroSpeech-24kHz

EuroSpeech 24 kHz Dataset Dataset Description EuroSpeech is a large-scale multilingual speech corpus containing high-quality aligned parliamentary…

10M–100M·MIT·Parquet
AudioMultimodalText

opendata-iisys-hui

HUI-Audio-Corpus-German Dataset Overview The HUI-Audio-Corpus-German is a high-quality Text-To-Speech (TTS) dataset developed by researchers at the…

10K–100K·MIT·Parquet
AudioMultimodalText

UrbanSound8K

UrbanSound8K This is an audio classification dataset for Sound Event Classification. Classes = 10 , Split = Ten-Fold…

10K–100K·MIT·CSV
Advertisement