Skip to content
Advertisement

Robotics / Physical AI

559 results

BasicAI

Labeling & Annotation

Irvine, CA-based data-annotation platform and services provider for image, video, 3D sensor-fusion, and LLM/GenAI data.

3D / Point CloudImageVideo

Objectways

Labeling & Annotation

Phoenix-based data-annotation and content-moderation provider, including embodied-AI datasets for robotics.

ImageTextVideo

Digital Divide Data

Labeling & Annotation

Social-enterprise data-annotation provider founded in Cambodia in 2001, offering impact-sourced labeling for physical and generative AI.

ImageTextVideo
Image

offroad-global-nav

Offroad-global-nav Geospatial Dataset Overview This repository contains the dataset introduced in: “Learning Traversability-Aware Global Planners for…

<1K·CC-BY-NC·Images (folder)
ImageMultimodalText

BEAF

BEAF: Before-After Changes for Hallucination Evaluation BEAF is a benchmark for evaluating object hallucination in vision-language models using…

10K–100K·CC-BY·Parquet
ImageMultimodalTabular

AgentVQA

AgentVQA: A Multi-Domain Visual Question Answering Dataset AgentVQA is a comprehensive dataset for training and evaluating visual agents…

10K–100K·MIT
MultimodalTextVideo

UrbanVideo-Bench

ACL'25 Oral UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban Spaces This repository contains…

1K–10K·MIT·Parquet
ImageMultimodalText

FineBench

Dataset Card for FineBench FineBench is a large-scale, multiple-choice Video Question Answering (VQA) dataset designed specifically to evaluate…

100K–1M·MIT·JSON
Image

SandThink

SandThink Dataset (v1.0) SandThink 是一个专为具身智能 (Embodied AI) 任务设计的大规模指令微调与偏好对齐数据集。该数据集通过结构化的 Chain-of-Thought (CoT) 推理过程,显著提升了 Vision-Language-Action…

<1K·MIT·Images (folder)
MultimodalTextVideo

UrbanVideo-Bench

ACL'25 Oral UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban Spaces This repository contains…

1K–10K·MIT·Parquet
Advertisement