nuScenes
1,000 driving scenes with 360° camera, LiDAR and radar.
CoLT Training Data This repository contains the training data for CoLT, a latent reasoning model without relying on auxiliary image annotations. It…
CoLT Training Data This repository contains the training data for CoLT, a latent reasoning model without relying on auxiliary image annotations. It could notably reduce the inference time by 10.3× and text decoding time by 20.4×, achieving superior efficiency. Code: Our training data is built from OneThinker, an all-in-one reasoning model for image and video, as presented in the paper OneThinker: All-in-one Reasoning Model for Image and Video.… See the full description on the dataset page: Train Dataset.
Source: Hugging Face Hub (hulianyuyy/CoLT_Train_Dataset). Metadata imported from the dataset’s Hub tags.