latex-formulas-80M
For more details, please refer to the 𝐓𝐞𝐱𝐓𝐞𝐥𝐥𝐞𝐫 GitHub repository. IMPORTANT NOTE!!! The handwritten subset of this dataset…
807 results
For more details, please refer to the 𝐓𝐞𝐱𝐓𝐞𝐥𝐥𝐞𝐫 GitHub repository. IMPORTANT NOTE!!! The handwritten subset of this dataset…
SpatialEdit-500K SpatialEdit-500K is a synthetic training dataset for fine-grained image spatial editing. It is built for learning geometry-aware…
personalized visual instruction tuning
Leopard-Instruct Paper Github Models-LLaVA Models-Idefics2 Summaries Leopard-Instruct is a large instruction-tuning dataset, comprising 925K…
LLaVA-OneVision-1.5 Instruction Data Paper Code 📌 Introduction This dataset, LLaVA-OneVision-1.5-Instruct, was collected and integrated during the…
This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v2.0", "robot type": "kuka iiwa", "total…
This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v2.0", "robot type": "widowx", "total episodes":…
LLaVA-OneVision-2-Data Training data for the LLaVA-OneVision-2 multimodal model family, covering large-scale video and spatial reasoning corpora used…
This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v2.0", "robot type": "google robot", "total…
Open-source, Apache-2.0 labeling tool by HumanSignal spanning text, images, audio, video, and time series, with configurable UIs and ML backend…
Multimodal