Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning Repo: Paper: Introduction We introduce MathCoder-VL, a series…
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning Repo: Paper: Introduction We introduce MathCoder-VL, a series of open-source large multimodal models (LMMs) specifically tailored for general math problem-solving. We also introduce FigCodifier-8B, an image-to-code model. Base Model Ours Mini-InternVL-Chat-2B-V1-5 MathCoder-VL-2B InternVL2-8B… See the full description on the dataset page:
Source: Hugging Face Hub (MathLLMs/ImgCode-8.6M). Metadata imported from the dataset’s Hub tags.