Skip to content
Advertisement
ImageMultimodalText

ImgCode-8.6M

MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning Repo: Paper: Introduction We introduce MathCoder-VL, a series…

MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning Repo: Paper: Introduction We introduce MathCoder-VL, a series of open-source large multimodal models (LMMs) specifically tailored for general math problem-solving. We also introduce FigCodifier-8B, an image-to-code model. Base Model Ours Mini-InternVL-Chat-2B-V1-5 MathCoder-VL-2B InternVL2-8B… See the full description on the dataset page:

Source: Hugging Face Hub (MathLLMs/ImgCode-8.6M). Metadata imported from the dataset’s Hub tags.

Advertisement