InternVL3 5 2B Instruct [\[๐ GitHub\]](https://github.com/OpenGVLab/InternVL) [\[๐ InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[๐ InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[๐ InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[๐ InternVL2.5 MPO\]](https://huggingface.co/papers/2411.10442) [\[๐ InternVL3\]](https://huggingface.co/papers/2504.10479) [\[๐ InternVL3.5\]](https://huggingface.co/papers/2508.18265) [\[๐ Blog\]](https://internvl.github.io/blog/) [\[๐จ๏ธ Chat Demo\]](https://chat.intern ai.org.cn/) [\[๐ Quick Start\]]( quick start) [\[๐ Documents\]](https://internvl.readthedocs.io/en/latest/) Introduction We introduce InternVL3.5 , a new family of open source multimodal models that significantly advances versatility, reasoning capability, and inference efficiency along the InternVL series. A key innovation is the Cascade Reinforcement Learning (Cascade RL) framework, which enhances reasoning through a two stage process: offline RL for stable convergence and online RL for refined alignment. This coarse to fine training strategy leads to substantial improvements on downstream reasoning tasks, e.g., MMMU and MathVista. To optโฆ
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy