Mini InternVL Chat 2B V1 5 [\[π GitHub\]](https://github.com/OpenGVLab/InternVL) [\[π InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[π InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[π Mini InternVL\]](https://arxiv.org/abs/2410.16261) [\[π InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[π Blog\]](https://internvl.github.io/blog/) [\[π¨οΈ Chat Demo\]](https://internvl.opengvlab.com/) [\[π€ HF Demo\]](https://huggingface.co/spaces/OpenGVLab/InternVL) [\[π Quick Start\]]( quick start) [\[π Documents\]](https://internvl.readthedocs.io/en/latest/) Introduction You can run multimodal large models using a 1080Ti now. We are delighted to introduce the Mini InternVL Chat series. In the era of large language models, many researchers have started to focus on smaller language models, such as Gemma 2B, Qwen 1.8B, and InternLM2 1.8B. Inspired by their efforts, we have distilled our vision foundation model InternViT 6B 448px V1 5 down to 300M and used InternLM2 Chat 1.8B or Phi 3 mini 128k instruct as our language model. This resulted in a small multimodal model with excellent performance. As shown in the figure below, we adopted the same model archβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy