GitHub Demo News [2025.01.14] 🔥 We open source MiniCPM o 2.6 , with significant performance improvement over MiniCPM V 2.6 , and support real time speech to speech conversation and multimodal live streaming. Try it now. [2024.08.06] 🔥 We open source MiniCPM V 2.6 , which outperforms GPT 4V on single image, multi image and video understanding. It advances popular features of MiniCPM Llama3 V 2.5, and can support real time video understanding on iPad. [2024.05.20] 🔥 The GPT 4V level multimodal model MiniCPM Llama3 V 2.5 is out. [2024.04.23] MiniCPM V 2.0 supports vLLM now! [2024.04.18] We create a HuggingFace Space to host the demo of MiniCPM V 2.0 at here! [2024.04.17] MiniCPM V 2.0 supports deploying WebUI Demo now! [2024.04.15] MiniCPM V 2.0 supports fine tuning with the SWIFT framework! [2024.04.12] We open source MiniCPM V 2.0, which achieves comparable performance with Gemini Pro in understanding scene text and outperforms strong Qwen VL Chat 9.6B and Yi VL 34B on OpenCompass , a comprehensive evaluation over 11 popular benchmarks. Click here to view the MiniCPM V 2.0 technical blog. MiniCPM V 2.0 MiniCPM V 2.8B is a strong multimodal large language model for efficient end s…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy