A GPT 4V Level Multimodal LLM on Your Phone GitHub Demo WeChat News π Pinned [2025.01.14] π₯π₯ π₯ We open source MiniCPM o 2.6 , with significant performance improvement over MiniCPM V 2.6 , and support real time speech to speech conversation and multimodal live streaming. Try it now. [2024.08.10] πππ MiniCPM Llama3 V 2.5 is now fully supported by official llama.cpp! GGUF models of various sizes are available here. [2024.08.06] π₯π₯π₯ We open source MiniCPM V 2.6 , which outperforms GPT 4V on single image, multi image and video understanding. It advances popular features of MiniCPM Llama3 V 2.5, and can support real time video understanding on iPad. Try it now! [2024.08.03] MiniCPM Llama3 V 2.5 technical report is released! See here. [2024.07.19] MiniCPM Llama3 V 2.5 supports vLLM now! See here. [2024.05.28] π« We now support LoRA fine tuning for MiniCPM Llama3 V 2.5, using only 2 V100 GPUs! See more statistics here. [2024.05.23] π₯π₯π₯ MiniCPM V tops GitHub Trending and HuggingFace Trending! Our demo, recommended by Hugging Face Gradioβs official account, is available here. Come and try it out! [2024.05.20] We open soure MiniCPM Llama3 V 2.5, it has improved OCR capability andβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy