🤗 1.5 HF Models     📕 Blog   Table of Contents Kanana 1.5 v 3b instruct Intended Use Model Details Evaluation Model Configuration Summary Overview Image Benchmarks (EN) Image Benchmarks (KO) Multimodal Instruction Following Benchmarks (EN, KO) Note on Benchmarking Methodology Usage Requirements Quickstart Limitations Contributors Contact kanana 1.5 v 3b instruct The Unified Foundation Model (UFO) task force of Kanana at Kakao developed and released the Kanana V family of multimodal large language models (MLLMs), a collection of pretrained text/image to text (TI2T) models. Intended Use kanana 1.5 v 3b instruct is intended for research and application development in multimodal understanding and text generation tasks. Typical use cases include image captioning, document understanding, OCR based reasoning, and multimodal instruction following in both English and Korean. The model is optimized for both general purpose and Korea specific benchmarks, making it suitable for bilingual environments. Model Details Developed by: Unified Foundation Model (UFO) TF at Kakao Language(s) : ['en', 'ko'] Model Architecture: kanana 1.5 v 3b instruct has 3.6B parameters and contains image…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy