LLaVA Med v1.5, using mistralai/Mistral 7B Instruct v0.2 as LLM for a better commercial license Large Language and Vision Assistant for bioMedicine (i.e., “LLaVA Med”) is a large language and vision model trained using a curriculum learning method for adapting LLaVA to the biomedical domain. It is an open source release intended for research use only to facilitate reproducibility of the corresponding paper which claims improved performance for open ended biomedical questions answering tasks, including common visual question answering (VQA) benchmark datasets such as PathVQA and VQA RAD. LLaVA Med was proposed in LLaVA Med: Training a Large Language and Vision Assistant for Biomedicine in One Day by Chunyuan Li, Cliff Wong, Sheng Zhang, Naoto Usuyama, Haotian Liu, Jianwei Yang, Tristan Naumann, Hoifung Poon, Jianfeng Gao. Model date: LLaVA Med v1.5 Mistral 7B was trained in April 2024. Paper or resources for more information: https://aka.ms/llava med Where to send questions or comments about the model: https://github.com/microsoft/LLaVA Med/issues License mistralai/Mistral 7B Instruct v0.2 license. Intended use The data, code, and model checkpoints are intended to be used solely for…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy