Med42 v2 A Suite of Clinically aligned Large Language Models Med42 v2 is a suite of open access clinical large language models (LLM) instruct and preference tuned by M42 to expand access to medical knowledge. Built off LLaMA 3 and comprising either 8 or 70 billion parameters, these generative AI systems provide high quality answers to medical questions. Key performance metrics: Med42 v2 70B outperforms GPT 4.0 in most of the MCQA tasks. Med42 v2 70B achieves a MedQA zero shot performance of 79.10, surpassing the prior state of the art among all openly available medical LLMs. Med42 v2 70B sits at the top of the Clinical Elo Rating Leaderboard. Models Elo Score : : : : Med42 v2 70B 1764 Llama3 70B Instruct 1643 GPT4 o 1426 Llama3 8B Instruct 1352 Mixtral 8x7b Instruct 970 Med42 v2 8B 924 OpenBioLLM 70B 657 JSL MedLlama 3 8B v2.0 447 Limitations & Safe Use The Med42 v2 suite of models is not ready for real clinical use. Extensive human evaluation is undergoing as it is required to ensure safety. Potential for generating incorrect or harmful information. Risk of perpetuating biases in training data. Use this suite of models responsibly! Do not rely on them for medical usage without rig…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy