StableLM Zephyr 3B Please note: For commercial use, please refer to https://stability.ai/license. Model Description StableLM Zephyr 3B is a 3 billion parameter instruction tuned inspired by HugginFaceH4's Zephyr 7B training pipeline this model was trained on a mix of publicly available datasets, synthetic datasets using Direct Preference Optimization (DPO), evaluation for this model based on MT Bench and Alpaca Benchmark Usage StableLM Zephyr 3B uses the following instruction format: This format is also available through the tokenizer's apply chat template method: You can also see how to run a performance optimized version of this model here using OpenVINO from Intel. Model Details Developed by : Stability AI Model type : StableLM Zephyr 3B model is an auto regressive language model based on the transformer decoder architecture. Language(s) : English Library : Alignment Handbook Finetuned from model : stabilityai/stablelm 3b 4e1t License : StabilityAI Community License. Commercial License : to use this model commercially, please refer to https://stability.ai/license Contact : For questions and comments about the model, please email lm@stability.ai Training Dataset The dataset is co…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy