Suzume [Paper] [Dataset] This Suzume 8B, a multilingual finetune of Llama 3 (meta llama/Meta Llama 3 8B Instruct). Llama 3 has exhibited excellent performance on many English language benchmarks. However, it also seemingly been finetuned on mostly English data, meaning that it will respond in English, even if prompted in other languages. We have fine tuned Llama 3 on almost 90,000 multilingual conversations meaning that this model has the smarts of Llama 3 but has the added ability to chat in more languages. Please feel free to comment on this model and give us feedback in the Community tab! We will release a paper in the future describing how we made the training data, the model, and the evaluations we have conducted of it. How to use The easiest way to use this model on your own computer is to use the GGUF version of this model (lightblue/suzume llama 3 8B multilingual gguf) using a program such as jan.ai or LM Studio. If you want to use this model directly in Python, we recommend using vLLM for the fastest inference speeds. Evaluation scores We achieve the following MT Bench scores across 6 languages: meta llama/Meta Llama 3 8B Instruct lightblue/suzume llama 3 8B multilingual N…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy