A large language model was developed with goals including excellent multilingual support, superior knowledge capabilities and cost efficiency. Introduction Ghost 8B Beta (Llama 3 Ghost 8B Beta) is a large language model developed with goals that include excellent multilingual support, superior knowledge capabilities, and cost effectiveness. The model comes in two context length versions, 8k and 128k, along with multilingual function tools support by default. The Ghost 8B Beta model outperforms prominent models such as Llama 3.1 8B Instruct, GPT 3.5 Turbo in the lc winrate score. In addition, it also outperforms Claude 3 Opus, Claude 3 Sonnet, GPT 4, and Mistral Large when comparing the winrate score of AlpacaEval 2.0, \ . Updates 16 Aug 2024 : The model has been released to version 160824, expanding support from 9 languages to 16 languages. The model has improved math, reasoning, and following instructions better than the previous version. Thoughts We believe that it is possible to optimize language models that are not too large to achieve better capabilities in terms of cross linguistic understanding and solving complex tasks. The potential of these models is often mentioned as…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy