Xiangxin 2XL Chat 1048k 我们提供私有化模型训练服务,如果您需要训练行业模型、领域模型或者私有模型,请联系我们: wanglei@xiangxinai.cn We offer customized model training services. If you need to train industry specific models, domain specific models, or private models, please contact us at: wanglei@xiangxinai.cn. 模型介绍/Introduction Xiangxin 2XL Chat 1048k是象信AI基于Meta Llama 3 70B Instruct模型和Gradient AI的扩充上下文的工作,利用自行研发的中文价值观对齐数据集进行ORPO训练而形成的Chat模型。该模型具备更强的中文能力和中文价值观,其上下文长度达到100万字。在模型性能方面,该模型在ARC、HellaSwag、MMLU、TruthfulQA mc2、Winogrande、GSM8K flex、CMMLU、CEVAL VALID等八项测评中,取得了平均分70.22分的成绩,超过了Gradientai Llama 3 70B Instruct Gradient 1048k。我们的训练数据并不包含任何测评数据集。 Xiangxin 2XL Chat 1048k is a Chat model developed by Xiangxin AI, based on the Meta Llama 3 70B Instruct model and expanded context from Gradient AI. It was trained using a proprietary Chinese value aligned dataset through ORPO training, resulting in enhanced Chinese proficiency and alignment with Chinese values. The model has a context length of up to 1 million words. In terms of performance, it surpassed the Gradientai Llama 3 70B Instruct Gradient 1048k model with an average score of 70.22 across eight evaluations including ARC, HellaSwag, MMLU, TruthfulQA mc2, Winogrande, GSM…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy