InternLM InternLM HOT 💻Github Repo • 🤔Reporting Issues • 📜Technical Report 👋 join us on Discord and WeChat Introduction InternLM2.5 has open sourced a 20 billion parameter base model and a chat model tailored for practical scenarios. The model has the following characteristics: Outstanding reasoning capability : State of the art performance on Math reasoning, surpassing models like Llama3 and Gemma2 27B. Stronger tool use : InternLM2.5 supports gathering information from more than 100 web pages, corresponding implementation has be released in MindSearch. InternLM2.5 has better tool utilization related capabilities in instruction following, tool selection and reflection. See examples. InternLM2.5 20B Chat Performance Evaluation We conducted a comprehensive evaluation of InternLM using the open source evaluation tool OpenCompass. The evaluation covered five dimensions of capabilities: disciplinary competence, language competence, knowledge competence, inference competence, and comprehension competence. Here are some of the evaluation results, and you can visit the OpenCompass leaderboard for more evaluation results. Benchmark InternLM2.5 20B Chat Gemma2 27B IT MMLU…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy