💡 Found this resource helpful? Creating and maintaining open source AI models and datasets requires significant computational resources. If this work has been valuable to you, consider supporting my research to help me continue building tools that benefit the entire AI community. Every contribution directly funds more open source innovation! ☕ Model Architecture Base Model: Meta Llama 3 8B Specialization: Italian Language Evaluation For a detailed comparison of model performance, check out the Leaderboard for Italian Language Models. Here's a breakdown of the performance metrics: Metric hellaswag it acc norm arc it acc norm m mmlu it 5 shot acc Average : : : : : Accuracy Normalized 0.6518 0.5441 0.5729 0.5896 How to Use Developer [Michele Montebovi] Open LLM Leaderboard Evaluation Results Detailed results can be found here Metric Value : Avg. 26.58 IFEval (0 Shot) 75.30 BBH (3 Shot) 28.08 MATH Lvl 5 (4 Shot) 5.36 GPQA (0 shot) 7.38 MuSR (0 shot) 11.68 MMLU PRO (5 shot) 31.69
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy