HAKARI Bench Results This dataset stores raw benchmark result artifacts generated by HAKARI Bench. Raw results: per task JSON (.xz) result files measured by HAKARI bench. Leaderboard: https://huggingface.co/spaces/hakari bench/leaderboard GitHub repository: https://github.com/hakari bench/hakari bench Contributing official model results: follow the new model evaluation workflow to evaluate a model and submit results for HAKARI Bench review: https://github.com/hakari bench/hakari bench/blob/main/docs/new model results workflow.md Leaderboard database: transformed data derived from these raw JSON results for leaderboard use: https://huggingface.co/datasets/hakari bench/leaderboard database
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy