Daredevil 8B abliterated Abliterated version of mlabonne/Daredevil 8B using failspy's notebook. It based on the technique described in the blog post "Refusal in LLMs is mediated by a single direction". Thanks to Andy Arditi, Oscar Balcells Obeso, Aaquib111, Wes Gurnee, Neel Nanda, and failspy. ๐ Applications This is an uncensored model. You can use it for any application that doesn't require alignment, like role playing. Tested on LM Studio using the "Llama 3" preset. โก Quantization GGUF : https://huggingface.co/mlabonne/Daredevil 8B abliterated GGUF ๐ Evaluation Open LLM Leaderboard Daredevil 8B abliterated is the second best performing 8B model on the Open LLM Leaderboard in terms of MMLU score (27 May 24). Nous Evaluation performed using LLM AutoEval. See the entire leaderboard here. Model Average AGIEval GPT4All TruthfulQA Bigbench : : : : : mlabonne/Daredevil 8B ๐ 55.87 44.13 73.52 59.05 46.77 mlabonne/Daredevil 8B abliterated ๐ 55.06 43.29 73.33 57.47 46.17 mlabonne/Llama 3 8B Instruct abliterated dpomix ๐ 52.26 41.6 69.95 54.22 43.26 meta llama/Meta Llama 3 8B Instruct ๐ 51.34 41.22 69.86 51.65 42.64 failspy/Meta Llama 3 8B Instruct abliterated v3 ๐ 51.21 40.23 69.5 5โฆ
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy