DeepSeek V4 Flash · Abliterated · GGUF CyberNeurova research — cyberneurova.ai Release: v2 (three direction multi turn aware ablation, 1338 prompt capture corpus) A permanently abliterated version of deepseek ai/DeepSeek V4 Flash (284 B FP8 MoE) packaged as GGUF for llama.cpp . The abliteration is baked into the weights at conversion time — no runtime hooks, no slowdown, no reliance on the inference framework supporting custom code paths. Status: experimental research artifact. Built with antirez's llama.cpp DeepSeek V4 Flash fork, which is itself experimental. Use at your own discretion. What's new in v2 v1 v2 Direction stack 2 3 (added a residual targeting direction) Capture corpus 33 prompts 1338 prompts (AdvBench + JBB + HarmfulQA + SafeRLHF + MaliciousInstruct + bundled) Refusal rate (8 bench safety) 0.0% 0.0% Refusal rate (55 prompt OOD probe) not measured 3.6% (baseline: 81.8%) Tool calling format compliance 74.2% 99.2% Bug finding 78.3% 85.0% Hacking compliance 88.7% 90.0% Cyber weapons compliance 87.3% 90.0% Coding / coherence / reasoning unchanged unchanged Plain English: v2 strictly dominates v1 on every dimension we measure. The big ticket fixes are the OOD soft refusal…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy