Qwen3.6 27B AEON Ultimate Uncensored Multimodal NVFP4 MTP Deployment, operations & benchmarks → github.com/AEON 7/Qwen3.6 27B AEON Ultimate Uncensored DFlash The GitHub repo is the source of truth for the production deployment guide, hardware tuned docker compose configs, full configuration reference, measured benchmarks, and AGENTS.md — an operator's manual that pre empts common stale documentation traps. 🙏 Reference recipe credit: The modelopt + MTP graft pipeline used to build this variant is based on sakamakismile 's validated Qwen3.6 27B NVFP4 MTP series (22K+ downloads). They worked out the modelopt config, the per projection quantization choices, and the MTP head graft technique on the un abliterated base; we adapted the same recipe to AEON Ultimate's abliterated weights. The reference benchmark numbers cited below are theirs. Full credit for the recipe → sakamakismile. 🆕 AEON vLLM Ultimate container (2026 06 04) ghcr.io/aeon 7/aeon vllm ultimate:latest — vLLM 0.24.0 (= :2026 07 01 v0.24.0 ) + PR 44389 NVFP4 KV cache (~3× capacity) + DFlash + TurboQuant K8V4 + AEON sm 121a patches. Same recipe family as the Multimodal NVFP4 MTP XS sibling which has been benchmarked end to…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy