This is a GGUF quantized version of FLUX.2 klein 9B. unsloth/FLUX.2 klein 9B GGUF uses Unsloth Dynamic 2.0 methodology for SOTA performance. Important layers are upcasted to higher precision. Uses tooling from ComfyUI GGUF by city96. The FLUX.2 [klein] model family are our fastest image models to date. FLUX.2 [klein] unifies generation and editing in a single compact architecture, delivering state of the art quality with end to end inference in as low as under a second . Built for applications that require real time image generation without sacrificing quality. FLUX.2 [klein] 9B is a 9 billion parameter rectified flow transformer capable of generating images from text descriptions and supports multi reference editing capabilities. Our flagship small model. Defines the Pareto frontier for quality vs. latency across text to image, single reference editing, and multi reference generation. Matches or exceeds models 5x its size—in under half a second. Built on a 9B flow model with 8B Qwen3 text embedder, step distilled to 4 inference steps. For more information, please read our blog post. Key Features 1. A distilled model for sub second image generation with outstanding quality. 2. Text…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy