Krea 2 OSS Optimized FP8 Weights (Turbo) This repository provides an optimized FP8 (float8 e4m3fn) weight only quantized version of the newly released Krea 2 OSS (Turbo) transformer. This optimization reduces the model size from the original 24.76 GiB (BF16) down to 12.01 GiB , making it highly accessible and runnable on standard consumer hardware (such as 16GB and 24GB GPUs) without sacrificing output quality. ⚠️ Licensing & Disclaimer Original Model Creators : All credit goes to KREA.ai for the original research, architecture, and weights. License : This model is subject to the KREA 2 License Agreement . Please read and comply with the official license terms before using these weights: KREA 2 Licensing Terms. Purpose : This repository is a community contributed utility. It does not claim ownership of the original model or architecture. Its sole purpose is to provide optimized, consumer hardware friendly weights for the open source community. 🛠️ Quantization Details (Quality First FP8) Unlike generic global quantization scripts that aggressively convert every parameter (which often degrades generation details or introduces NaN/promotion calculation errors in neural networks), thi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy