This is an NF4 quantized model of Qwen image edit so it can run on GPUs using 20GB VRAM. You can run it on lower VRAM like 16GB. There were other NF4 models but they made the mistake of blindly quantizing all layers in the transformer. This one does not. We retain some layers at full precision in order to ensure that we get quality output. You can use the original Qwen Image Edit parameters. This model is not yet available for inference at JustLab.ai Model tested: Working perfectly even with 10 steps. Contact: JustLab.ai for commercial support Performance on rtx4090 20 steps about 78 seconds. 10 steps about 40 seconds. Interestingly I was under the impression that the Qwen VL could not be quantized which is why several projects use the full 15Gb model. Here I have quantized it too and it seems to be workign fine. Sample script. (min 20GB VRAM) The original Qwen Image attributions are included verbatim below. 💜 Qwen Chat      🤗 Hugging Face      🤖 ModelScope       📑 Tech Report       📑 Blog    🖥️ Demo      💬 WeChat (微信)      🫨 Discord       Github  &nbs…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy