Quants of Nex N2 Pro, a fine tune built on Qwen 3.5 397B A17B. Basically the Qwen 3.6 397B that we never got. Comes with mmproj for vision, but isn't shipped with MTP. All quants target 16/24/32GB GPUs, with varying amounts of RAM depending on the quant. Specific quant details: IQ5 KS ik fork only Only works on ik llama.cpp, targets a 256GB RAM system + nvidia GPU 24/32GB. Will eat 20822MB of VRAM and 214GB of RAM with this config (needs a strong CPU, like 9950x3d, or PP will be slower): Will eat 23500MB of VRAM and 214GB of RAM with this config (increases PP speed for weaker CPUs at the cost of more VRAM usage): Details: IQ4 KSS ik fork only Works with ik only, targets a 192GB RAM system + any GPU 24GB. Will eat 19450MB of VRAM and 182GB of RAM with standard config: Details: IQ3 M mainline compatible Works with mainline and ik, targets a 196GB RAM system + any GPU 24GB. Will eat 19600MB of VRAM and 180GB of RAM with standard config: Details: IQ3 XXS mainline compatible Works with mainline and ik, targets a 196GB RAM system + any GPU 24GB. Will eat 18930MB of VRAM and 151GB of RAM with standard config: Details: IQ2 M mainline compatible Works with mainline and ik, targets a 196GB R…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy