Qwen3.5 9B MLX 4bit This is a quantized MLX version of Qwen/Qwen3.5 9B for Apple Silicon. Model Details Original Model: Qwen/Qwen3.5 9B Quantization: 4 bit (~5.059 bits per weight) Group Size: 64 Format: MLX SafeTensors Framework: mlx vlm Conversion Details This model was converted using mlx vlm with 4 bit quantization. Conversion command: Important Note A better, more optimized conversion may be available from @Prince (@Blaizzy) in the MLX VLM community. Check the mlx community organization for updated versions as official Qwen3.5 support is merged into the main mlx vlm branch. Usage Or from the command line: Performance Disk Size: ~5.6 GB Runs efficiently on Apple Silicon Macs (M1/M2/M3/M4) Lower memory footprint compared to 8 bit quantization License This model inherits the Apache 2.0 license from the original Qwen3.5 9B model.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy