Qwen3.5 0.8B MLX 4bit This is a 4 bit quantized MLX version of Qwen/Qwen3.5 0.8B for Apple Silicon. Model Details Original Model: Qwen/Qwen3.5 0.8B Quantization: 4 bit (5.863 bits per weight) Group Size: 64 Format: MLX SafeTensors Framework: mlx vlm Disk Size: ~622M Conversion Details This model was converted using mlx vlm from the pc/fix qwen35 predicate branch, which includes fixes for Qwen3.5 model support (proper handling of MoE gate layers, shared expert gate , and A log casting). Conversion command: Important Note A better, more optimized conversion may be available from @Prince (@Blaizzy) in the MLX VLM community. Check the mlx community organization for updated versions as official Qwen3.5 support is merged into the main mlx vlm branch. Related Models bf16 (full precision): mlx community/Qwen3.5 0.8B MLX bf16 8 bit quantized: mlx community/Qwen3.5 0.8B MLX 8bit Original: Qwen/Qwen3.5 0.8B Usage CLI: License This model inherits the Apache 2.0 license from the original Qwen model.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy