⚠️ DEPRECATED — Recommended: leonsarmiento/Ornith 1.0 35B 5bit XL mlx This model has been superseded by the BaseQuant XL variant, which keeps routing critical layers (MoE router gate, shared expert, lm head) in full bf16 precision for improved quality. Benchmark comparisons are available in the XL model card. → Download leonsarmiento/Ornith 1.0 35B 5bit XL mlx leonsarmiento/Ornith 1.0 35B 5bit mlx This model was converted to MLX format from deepreinforce ai/Ornith 1.0 35B using mixed 5/8 bit quantization optimized for Apple Silicon. The vision encoder is preserved and quantized at 5 bit, making this a full multimodal model. Ornith 1.0 35B is a 35B parameter MoE (Mixture of Experts) model fine tuned from Qwen3.5 35B A3B by DeepReinforce AI, using a self improving RL training framework that jointly optimizes scaffold and solution rollouts for agentic coding tasks. Despite 35B total parameters, only ~3B are activated per token. It features 256 experts (8 active per token + 1 shared expert), hybrid full + linear (Gated DeltaNet) attention, and a vision encoder. Benchmark Highlights Benchmark Ornith 1.0 35B Qwen3.5 35B Qwen3.6 35B Terminal Bench 2.1 (Terminus 2) 64.2 41.4 52.5 Terminal…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy