gpt oss 120b Fable 5 Distilled A LoRA fine tuned MoE coding agent model distilled from real world Claude Code programming sessions. Merged into a single complete model — download and run directly on Apple Silicon via MLX. Recommended temperature: 0.95 — this model was trained on opencode agent traces where the assistant alternates between reasoning, tool calls, and code generation. Slightly higher temperature preserves this multi modal behavior. GGUF format is ready (autotrust/gpt oss 120b Fable 5 Distilled GGUF) Test Results (HumanEval PASS@1) Model Quant Score Pass/Fail Fable 5 Distilled Q5 0 (67 GB) 99.39% 163/164 Fable 5 Distilled Q8 0 (115 GB) 98.78% 162/164 gpt oss 120b (base) MXFP4 83.54% 137/164 Test Details(cloudyu/gpt oss 120b Fable 5 Distilled GGUF) Model Summary gpt oss 120b Fable 5 Distilled is a fully merged LoRA fine tune of gpt oss 120b heretic v2 mxfp4 q8 hi mlx , trained using the Muon optimizer with Newton Schulz orthogonalization on Apple MLX. The fine tuning data consists of 352 real world opencode programming sessions extracted from armand0e/claude fable 5 claude code . This is a complete model upload — no separate adapter files or base model are needed. The m…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy