Ornith 1.0 9B: 1M Context + MTP + Vision 181717?logo=github) Mirrors: Hugging Face ModelScope (full quant ladder always available on ModelScope) DeepReinforce's Ornith 1.0 9B (9B dense, Qwen3.5 family) with YaRN RoPE scaling baked into the GGUF metadata for a 1,048,576 token context window , 4x the native 262,144, shipping with the MTP speculative decoding layer baked in and a vision tower alongside. Weights bit identical to the source builds; llama.cpp and Ollama apply the context extension with no extra flags. Files: MTP first File Size Pick it when ornith 1.0 9b 1M MTP Q4 K M.gguf 5.8 GB Default. MTP layer baked in: +15 to 38 percent decode via llama.cpp speculative decoding ornith 1.0 9b 1M MTP Q8 0.gguf 9.8 GB Max quality. This is the config that scored 10/10 at the full 1M with f16 KV mmproj ornith 9b f16.gguf 918 MB Vision, attach with mmproj This repo intentionally carries only the MTP builds. The full 10 quant ladder (IQ2 M through bf16) lives on the ModelScope mirror. Every file, every mirror Nothing was discontinued: every quant is one click away. Hugging Face carries the curated picks, ModelScope always carries everything, and Ollama serves ready to run tags. On Ollama…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy