MiniCPM5 1B Claude Opus Fable5 Thinking GGUF 📢 V2.0 is available — We have released an updated model with enhanced tool calling capabilities. Welcome to try the new version: Transformers: MiniCPM5 1B Claude Opus Fable5 V2 Thinking GGUF: MiniCPM5 1B Claude Opus Fable5 V2 Thinking GGUF GGUF quantizations of MiniCPM5 1B Claude Opus Fable5 Thinking for llama.cpp, Ollama, LM Studio, jan, KoboldCpp, and other GGUF runtimes. 中文说明 This repository provides local deployment builds of a 1B Thinking model fine tuned on Fable 5 data atop openbmb/MiniCPM5 1B. The GGUF files embed MiniCPM5's native chat template for llama.cpp compatible runtimes. Transformers checkpoint: MiniCPM5 1B Claude Opus Fable5 Thinking Files File Quant Size Notes MiniCPM5 1B Claude Opus Fable5 Thinking Q4 K M.gguf Q4 K M ~657 MB smallest footprint MiniCPM5 1B Claude Opus Fable5 Thinking Q5 K M.gguf Q5 K M ~751 MB balanced quality / size MiniCPM5 1B Claude Opus Fable5 Thinking Q8 0.gguf Q8 0 ~1.1 GB recommended default MiniCPM5 1B Claude Opus Fable5 Thinking F16.gguf F16 ~2.1 GB full precision conversion base Q8 0 is the recommended default quant for this 1B model. Quick start llama.cpp ( llama cli ) The model supports up…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy