[!TIP] Support this work → · X · GitHub · REAP paper · Cerebras REAP DeepSeek V4 Flash 162B GGUF GGUF quantization of 0xSero/DeepSeek V4 Flash 162B. At a glance Base model 0xSero/DeepSeek V4 Flash 162B Format GGUF Total params 162B Active / token — Experts / layer — Layers — Hidden size — Context — On disk size 149 GB Which variant should I pick? Variant Format Link DeepSeek V4 Flash 162B BF16 link DeepSeek V4 Flash 162B GGUF (this) GGUF link DeepSeek V4 Flash 180B BF16 link DeepSeek V4 Flash 180B GGUF GGUF link DeepSeek V4 Flash 213B BF16 link This repository contains DS4/DwarfStar GGUF conversions of DeepSeek V4 Flash Spark Mini . The GGUFs point back to the original Spark Hugging Face model: Original Spark model: https://huggingface.co/0xSero/DeepSeek V4 Flash 162B Conversion source checkpoint: https://huggingface.co/0xSero/DeepSeek V4 Flash 162B codex K144 REAP Runtime/converter repo: https://github.com/antirez/ds4 Spark deployment repo: https://github.com/0xSero/deepseek spark Files File Size SHA256 : DeepSeek V4 Flash Spark Mini Q2 REAP ds4.gguf 48.98 GiB e917278028d7a9e25dfc9d04bf5848375dad7573c5aeab1720d6a83714352406 Quantization Q2 REAP ds4 : compact DS4 profile using IQ2…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy