See our collection for all versions of gpt oss including GGUF, 4 bit & 16 bit formats. Learn to run gpt oss correctly Read our Guide . See Unsloth Dynamic 2.0 GGUFs for our quantization benchmarks. ✨ Read our gpt oss Guide here ! Read our Blog about gpt oss support: unsloth.ai/blog/gpt oss View the rest of our notebooks in our docs here. Thank you to the llama.cpp team for their work on supporting this model. We wouldn't be able to release quants without them! gpt oss 20b Details Try gpt oss · Guides · System card · OpenAI blog Welcome to the gpt oss series, OpenAI’s open weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases. We’re releasing two flavors of the open models: gpt oss 120b — for production, general purpose, high reasoning use cases that fits into a single H100 GPU (117B parameters with 5.1B active parameters) gpt oss 20b — for lower latency, and local or specialized use cases (21B parameters with 3.6B active parameters) Both models were trained on our harmony response format and should only be used with the harmony format as it will not work correctly otherwise. [!NOTE] This model card is dedicated to the smaller gpt oss 20b mo…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy