Model Overview Description: The NVIDIA gpt oss 120b is an Eagle head variant of OpenAI’s gpt oss 120b model, an autoregressive language model that uses a mixture of experts (MoE) architecture with 5 billion activated parameters and 120 billion total parameters. For more information, please check here. The NVIDIA gpt oss 120b Eagle3 model incorporates Eagle speculative decoding with Model Optimizer. This model is ready for commercial/non commercial use. License/Terms of Use: Use of these model weights is governed by the nvidia open model license. Additional Information: Apache License 2.0. Deployment Geography: Global Use Case: Developers designing AI Agent systems, chatbots, RAG systems, and other AI powered applications. Also suitable for typical instruction following tasks. This model improves the accuracy over all of gpt oss 120b Eagle3 short context, gpt oss 120b Eagle3 long context, and gpt oss 120b Eagle3 throughput. This model is recommended for all use cases where one of the previous models is used. Release Date: Hugging Face May 8 2026 via https://huggingface.co/nvidia/gpt oss 120b Eagle3 v3 Model Architecture: Architecture Type: Transformer Network Architecture: gpt oss 1…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy