Model Overview Description: The NVIDIA gpt oss 120b Eagle model is the Eagle head of the OpenAI’s gpt oss 120b model, which is an auto regressive language model that uses a mixture of experts (MoE) architecture with 5 billion activated parameters and 120 billion total parameters. For more information, please check here. The NVIDIA gpt oss 120b Eagle3 model incorporates Eagle speculative decoding with Model Optimizer. This model is ready for commercial/non commercial use. For use cases of less than 8k context length please consider using gpt oss 120b Eagle3 short context License/Terms of Use: nvidia open model license Apache License 2.0 Deployment Geography: Global Use Case: Developers designing AI Agent systems, chatbots, RAG systems, and other AI powered applications. Also suitable for typical instruction following tasks. Release Date: Huggingface: Aug 20th, 2025 via [https://huggingface.co/nvidia/gpt oss 120b Eagle3 long context] Model Architecture: Architecture Type: Transformers Network Architecture: gpt oss 120b Computational Load Cumulative Compute: 4.8x10^20 Estimated Energy and Emissions for Model Training: Total kWh = 2500 Total Emissions (tCO2e) = 0.8075 Input: Input Type…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy