Model Overview Description: The NVIDIA Llama 3.1 8B Instruct FP4 model is the quantized language model of the Meta's Llama 3.1 8B model, which is an auto regressive language model. For more information, please check here. This model is ready for commercial and non commercial use. Third Party Community Consideration This model is not owned or developed by NVIDIA. This model has been developed and built to a third party’s requirements for this application and use case; see link to Non NVIDIA (Llama 3.1 8B Instruct) Model Card. License/Terms of Use: GOVERNING TERMS: Use of this model is governed by the NVIDIA Open Model License. ADDITIONAL INFORMATION: Llama3 Community License Agreement. Built with Llama. Deployment Geography: Global, except in European Union Use Case: Developers looking to take off the shelf pre quantized models for deployment in AI Agent systems, chatbots, RAG systems, and other AI powered applications. Release Date: Huggingface 09/15/2025 via [https://huggingface.co/nvidia/Llama 3.1 8B Instruct FP4] Model Architecture: Architecture Type: Transformers Network Architecture: Llama3 This model was developed based on Llama3.1 8B Instruct Number of model parameters 8.0 1…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy