Nemotron H 8B Base 8K Model Overview NVIDIA Nemotron H 8B Base 8K is a large language model (LLM) developed by NVIDIA that is designed as a completion model for a given piece of text. It uses a hybrid model architecture that consists primarily of Mamba 2 and MLP layers combined with just four Attention layers. The model features a context length of 8K. The supported languages include: English, German, Spanish, French, Italian, Korean, Portuguese, Russian, Japanese, and Chinese. For more detailed information on the model architecture, training, and evaluation, please see the project page and the technical report. For best performance on a given task, users are encouraged to customize the model using the NeMo Framework suite of customization tools including Parameter Efficient Fine Tuning (P tuning, Adapters, LoRA, and more), and Model Alignment (SFT, SteerLM, RLHF, and more) using NeMo Aligner. This model is for research and development only. This model is part of the Nemotron H Collection. You can find the models in this family here: Nemotron H 56B Base 8K Nemotron H 47B Base 8K Nemotron H 8B Base 8K License/Terms of Use GOVERNING TERMS: Use of this model is governed by the NVIDIA…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy