llm jp 4 32b a3b thinking LLM jp 4 is a series of large language models developed by the Research and Development Center for Large Language Models at the National Institute of Informatics. This repository provides the llm jp 4 32b a3b thinking model. For an overview of the LLM jp 4 models across different parameter sizes, please refer to: LLM jp 4 Models Base models are trained with pre training and mid training only. Post trained models are aligned using supervised fine tuning (SFT) and direct preference optimization (DPO), without reinforcement learning. For practical usage examples and detailed instructions on how to use the models, please also refer to our cookbook. To support the continued development of LLM jp, we would greatly appreciate it if you could share how you utilize LLM jp outcomes via the survey form. Usage Please refer to our cookbook for practical usage examples and detailed instructions on how to use the models. Model Details Model type: Transformer based Language Model Architectures: Dense model: Params Layers Hidden size Heads Context length Embedding parameters Non embedding parameters Total parameters : : : : : : : : : : : : : : : : 8B 32 4,096 32 65,536 805…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy