llama 3 8b instruct simpo This model is a fine tuned version of meta llama/Meta Llama 3 8B Instruct on the princeton nlp/llama3 ultrafeedback dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning rate: 1e 06 train batch size: 2 eval batch size: 4 seed: 42 distributed type: multi GPU num devices: 16 gradient accumulation steps: 8 total train batch size: 256 total eval batch size: 64 optimizer: Adam with betas=(0.9,0.999) and epsilon=1e 08 lr scheduler type: cosine lr scheduler warmup ratio: 0.1 num epochs: 1 Training results Framework versions Transformers 4.41.2 Pytorch 2.3.1+rocm6.0 Datasets 2.19.2 Tokenizers 0.19.1
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy