ModernBERT base cefr all classifier This model is a fine tuned version of answerdotai/ModernBERT base on an unknown dataset. It achieves the following results on the evaluation set: Loss: 0.1620 F1: 0.9528 Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning rate: 3.6e 05 train batch size: 2 eval batch size: 3 seed: 42 gradient accumulation steps: 16 total train batch size: 32 optimizer: Use adamw torch fused with betas=(0.9,0.999) and epsilon=1e 08 and optimizer args=No additional optimizer arguments lr scheduler type: linear lr scheduler warmup ratio: 0.1 num epochs: 3 mixed precision training: Native AMP Training results Training Loss Epoch Step Validation Loss F1 : : : : : : : : : : 2.1765 1.0 13623 0.1331 0.9438 1.604 2.0 27246 0.1180 0.9503 1.1179 2.9998 40866 0.1620 0.9528 Framework versions Transformers 4.51.0 Pytorch 2.5.1+cu121 Datasets 3.1.0 Tokenizers 0.21.0
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy