EuroBERT 210m Table of Contents 1. Overview 2. Usage 3. Evaluation 4. License 5. Citation Overview EuroBERT is a family of multilingual encoder models designed for a variety of tasks such as retrieval, classification and regression supporting 15 languages, mathematics and code, supporting sequences of up to 8,192 tokens. EuroBERT models exhibit the strongest multilingual performance across domains and tasks compared to similarly sized systems. It is available in 3 sizes: EuroBERT 210m 210 million parameters EuroBERT 610m 610 million parameters EuroBERT 2.1B 2.1 billion parameters For more information about EuroBERT, please check our blog post and the arXiv preprint. Usage 💻 You can use these models directly with the transformers library starting from v4.48.0: 🏎️ If your GPU supports it, we recommend using EuroBERT with Flash Attention 2 to achieve the highest efficiency. To do so, install Flash Attention 2 as follows, then use the model as normal: Evaluation We evaluate EuroBERT on a suite of tasks to cover various real world use cases for multilingual encoders, including retrieval performance, classification, sequence regression, quality estimation, summary evaluation, code rela…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy