jina embeddings v5 text nano jina embeddings v5 text nano is the fifth generation of Jina AI's multilingual embedding models, released on February 18, 2026. For higher performance at a larger size, see jina embeddings v5 text small. Elastic Inference Service ArXiv Release Note Blog Model Overview jina embeddings v5 text nano scores 71.0 average on MTEB English v2 and 65.5 on MMTEB with only 239M parameters, matching or exceeding all other sub 500M embedding models including KaLM mini v2.5 (494M) and Gemma 300M (308M). Built on EuroBERT 210M and trained by combining embedding distillation from Qwen3 Embedding 4B with task specific contrastive losses, it supports multilingual text up to 32K tokens and produces embeddings robust under truncation and binary quantization. Feature Value Parameters 239M Supported Tasks retrieval , text matching , clustering , classification Max Sequence Length 8192 Embedding Dimension 768 Matryoshka Dimensions 32, 64, 128, 256, 512, 768 Pooling Strategy Last token pooling Base Model EuroBERT/EuroBERT 210m Training and Evaluation For training details and evaluation results, see our technical report. Usage Requirements The following Python packages are requ…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy