jina embeddings v5 text small jina embeddings v5 text small is the fifth generation of Jina AI's multilingual embedding models, released on February 18, 2026. For a lighter alternative, see jina embeddings v5 text nano (239M parameters). Elastic Inference Service ArXiv Release Note Blog Model Overview jina embeddings v5 text small scores 71.7 average on MTEB English v2 and 67.7 on MMTEB with 677M parameters, the highest among multilingual embedding models under 1B parameters. Built on Qwen3 0.6B Base and trained by combining embedding distillation from Qwen3 Embedding 4B with task specific contrastive losses, it supports 119+ languages with up to 32K tokens and produces embeddings robust under truncation and binary quantization. It is part of the jina embeddings v5 text model family, which also includes jina embeddings v5 text nano, a smaller model for resource constrained use cases. Feature Value Parameters 677M Supported Tasks retrieval , text matching , clustering , classification Max Sequence Length 32768 Embedding Dimension 1024 Matryoshka Dimensions 32, 64, 128, 256, 512, 768, 1024 Pooling Strategy Last token pooling Base Model Qwen/Qwen3 0.6B Base Training and Evaluation For…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy