jina embeddings v5 text small text matching : Text Matching Targeted Embedding Distillation Elastic Inference Service ArXiv Release Note Blog Model Overview jina embeddings v5 text small text matching is a compact, high performance text embedding model designed for text matching. It is part of the jina embeddings v5 text model family, which also includes jina embeddings v5 text nano, a smaller model for more resource constrained use cases. Trained using a novel approach that combines distillation with task specific contrastive losses, jina embeddings v5 text small text matching outperforms existing state of the art models of similar size across diverse embedding benchmarks. Feature Value Parameters 677M Supported Tasks text matching Max Sequence Length 32768 Embedding Dimension 1024 Matryoshka Dimensions 32, 64, 128, 256, 512, 768, 1024 Pooling Strategy Last token pooling Base Model jinaai/jina embeddings v5 text small Training and Evaluation For training details and evaluation results, see our technical report. Usage Requirements The following Python packages are required: transformers =5.1.0 torch =2.8.0 peft =0.15.2 vllm =0.15.1 Optional / Recommended flash attention : Installin…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy