Trained by Jina AI . JinaColBERT V2: A General Purpose Multilingual Late Interaction Retriever. JinaColBERT V2 ( jina colbert v2 ) is a new model based on the JinaColBERT V1 that expands on the capabilities and performance of the jina colbert v1 en model. Like the previous release, it has Jina AI’s 8192 token input context and the improved efficiency, performance, and explainability of token level embeddings and late interaction. This new release adds new functionality and performance improvements: Multilingual support for dozens of languages, with strong performance on major global languages. Matryoshka embeddings, which allow users to trade between efficiency and precision flexibly. Superior retrieval performance when compared to the English only jina colbert v1 en . JinaColBERT V2 offers three different versions for different embeddings dimensions: jinaai/jina colbert v2 : 128 dimension embeddings jinaai/jina colbert v2 96 : 96 dimension embeddings jinaai/jina colbert v2 64 : 64 dimension embeddings Usage Installation jina colbert v2 is trained with flash attention and therefore requires einops and flash attn to be installed. To use the model, you could either use the Standford…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy