jina embeddings v4 gguf A collection of GGUF and quantizations for jina embeddings v4 . [!IMPORTANT] We highly recommend to first read this blog post for more technical details and customized llama.cpp build. [!TIP] Multimodal v4 GGUF is now available, check out this blog post for the walkthrough. Overview jina embeddings v4 is a cutting edge universal embedding model for multimodal multilingual retrieval. It's based on qwen2.5 vl 3b instruct with three LoRA adapters: retrieval (optimized for retrieval tasks), text matching (optimized for sentence similarity tasks), and code (optimized for code retrieval tasks). It is also heavily trained for visual document retrieval and late interaction style multi vector output. Text Only Task Specific Models We removed the visual components of qwen2.5 vl and merged all LoRA adapters back into the base language model. This results in three task specific v4 models with 3.09B parameters, downsized from the original jina embeddings v4 3.75B parameters: HuggingFace Repo Task jinaai/jina embeddings v4 text retrieval GGUF Text retrieval jinaai/jina embeddings v4 text code GGUF Code retrieval jinaai/jina embeddings v4 text matching GGUF Sentence simila…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy