jina embeddings v5 omni nano : Multi Task Omni Embedding Base (Nano) ArXiv Blog Average score vs. parameter count across image (MIEB Lite), video (MMEB V), and audio (MAEB) benchmarks — jina v5 omni nano and jina v5 omni small define the open weight frontier (Table 1 in the ArXiv report). Model Overview jina embeddings v5 omni nano is a multimodal embedding model that accepts text, images, video, and audio and produces embeddings in a shared vector space aligned with the text only jinaai/jina embeddings v5 text nano — so you can index with text and query with any modality, no reindexing. For higher performance at a larger size, see jinaai/jina embeddings v5 omni small . This is the base repository — it holds all task adapters (retrieval, classification, clustering, text matching). For a single task, pre merged task specific variants are also available: jinaai/jina embeddings v5 omni nano retrieval — query–document semantic search and RAG (raw transformers users prepend Query: / Document: to text; sentence transformers users call encode query() / encode document() ). jinaai/jina embeddings v5 omni nano classification — assigning labels via embedding similarity — zero shot and few sh…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy