The text embedding set trained by Jina AI . Quick Start The easiest way to starting using jina embeddings v2 base zh is to use Jina AI's Embedding API. Intended Usage & Model Info jina embeddings v2 base zh is a Chinese/English bilingual text embedding model supporting 8192 sequence length . It is based on a BERT architecture (JinaBERT) that supports the symmetric bidirectional variant of ALiBi to allow longer sequence length. We have designed it for high performance in mono lingual & cross lingual applications and trained it specifically to support mixed Chinese English input without bias. Additionally, we provide the following embedding models: jina embeddings v2 base zh 是支持中英双语的 文本向量 模型,它支持长达 8192字符 的文本编码。 该模型的研发基于BERT架构(JinaBERT),JinaBERT是在BERT架构基础上的改进,首次将ALiBi应用到编码器架构中以支持更长的序列。 不同于以往的单语言/多语言向量模型,我们设计双语模型来更好的支持单语言(中搜中)以及跨语言(中搜英)文档检索。 除此之外,我们也提供其它向量模型: jina embeddings v2 small en : 33 million parameters. jina embeddings v2 base en : 137 million parameters. jina embeddings v2 base zh : 161 million parameters Chinese English Bilingual embeddings (you are here) . jina embeddings v2 base de : 161 million parameters German English Bilingual embeddings. jina embeddings v2 base es : Sp…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy