gte Qwen2 1.5B instruct gte Qwen2 1.5B instruct is the latest model in the gte (General Text Embedding) model family. The model is built on Qwen2 1.5B LLM model and use the same training data and strategies as the gte Qwen2 7B instruct model. The model incorporates several key advancements: Integration of bidirectional attention mechanisms, enriching its contextual understanding. Instruction tuning, applied solely on the query side for streamlined efficiency Comprehensive training across a vast, multilingual text corpus spanning diverse domains and scenarios. This training leverages both weakly supervised and supervised data, ensuring the model's applicability across numerous languages and a wide array of downstream tasks. Model Information Model Size: 1.5B Embedding Dimension: 1536 Max Input Tokens: 32k Requirements Usage Sentence Transformers Observe the config sentence transformers.json to see all pre built prompt names. Otherwise, you can use model.encode(queries, prompt="Instruct: ...\nQuery: " to use a custom prompt of your choice. Transformers infinity emb Usage via infinity, MIT Licensed. Evaluation MTEB & C MTEB You can use the scripts/eval mteb.py to reproduce the followi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy