gte Qwen2 7B instruct gte Qwen2 7B instruct is the latest model in the gte (General Text Embedding) model family that ranks No.1 in both English and Chinese evaluations on the Massive Text Embedding Benchmark MTEB benchmark (as of June 16, 2024). Recently, the Qwen team released the Qwen2 series models, and we have trained the gte Qwen2 7B instruct model based on the Qwen2 7B LLM model. Compared to the gte Qwen1.5 7B instruct model, the gte Qwen2 7B instruct model uses the same training data and training strategies during the finetuning stage, with the only difference being the upgraded base model to Qwen2 7B. Considering the improvements in the Qwen2 series models compared to the Qwen1.5 series, we can also expect consistent performance enhancements in the embedding models. The model incorporates several key advancements: Integration of bidirectional attention mechanisms, enriching its contextual understanding. Instruction tuning, applied solely on the query side for streamlined efficiency Comprehensive training across a vast, multilingual text corpus spanning diverse domains and scenarios. This training leverages both weakly supervised and supervised data, ensuring the model's ap…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy