Update 2026/04/15: This new model is a complete replacement (smaller, faster and better). This Model2Vec model was created by using Tokenlearn, with nomic embed text v2 moe as a base. The output dimension is 768. The evaluation in the model card, was executed using this model (distilled), not the original. The process to create this one, was not a simple model2vec distill, this involved generating embeddings for 23M triplets (msmarco) with the original model, then training the tokenlearn model on it, with the nomic model as a base. Usage Load this model using model2vec library: Or using sentence transformers library:
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy