Octen Embedding 8B Octen Embedding 8B is a text embedding model developed by Octen for semantic search and retrieval tasks. This model is fine tuned from Qwen/Qwen3 Embedding 8B and supports multiple languages, providing high quality embeddings for various applications. Key Highlights 🥇 RTEB Leaderboard Champion (as of January 12, 2026) Octen Embedding 8B ranks 1 on the RTEB Leaderboard with Mean (Task) score of 0.8045 Excellent performance on both Public (0.7953) and Private (0.8157) datasets Demonstrates true generalization capability without overfitting to public benchmarks Industry Oriented Vertical Domain Expertise Legal : Legal document retrieval Finance : Financial reports, Q&A, and personal finance content Healthcare : Medical Q&A, clinical dialogues, and health consultations Code : Programming problems, code search, and SQL queries Ultra Long Context Support Supports up to 32,768 tokens context length Suitable for processing long documents in legal, healthcare, and other domains High dimensional embedding space for rich semantic representation Multilingual Capability Supports 100+ languages Includes various programming languages Strong multilingual, cross lingual, and cod…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy