🔎 KoE5 Introducing KoE5, a model with advanced retrieval abilities. It has shown remarkable performance in Korean text retrieval. For details, visit the KURE repository Model Versions Model Name Dimension Sequence Length Introduction : : : : : : : : KURE v1 1024 8192 Fine tuned BAAI/bge m3 with Korean data via CachedGISTEmbedLoss KoE5 1024 512 Fine tuned intfloat/multilingual e5 large with ko triplet v1.0 via CachedMultipleNegativesRankingLoss Model Description This is the model card of a 🤗 transformers model that has been pushed on the Hub. Developed by: NLP&AI Lab Language(s) (NLP): Korean, English License: MIT Finetuned from model: intfloat/multilingual e5 large Finetuned dataset: ko triplet v1.0 Example code Install Dependencies First install the Sentence Transformers library: Python code Then you can load this model and run inference. Training Details Training Data ko triplet v1.0 Korean query document hard negative data pair (open data) About 700000+ examples used totally Training Procedure loss: Used CachedMultipleNegativesRankingLoss by sentence transformers batch size: 512 learning rate: 1e 05 epochs: 1 Evaluation Metrics Recall, Precision, NDCG, F1 Benchmark Datasets Ko…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy