Arabic SBERT 100K This is a sentence transformers model finetuned from aubmindlab/bert base arabertv02. It maps sentences & paragraphs to a 768 dimensional dense vector space and can be used for semantic textual similarity, semantic search, paraphrase mining, text classification, clustering, and more. This model is trained on 100K samples filtered from the akhooli/arabic triplets 1m curated sims len dataset with 75K training and 25K validation. Trained for 5 epochs, with final training loss of 0.133 (using MatryoshkaLoss). The rest of this file is auto generated. ======================================================================== Model Details Model Description Model Type: Sentence Transformer Base model: aubmindlab/bert base arabertv02 Maximum Sequence Length: 512 tokens Output Dimensionality: 768 tokens Similarity Function: Cosine Similarity Model Sources Documentation: Sentence Transformers Documentation Repository: Sentence Transformers on GitHub Hugging Face: Sentence Transformers on Hugging Face Full Model Architecture Usage Direct Usage (Sentence Transformers) First install the Sentence Transformers library: Then you can load this model and run inference. Click to see t…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy