BGE M3 MLX (FP16) This is the BAAI/bge m3 model converted to MLX format for Apple Silicon. Model Description BGE M3 is a versatile embedding model capable of: Dense retrieval Sparse retrieval Multi vector (ColBERT) retrieval This MLX conversion enables efficient inference on Apple Silicon Macs. Model Details Property Value Architecture XLM RoBERTa Precision FP16 (float16) Embedding Dimension 1024 Max Sequence Length 8192 Model Size ~1.1 GB Languages 100+ languages Usage With MLX With oMLX With Python (OpenAI compatible) Performance Tested on macOS with Apple Silicon: Successfully generates 1024 dimensional embeddings Supports multilingual text (English, Chinese, Japanese, Korean, etc.) Compatible with oMLX embedding endpoint Conversion Details This model was converted from the original BAAI/bge m3 using: mlx embeddings conversion tool FP16 precision for balanced size and quality License This model inherits the MIT license from the original BAAI/bge m3 model. Citation Disclaimer This is an unofficial MLX conversion of the BAAI/bge m3 model. For the original model and official implementations, please refer to BAAI/bge m3.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy