Quantization made by Richard Erkhov. Github Discord Request more models gte Qwen2 7B instruct GGUF Model creator: https://huggingface.co/Alibaba NLP/ Original model: https://huggingface.co/Alibaba NLP/gte Qwen2 7B instruct/ Name Quant method Size gte Qwen2 7B instruct.Q2 K.gguf Q2 K 2.81GB gte Qwen2 7B instruct.IQ3 XS.gguf IQ3 XS 3.11GB gte Qwen2 7B instruct.IQ3 S.gguf IQ3 S 3.26GB gte Qwen2 7B instruct.Q3 K S.gguf Q3 K S 3.25GB gte Qwen2 7B instruct.IQ3 M.gguf IQ3 M 3.33GB gte Qwen2 7B instruct.Q3 K.gguf Q3 K 3.55GB gte Qwen2 7B instruct.Q3 K M.gguf Q3 K M 3.55GB gte Qwen2 7B instruct.Q3 K L.gguf Q3 K L 3.81GB gte Qwen2 7B instruct.IQ4 XS.gguf IQ4 XS 3.96GB gte Qwen2 7B instruct.Q4 0.gguf Q4 0 4.13GB gte Qwen2 7B instruct.IQ4 NL.gguf IQ4 NL 4.15GB gte Qwen2 7B instruct.Q4 K S.gguf Q4 K S 4.15GB gte Qwen2 7B instruct.Q4 K.gguf Q4 K 4.36GB gte Qwen2 7B instruct.Q4 K M.gguf Q4 K M 4.36GB gte Qwen2 7B instruct.Q4 1.gguf Q4 1 4.54GB gte Qwen2 7B instruct.Q5 0.gguf Q5 0 4.95GB gte Qwen2 7B instruct.Q5 K S.gguf Q5 K S 4.95GB gte Qwen2 7B instruct.Q5 K.gguf Q5 K 5.07GB gte Qwen2 7B instruct.Q5 K M.gguf Q5 K M 5.07GB gte Qwen2 7B instruct.Q5 1.gguf Q5 1 5.36GB gte Qwen2 7B instruct.Q6 K.gg…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy