This model was quantized by SanctumAI. To leave feedback, join our community in Discord. Meta Llama 3 8B Instruct GGUF Model creator: meta llama Original model : Meta Llama 3.1 8B Instruct Model Summary: The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multilingual dialogue use cases and outperform many of the available open source and closed chat models on common industry benchmarks. Prompt Template: If you're using Sanctum app, simply use Llama 3 model preset. Prompt template: Hardware Requirements Estimate Name Quant method Size Memory (RAM, vRAM) required meta llama 3.1 8b instruct.Q2 K.gguf Q2 K 3.18 GB 7.20 GB meta llama 3.1 8b instruct.Q3 K S.gguf Q3 K S 3.67 GB 7.65 GB meta llama 3.1 8b instruct.Q3 K M.gguf Q3 K M 4.02 GB 7.98 GB meta llama 3.1 8b instruct.Q3 K L.gguf Q3 K L 4.32 GB 8.27 GB meta llama 3.1 8b instruct.Q4 0.gguf Q4 0 4.66 GB 8.58 GB meta llama 3.1 8b instruct.Q4 K S.gguf Q4 K S 4.69 GB 8.61 GB meta llama 3.1 8b instruct.Q4 K M.gguf Q4 K M…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy