Llamacpp imatrix Quantizations of Pantheon RP 1.5 12b Nemo Using llama.cpp release b3509 for quantization. Original model: https://huggingface.co/Gryphe/Pantheon RP 1.5 12b Nemo All quants made using imatrix option with dataset from here Run them in LM Studio Prompt format No prompt format found, check original model page Download a file (not the whole branch) from below: Filename Quant type File Size Split Description Pantheon RP 1.5 12b Nemo f32.gguf f32 49.00GB false Full F32 weights. Pantheon RP 1.5 12b Nemo Q8 0.gguf Q8 0 13.02GB false Extremely high quality, generally unneeded but max available quant. Pantheon RP 1.5 12b Nemo Q6 K L.gguf Q6 K L 10.38GB false Uses Q8 0 for embed and output weights. Very high quality, near perfect, recommended . Pantheon RP 1.5 12b Nemo Q6 K.gguf Q6 K 10.06GB false Very high quality, near perfect, recommended . Pantheon RP 1.5 12b Nemo Q5 K L.gguf Q5 K L 9.14GB false Uses Q8 0 for embed and output weights. High quality, recommended . Pantheon RP 1.5 12b Nemo Q5 K M.gguf Q5 K M 8.73GB false High quality, recommended . Pantheon RP 1.5 12b Nemo Q5 K S.gguf Q5 K S 8.52GB false High quality, recommended . Pantheon RP 1.5 12b Nemo Q4 K L.gguf Q4 K L…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy