GLM 5.2 Q2 K MTP Q8 Distributed GGUF inference package for Mesh LLM GGUF layer package for running GLM 5.2 Q2 K MTP Q8 across a local Mesh LLM cluster. This package is derived from meshllm/GLM 5.2 Q2 K MTP Q8 GGUF and keeps the original GGUF distribution split into per layer artifacts for distributed inference. Highlights Run locally Pool multiple machines OpenAI compatible Package variant Private inference on your hardware Split layers across peers Serve /v1/chat/completions locally Q2 K layer package Model Overview Property Value Source model meshllm/GLM 5.2 Q2 K MTP Q8 GGUF Model id meshllm/GLM 5.2 Q2 K MTP Q8 GGUF:Q2 K MTP Q8 Family GLM Parameter scale not recorded Quantization Q2 K Layer count 79 Activation width 6144 Package size 260.3 GB Source file Q2 K MTP Q8/GLM 5.2 Q2 K MTP Q8 00001 of 00306.gguf Package repo meshllm/GLM 5.2 Q2 K MTP Q8 layers Recommended Use Local and private inference with Mesh LLM. Multi machine serving when the full GGUF is too large for one host. OpenAI compatible chat/completions workflows through Mesh LLM's local API. For upstream architecture details, chat template guidance, sampling recommendations, license terms, and benchmark notes, see the so…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy