Qwen3 VL 235B A22B Thinking UD Q4 K XL Distributed GGUF inference package for Mesh LLM GGUF layer package for running Qwen3 VL 235B A22B Thinking UD Q4 K XL across a local Mesh LLM cluster. This package is derived from unsloth/Qwen3 VL 235B A22B Thinking GGUF and keeps the original GGUF distribution split into per layer artifacts for distributed inference. Highlights Run locally Pool multiple machines OpenAI compatible Package variant Private inference on your hardware Split layers across peers Serve /v1/chat/completions locally UD Q4 K XL layer package Model Overview Property Value Source model unsloth/Qwen3 VL 235B A22B Thinking GGUF Model id unsloth/Qwen3 VL 235B A22B Thinking GGUF:UD Q4 K XL Family Qwen3 Parameter scale 235B A22B Quantization UD Q4 K XL Layer count 94 Activation width 4096 Package size 125.5 GB Source file UD Q4 K XL/Qwen3 VL 235B A22B Thinking UD Q4 K XL 00001 of 00003.gguf Package repo meshllm/Qwen3 VL 235B A22B Thinking UD Q4 K XL layers Recommended Use Local and private inference with Mesh LLM. Multi machine serving when the full GGUF is too large for one host. OpenAI compatible chat/completions workflows through Mesh LLM's local API. For upstream architect…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy