Lance 3B Video GGUF This repository contains the quantized GGUF weights for Lance 3B Video , originally developed by ByteDance Research. Lance is a lightweight, native unified multimodal model that supports image and video understanding, generation, and editing within a single framework. By quantizing these weights into GGUF format, the model becomes significantly more accessible for local inference, drastically reducing the massive 40GB VRAM requirement of the original unquantized model. 📊 Hardware Compatibility & Files These GGUF files are designed for efficient CPU/GPU offloading. Choose the quantization level that best fits your available memory and desired quality. Quantization Filename Size Description : : : : 4 bit Lance 3B Video Q4 K M.gguf 4.96 GB Optimal balance of speed and VRAM. Great for mid range hardware. 5 bit Lance 3B Video Q5 K M.gguf 5.53 GB Higher precision with minimal size increase. 6 bit Lance 3B Video Q6 K.gguf 6.12 GB Near unquantized quality for higher end local setups. 8 bit Lance 3B Video Q8 0.gguf 7.62 GB Maximum quality, requiring the most memory among these options. 🌟 Model Overview Rather than relying on massive parameter scaling, Lance achieves st…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy