The original Llama 3.3 70B Instruct model quantized using AutoAWQ. Follow the instruction here. vLLM serve Benchmark
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy