Update (2026 06 25): A Transformers native version of Xcodec2 has been released here! Update (2025 02 13): Add Llasa finetune instruction. Update (2025 02 07): Our paper has been released! Paper LLaSA: Scaling Train Time and Inference Time Compute for LLaMA based Speech Synthesis Codec Does Matter: Exploring the Semantic Shortcoming of Codec for Audio Language Model (AAAI 2025, xcodec 1.0) Getting Started with XCodec2 on Hugging Face XCodec2 is a speech tokenizer that offers the following key features: 1. Single Vector Quantization 2. 50 Tokens per Second 3. Multilingual Speech Semantic Support and High Quality Speech Reconstruction To use xcodec2 , ensure you have it installed. You can install it using the following command: Then, If you want to train your own xcodec2, batch inference, or large scale code extraction, the code is released here.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy