LTX 2.3 Model Card This model card focuses on the LTX 2.3 model, which is a significant update to the LTX 2 model with improved audio and visual quality as well as enhanced prompt adherence. LTX 2 was presented in the paper LTX 2: Efficient Joint Audio Visual Foundation Model. 💻💻 If you want to dive in right to the code it is available here. 💾💾 LTX 2.3 is a DiT based audio video foundation model designed to generate synchronized video and audio within a single model. It brings together the core building blocks of modern video generation, with open weights and a focus on practical, local execution. Model Checkpoints Name Notes ltx 2.3 22b dev The full model, flexible and trainable in bf16 ltx 2.3 22b distilled The distilled version of the full model, 8 steps, CFG=1 ltx 2.3 22b distilled 1.1 The distilled v1.1 version of the full model, 8 steps, CFG=1 A different aesthetic experience and improved audio compared to v1.0 ltx 2.3 22b distilled lora 384 A LoRA version of the distilled model applicable to the full model ltx 2.3 22b distilled lora 384 1.1 A LoRA version of the v1.1 distilled model applicable to the full model ltx 2.3 spatial upscaler x2 1.1 An x2 spatial upscaler for t…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy