LTX 2.3 NVFP4 Model Card This is the NVFP4 versions of the LTX 2.3 model. All information below is derived from the base model. This model card focuses on the LTX 2.3 model, which is a significant update to the LTX 2 model with improved audio and visual quality as well as enhanced prompt adherence. LTX 2 was presented in the paper LTX 2: Efficient Joint Audio Visual Foundation Model. 💻💻 If you want to dive in right to the code it is available here. 💾💾 LTX 2.3 is a DiT based audio video foundation model designed to generate synchronized video and audio within a single model. It brings together the core building blocks of modern video generation, with open weights and a focus on practical, local execution. Model Checkpoints Name Notes ltx 2.3 22b dev nvfp4 The full model, flexible and trainable, in nvfp4, trained by Quantization Aware Distillation for improved accuracy ltx 2.3 22b distilled nvfp4 (coming soon) The distilled version of the full model, 8 steps, CFG=1, in nvfp4 Model Details Developed by: Lightricks Model type: Diffusion based audio video foundation model Language(s): English Online demo LTX 2.3 is accessible right away via the API Playground. Run locally Direct use…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy