LTX 2 Model Card This model card focuses on the LTX 2 model, as presented in the paper LTX 2: Efficient Joint Audio Visual Foundation Model. The codebase is available here. LTX 2 is a DiT based audio video foundation model designed to generate synchronized video and audio within a single model. It brings together the core building blocks of modern video generation, with open weights and a focus on practical, local execution. Model Checkpoints Name Notes ltx 2 19b dev The full model, flexible and trainable in bf16 ltx 2 19b dev fp8 The full model in fp8 quantization ltx 2 19b dev fp4 The full model in nvfp4 quantization ltx 2 19b distilled The distilled version of the full model, 8 steps, CFG=1 ltx 2 19b distilled lora 384 A LoRA version of the distilled model applicable to the full model ltx 2 spatial upscaler x2 1.0 An x2 spatial upscaler for the ltx 2 latents, used in multi stage (multiscale) pipelines for higher resolution ltx 2 temporal upscaler x2 1.0 An x2 temporal upscaler for the ltx 2 latents, used in multi stage (multiscale) pipelines for higher FPS Model Details Developed by: Lightricks Model type: Diffusion based audio video foundation model Language(s): English Online…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy