Updated: Aug 14, 2026
base modelDownload
2 variants available
bf16 SafeTensor
ltx-2.5-22b-dev-transformer-bf16.safetensors
BF16, good balance • 39.13 GB
Verified: 2 days ago
SafeTensor
int8
ltx-2.5-22b-dev-transformer-comfy-int8-convrot.safetensors
8-bit integer, smaller file
Verified: 3 days ago
gemma4-12b-with-proj-ltx-2.5-comfy-int8-convrot.safetensors
.safetensors • 14.32 GB
Verified: 3 days ago
ltx-2.5-latent-temporal-upscaler-x2-bf16-1.0.safetensors
.safetensors • 249.81 MB
Verified: 3 days ago
5720 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
00 1 2 3 4 5 6 7 8 9
(58)
Aug 11, 2026

17.3K0 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K
1.4M0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9M
9.9M0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9M
LTX Video 2.5 and its derivatives, including LoRAs and fine-tunes, are licensed by Lightricks Ltd. under the LTX-2.x Community License Agreement and must be redistributed under that same agreement, with a copy included. Use is subject to the use restrictions in its Attachment A. Entities with annual revenues of at least $10,000,000 must obtain a paid commercial license from Lightricks before any commercial use.
LTX-2.5 is an open world model with open weights, built for local execution and fine-tuning. Its established use is generating synchronized, high-fidelity video and audio from text, image, and video inputs; applicability to emerging domains such as robotics and physical AI is developing.
Full control and customization — self-host on your own infrastructure. No per-generation billing, no per-seat lock-in, no forced API dependency. Revenue is measured across the whole entity, including subsidiaries and affiliates under common control. The full, binding terms live in LICENSE.
What's new in LTX-2.5
Native multishot generation — generate connected scenes in a single pass: multiple shots that hold character identity, environment, lighting, voice, and visual style across cuts (previous versions produced a single continuous shot).
Diffusion fidelity rendering — Instead of locking every scene to one compression rate, our model dynamically allocates compute by scene complexity and budget, rendering flawless detail where it matters, efficient everywhere else.
New diffusion video decoder — replaces the VAE reconstruction stage; sharper faces, textures, and on-screen text, better motion, and fewer artifacts in demanding scenes.
Custom Gemma 4 12B text encoder — holds complex prompts together (multiple characters, camera moves, lighting, actions) instead of dropping details across a longer sequence.
Prompt enhancer — expands a short prompt into richer cinematic instructions at minimal extra compute.
Duration predictor (optional) — an opt-in node predicts a clip's length from the prompt and sets the frame count for you, instead of relying on a fixed-duration parameter.
Substantially improved distilled model — retains much more of the full model's visual quality, prompt adherence, and motion consistency in a smaller, faster checkpoint.
