LTX-2.5 Open-Weights Video Model: 4K Lives on the Fast Endpoints, and an EXR Color Pipeline Is Built In

Want 4K output from LTX-2.5? You need the endpoint labeled "Fast," not "Pro." That is not a bug: Lightricks documentation states plainly that the Fast endpoints trade some fidelity for speed, 4K output, and clips up to 20 seconds, while the quality-optimized Pro endpoints top out at 1080p and 10 seconds. This counterintuitive design is a good entry point for understanding the product philosophy of LTX-2.5: it is not competing on "bigger and stronger," but on pushing an open-weights video model into real production pipelines.

Six Endpoints, One Generation, Audio Included

LTX-2.5 is Lightricks open-weights audio-video model, released under the LTX-2.x Community License and served on the fal platform through six endpoints: text-to-video, image-to-video, and audio-to-video, each in a quality-optimized Pro variant and a speed-optimized Fast variant. Every variant generates synchronized audio by default and supports 16:9 and 9:16 framing. Image-to-video also accepts an end frame, letting users pin where a shot finishes.

On specs, Fast covers 720p/1080p/1440p/2160p (4K) at 24/25/48/50fps with clips up to 20 seconds; Pro runs 720p/1080p at 24/25/50fps with 6-, 8-, or 10-second clips. The division of labor is telling: Fast handles rapid iteration and 4K delivery, Pro handles final high-fidelity output.

The Fidelity Ledger: Where the Compute Goes Matters

LTX-2.5 introduces Diffusion Fidelity Rendering, which puts more compute into complex scenes instead of spreading it evenly across every frame — crowds, fast motion, and dense detail hold together where they would otherwise soften. The feature runs on the Pro endpoints.

According to the official FAQ, the single biggest contributor to the fidelity jump over version 2.3 is the Diffusion Video Decoder: it replaces plain VAE decoding with a diffusion-based one, so faces stay sharp, on-screen text stays legible, and fast motion shows fewer smears.

The narrative-side increment is native multishot: a single generation yields multiple connected shots that hold character, environment, lighting, voice, and style across every cut. Supporting it is a custom Gemma 4 12B text encoder that tracks multiple subjects, actions, lighting cues, and camera direction through a complex prompt, while Auto Duration reads the described action to pick the right clip length before diffusion begins.

A More Industrial Step: EXR and the Raw Checkpoint

Two details show Lightricks is aiming at professional production. First, LTX-2.5 adds a native EXR workflow that reads and writes cinema-grade EXR inside professional color spaces including ACES and DaVinci Wide Gamut, with generative edits returning EXR — no lossy 8-bit round-trip in the middle of a color pipeline.

Second, the model ships with a raw pretrained checkpoint: a non-SFT base built for aggressive adaptation toward new data and objectives — robotics, synthetic AV, industrial digital twins, private domain models. Combined with LoRA training (fal already hosts more than 80 LTX fine-tunes), "open weights" here means not just downloadable, but modifiable.

Price and License Boundaries

Pricing on fal is per second: the Fast variant costs 0.09 USD per second at 720p, 0.13 USD at 1080p, 0.19 USD at 1440p, and 0.30 USD at 4K, with native audio included at every resolution; Pro image-to-video is 0.12 USD per second at 720p and 0.17 USD at 1080p.

On licensing, a bucket of cold water: the LTX-2.x Community License is not an OSI open-source license. It permits commercial use, but entities with annual revenue of at least 10 million US dollars must obtain a paid commercial license from Lightricks first, and the license carries use restrictions — such as not training competing models.

So What

Closed-API video models compete on the wow factor of "one prompt, one blockbuster." LTX-2.5 is competing on something else: being an asset that is self-hostable, fine-tunable, pluggable into an ACES color pipeline, and runnable on your own infrastructure. Putting 4K on the Fast endpoints is, at heart, treating high resolution as a delivery format while reserving Pro for image quality itself — that is production logic, not demo logic. The next round of competition among open-weights video models may be decided not by spec sheets, but by who gets stitched into a workflow first.

Source and model details: the LTX-2.5 page on fal.ai.