// AI TOOLS · LTX-2.5

LTX-2.5

LTX's open-weight video model — spun out of Lightricks — that turns an image into a 10-second clip in seconds, and runs on a GPU you already own.

Last verified · 2026-08-12 · by Moe Ameen

What LTX-2.5 is

LTX-2.5 is an open-weight video and world model released on August 11, 2026 by LTX, the open-model company spun out of Lightricks. Its headline claim is speed with ownership: on two NVIDIA GB200 chips it generates a 10-second, 720p image-to-video clip in about 6.8 seconds — faster than the clip itself runs — and through LTX's managed API it returns a 1080p clip in roughly 23.7 seconds. It does both text-to-video and image-to-video, and it keeps the synchronized native audio LTX introduced in the 2.3 generation, so a render can arrive with matching sound rather than as a silent file.

What sets LTX-2.5 apart from the closed frontier is that the weights are open. The model is free to use for organizations under $10 million in annual recurring revenue, with larger companies negotiating a license, and it ships day-one on Hugging Face, inside ComfyUI, and through fal.ai and the LTX API. LTX tuned it to run on hardware creators already have: it needs roughly 16GB of VRAM at minimum and is optimized for local NVIDIA RTX GPUs and the NVIDIA DGX Spark desktop, with a distilled variant for faster local inference. A frontier-class video model that runs on one desk instead of a rented cluster is the real story here.

On capability, LTX-2.5 adds native multishot — one generation produces several connected shots that hold character, environment, lighting, and voice across cuts, instead of you stitching separate clips — plus a new diffusion video decoder for cleaner high-motion footage, a custom Gemma-based text encoder with a prompt enhancer for stronger prompt-following, automatic clip duration, native 4K HDR support with a RAW finishing workflow, and a beta precise-editing mode. The honest framing: LTX-2.5 is a model, not a content studio. It generates a clip (with sound), and that is where it stops — there is no caption burner, no per-platform reframing, no brand or persona system, and no scheduler or publishing. Treat specific speeds, resolutions, and license terms as the launch snapshot and confirm them on ltx.io before you build against them.

What you can make with it

  • A 10-second, 720p image-to-video clip generated in seconds on local NVIDIA hardware, or 1080p through the LTX API
  • Text-to-video shots from a written prompt, with a prompt enhancer expanding short prompts into detailed instructions
  • Multishot sequences that hold one character, setting, and voice consistent across several cuts in a single generation
  • Clips with synchronized native audio — dialogue, effects, and ambient sound produced alongside the picture
  • Native 4K HDR output with a RAW workflow for a professional finishing pipeline
  • A fully local, open-weight generation pipeline inside ComfyUI on an RTX GPU or DGX Spark — no per-render cloud bill

How Kompozy turns LTX-2.5 output into content

The thing that makes LTX-2.5 special — open weights on your own GPU — is also what leaves you holding a folder of raw clips at the end of the night. Run it in [ComfyUI](/ai-tools/comfyui) on an RTX card and you can batch a week of b-roll and image-to-video shots locally, for free, while you sleep. What you wake up to is exactly the problem: a stack of 10-second MP4s with no captions, no aspect ratios, no hook on the opening frame, and no way to post them. Native multishot keeps a character consistent inside one clip, but nothing keeps Monday's render looking like Thursday's, or makes either read as your brand. [Kompozy](/) is the no-code layer that turns that local render farm into a published channel. Drop the clips in and it burns branded, on-style captions for the sound-off feed, reframes each one to 9:16, 1:1, and 16:9 per destination, and stacks a hook overlay through [HyperFrames](/glossary/hyperframes) so the muted autoplay opening lands — with the [Persona Brief](/glossary/persona-brief) enforcing one voice across every caption regardless of what prompt produced the clip.

Then it does the two things a raw model can't. It generates the formats LTX-2.5 will never make — [carousels](/glossary/output-buckets) that break a clip into beats, quote cards, face-locked persona photos, blog articles, and email newsletters — and its own [Persona Shorts](/glossary/persona-shorts) and [Persona Frames](/glossary/persona-frames) avatar video, so you aren't limited to what the model rendered. And it publishes: scheduling and fanning the finished set across the eight social platforms plus blog and email from one queue with a per-post review gate and [Autopilot](/glossary/autopilot). LTX-2.5 is the fastest way to make the clip cheaply and locally; Kompozy is the no-code way to make it a brand and ship it everywhere.

  1. Generate clips with LTX-2.5 — locally in ComfyUI on an RTX GPU or DGX Spark, or through the LTX API or fal.ai for 1080p.
  2. Bring the finished MP4s into Kompozy — no code required.
  3. Let Kompozy burn in branded captions, reframe each clip per platform, and stack a hook overlay through HyperFrames so every render matches your brand.
  4. Fan the same idea into a carousel, quote card, persona photo, blog, and newsletter, all governed by your Persona Brief.
  5. Schedule and publish the set across the eight social platforms plus blog and email from one queue with Autopilot.

Frequently asked questions

What is LTX-2.5?

LTX-2.5 is an open-weight AI video and world model released August 11, 2026 by LTX, the open-model company spun out of Lightricks. It generates a 10-second, 720p image-to-video clip in about 6.8 seconds on NVIDIA GB200 hardware, supports text-to-video and synchronized native audio, and runs locally on RTX GPUs.

Is LTX-2.5 free and open source?

The weights are open and free to use for organizations under $10 million in annual recurring revenue; larger companies negotiate a license. It is available on Hugging Face, inside ComfyUI, and through fal.ai and the LTX API. Confirm current license terms on ltx.io.

What hardware do I need to run LTX-2.5 locally?

LTX tuned it for local inference on NVIDIA RTX GPUs and the DGX Spark desktop, with a minimum of roughly 16GB of VRAM and a distilled variant for faster generation. On two GB200 chips it renders a 10-second clip in about 6.8 seconds; local speed depends on your GPU.

Does LTX-2.5 generate audio with the video?

Yes. It keeps the synchronized native audio introduced in LTX-2.3, so a clip can arrive with matching dialogue, effects, and ambient sound in the same generation rather than added afterward. It also adds native multishot for consistency across cuts.

How do I turn LTX-2.5 clips into finished, published posts?

Generate the clips in ComfyUI or via the LTX API, then bring them into Kompozy — no code needed. Kompozy adds branded captions, reframes per platform, and unifies brand styling, then fans the idea into a carousel, quote card, blog, and captions in your voice, scheduled and published across the eight social platforms plus blog and email.

Related tools

  • ComfyUIThe open-source, node-based interface for running generative AI models on your own machine — now with day-0 native support for MiniMax H3, so you can generate 2K video with native audio locally.
  • Google Veo 3Google DeepMind's video model that was the first to generate synchronized native audio — dialogue, sound effects, and music — inside the same pass as the video, with lip sync.
  • Kling AI 3.0Kuaishou's flagship Kling 3.0 model — a multi-shot "director" video model that generates a scripted sequence with native audio and native 4K/60fps in a single pass, plus 2K/4K images.
  • ByteDance Seedance 2.5AI video model that generates a 30-second clip in one pass — no stitching.
  • Hailuo AI Video GeneratorMiniMax's AI video generator, known for physically believable motion and strong instruction following from text or a single image.
  • RunwayThe AI video platform behind the Lionsgate partnership — cinematic text-, image-, and video-to-video generation with consistent characters and scenes.

← All AI tools · Get started →