Sign In

LTX Video

Text/Image-to-VideoBy LightricksOpen weights

LTX Video (LTXV) is an open-weight video model from Lightricks built around a diffusion transformer tuned for speed — it generates coherent, cinematic clips from a text prompt or a starting image at close to real-time rates. The family spans the light, fast 2B model up to the latest LTX Video 2.3. Generate video right here on Civitai — no GPU, no install.

705+LTX Video models
371K+Videos generated
321+LTX Video LoRAs

About LTX Video

LTX Video (LTXV) comes from Lightricks, the team behind the LTX-Video research line. It is built on a diffusion transformer (DiT) rather than the U-Net used by the SD/SDXL lineage, and it was the first DiT-based video model able to generate high-quality clips in real time — the original release produces 24 FPS video faster than it can be watched. Like other transformer models it reads long, natural-language descriptions through a T5 text encoder, and it handles both text-to-video and image-to-video from a single starting frame.

The family has grown from a fast, lightweight base into a full audio-visual foundation model. The 2B model (0.9.x) is the light, near-real-time option; LTX Video 2 (19B) is a joint audio-video model that generates synchronized video and audio together within one network; and LTX Video 2.3 is a significant update to LTX-2 with improved audio and visual quality and stronger prompt adherence. All are open-weight and built for practical local execution, and the ecosystem has already spawned community fine-tunes — such as the uncensored Sulphur 2, built directly on the LTX 2.3 pipeline.

Choose LTXV when iteration speed matters most: its diffusion-transformer design turns prompts and starting images into coherent clips far faster than heavier video models, and newer versions add synchronized audio in the same pass. For the deepest LoRA library and the most detailed motion control, the Wan ecosystem still leads; Hunyuan Video leans toward cinematic realism, and Kling is a closed cloud API. Because even the larger LTXV models are engineered for efficiency, they generally cost less Buzz per second of video than the heaviest open alternatives.

How to prompt LTX Video

  • Write in natural language, not tags — LTXV uses a T5 text encoder, and moderate detail of two to four sentences works best. A reliable template is: [Subject and action]. [Setting]. [Camera movement]. [Lighting and style].
  • Describe subject movement and camera movement separately, and always include a camera cue — e.g. "camera panning slowly to the right," "slow zoom in," or "static wide shot." Missing camera direction is the most common weak spot in LTXV prompts.
  • Skip weight syntax — (word:1.5) and similar emphasis markers are not used. Say what you want directly instead.
  • Negative prompts are supported (applied via CFG). A solid default is "worst quality, blurry, jittery, distorted, watermark, low resolution, inconsistent motion."
  • Keep clips short and describe one continuous action. Packing many sequential events into a short clip hurts temporal consistency.

AI models move fast — new versions ship often, and a model’s capabilities or Buzz cost can change. For the latest, check the model’s own page before you generate.

Featured LTX Video models

Curated — the models worth generating with first.

Checkpoint
Lightricks LTX Video 2.3

2K

9

Civitai-hosted · default

Checkpoint
Lightricks LTX Video 2 (19B)

4K

2K

Civitai-hosted · higher fidelity

Checkpoint
Lightricks LTXV 2B (0.9.5)

11K

225K

Civitai-hosted · fast / lightweight

Checkpoint
Sulphur 2 Base

6K

94K

Civitai-hosted · LTXV 2.3 fine-tune

Popular LTX Video LoRAs & add-ons

Top LoRAs by downloads — live data, refreshed daily. Stack them on any checkpoint.

Example videos

Curated, safe-for-work showcase — every clip ships with its prompt and settings.

Cinematic medium shot of a bartender in a posh 1950s bar picking up an old telephone
LTX Video 2.3 · 1280×704 · 24fps
Cinematic full shot of a young businessman presenting beside a chart with a red downward curve
LTX Video 2.3 · 1280×704 · 24fps
A man in a hoodie and cap pushing through a crowd in a busy subway station
LTX Video 2.3 · 1280×704 · 24fps
Cinematic close-up of a young woman seated in a dark confession booth, side view
LTX Video 2.3 · 1280×704 · 24fps
An elderly bishop preaching from the pulpit of a gothic cathedral in rich vestments
LTX Video 2.3 · 1280×704 · 24fps
A wizard seated on a stone floor at the center of a chalk pentagram, candlelit
LTX Video 2.3 · 1280×704 · 24fps

How to run LTX Video

Two paths — one takes ten seconds, one takes an afternoon.

🖥️ Run it locally

For power users who want full control.

Full control over the workflow
Batch and automate
Needs a ~12GB+ VRAM (2B model) GPU
Download ~6GB (2B) of weights
Set up ComfyUI yourself

No graphics card? The Civitai path above skips all of this.

LTX Video vs other ecosystems

Backed by Civitai usage data.

FeatureLTX VideoWanHunyuan VideoKling
Best forFast, near real-time videoDetailed motion & LoRAsCinematic realismPolished cinematic clips
Prompt adherenceVery goodExcellentVery goodExcellent
Image-to-videoYesYesYesYes
Speed on CivitaiFastest (near real-time)MediumMediumAPI (cloud)
LoRA ecosystem500+LargeSmallNone (closed)
Available on Civitai✓ Yes✓ Yes✓ YesPaid API

Frequently asked questions

What is LTX Video best at?

LTX Video is built for speed — its diffusion-transformer design renders coherent clips close to real time, making it ideal for fast iteration on text- and image-to-video. Try it in the Civitai generator.

How much does it cost to generate with LTX Video?

Generation on Civitai runs on Buzz. You can claim free Blue Buzz every day — through actions like reacting to images and other on-site activity — and put it straight toward generating, no real money required. LTX Video is engineered for efficiency, so it tends to cost less Buzz per clip than heavier video models: the light 2B model stretches your daily Blue Buzz especially far, while the larger 19B and 2.3 versions cost a bit more per generation. For heavier use, let your Blue Buzz accumulate or add a membership for higher limits.

What's the difference between the LTXV versions?

The 2B model is the lightest and fastest; LTX Video 2 (19B) and LTX Video 2.3 trade some speed for higher fidelity and better motion. All are hosted on Civitai — pick any from the generator and compare.

Can I use LoRAs and image-to-video with LTX Video?

Yes. Civitai hosts 500+ LTX Video LoRAs for styles, camera moves, and characters, and LTXV supports starting from an image. Remix an example to see how the settings carry over.

Do I need a GPU to run LTX Video?

Not on Civitai — we run the compute for you. Locally, the 2B model is unusually light for video and runs in ComfyUI on a consumer GPU, while the larger 19B/2.3 models want more VRAM.

Start generating with LTX Video now

No installation. No GPU. Runs in your browser.

Want more daily generations and a faster queue? Explore Civitai membership.