LTX-2 is a 19B parameter DiT-based audio-video foundation model by Lightricks. It generates synchronized video and audio in a single pass, creating cohesive experiences where motion, dialogue, background noise, and music are produced together.
Key features
- Synchronized audio-video generation: Generates motion, dialogue, SFX, and music together in one pass
- Multiple generation modes: Text-to-video, image-to-video, and video-to-video
- Control options: Canny, Depth, and Pose video-to-video control via IC-LoRAs
- Keyframe-driven generation: Interpolate between keyframe images
- Native upscaling: Spatial (2x) and temporal (2x) upscalers for higher resolution and FPS
- Prompt enhancement: Automatic prompt enhancement support
Model checkpoints
Getting started
LTX-2 is natively supported in ComfyUI. To get started:- Update ComfyUI to the latest version
- Go to Template Library > Video > choose any LTX-2 workflow
- Follow the pop-up to download models and run the workflow
Workflows
Text-to-video
Generate videos from text prompts.
Text-to-Video
Download workflow
Run on Comfy Cloud
Open in cloud
Text-to-Video Distilled
Download workflow
Run on Comfy Cloud
Open in cloud
Image-to-video
Generate videos from an input image.
Image-to-Video
Download workflow
Run on Comfy Cloud
Open in cloud
Input Image
Download the default input image, or use your own image.
Image-to-Video Distilled
Download workflow
Run on Comfy Cloud
Open in cloud
Input Image
Download the default input image, or use your own image.
Control-to-video
Generate videos with structural control using IC-LoRAs. Depth control:
Depth-to-Video
Download workflow
Run on Comfy Cloud
Open in cloud
Input Image
Download the default depth input image, or use your own.
Input Video
Download the default input video, or use your own.
Canny-to-Video
Download workflow
Run on Comfy Cloud
Open in cloud
Input Image
Download the default canny input image, or use your own.
Input Video
Download the default input video, or use your own.
Pose-to-Video
Download workflow
Run on Comfy Cloud
Open in cloud
Input Image
Download the default pose input image, or use your own.
Input Video
Download the default input video, or use your own.
Prompting tips
When writing prompts for LTX-2, focus on detailed, chronological descriptions of actions and scenes. Include specific movements, appearances, camera angles, and environmental details in a single flowing paragraph. Start directly with the action and keep descriptions literal and precise. Structure your prompts with:- Main action in a single sentence
- Specific details about movements and gestures
- Character/object appearances
- Background and environment details
- Camera angles and movements
- Lighting and colors
- Any changes or sudden events