# New #1 open-source AI video generator is here!

Source: https://www.youtube.com/watch?v=hkfSfr-hMWs
Recap page: https://rapidrecap.app/video/hkfSfr-hMWs
Generated: 2026-01-31T20:35:17.231+00:00

---
## Quick Overview

Lightricks released the LTX-2 open-source audio-video generative model, which offers significantly improved performance and control over previous models, featuring native 4K 50 FPS generation, 20-second clips, and extensive LoRA support for style and camera motion control, making it a new standard for AI video generation.

**Key Points:**
- LTX-2 is a new open-source audio-video generative model released by Lightricks, representing a new standard for AI video generation.
- The model supports native 4K resolution at 50 frames per second (FPS) and can generate clips up to 20 seconds long.
- LTX-2 is natively multimodal, supporting text-to-video, image-to-video, and audio-conditioned generation.
- It features extensive LoRA support, including specialized Camera LoRAs for precise control over dolly, jib, and pan movements.
- The model is optimized for NVIDIA RTX GPUs, with the reviewer testing it successfully on an RTX 4090.
- The base model weights and accompanying LoRAs are available for local download and use via ComfyUI.
- Performance benchmarks show the distilled model generating a 5-second clip in 53 seconds on an RTX 4090, significantly faster than previous models.

![Screenshot at 00:16: The official LTX-2 webpage is displayed, highlighting the model's capabilities such as 'Audio to Video', '20 sec Clip', '50 FPS Performance', and 'Native 4K 50 FPS', alongside download and API key access buttons.](https://ss.rapidrecap.app/screens/hkfSfr-hMWs/00-00-16.jpg)

**Context:** The video introduces LTX-2, the latest AI video generation model released by Lightricks, which they claim sets a new standard for the field. The presenter details the model's capabilities, including high frame rates, longer clip generation, and enhanced control mechanisms like LoRAs for camera motion. The presenter runs demonstrations using the ComfyUI interface, comparing the performance of the full model versus the distilled model on his local RTX 4090 workstation.

## Detailed Analysis

The video announces the release of LTX-2, Lightricks' new open-source audio-video generative model, emphasizing its superior capabilities over previous iterations. The model achieves native 4K resolution at 50 FPS and supports up to 20-second video clips. It is natively multimodal, handling text-to-video, image-to-video, and audio conditioning. A key feature is the extensive support for LoRAs, including specialized Camera LoRAs (like Dolly-Left, Dolly-Right, Jib-Down) that allow precise control over camera movement, structure, and style, which is demonstrated through the ComfyUI node graph. The presenter confirms he is running the model locally on an NVIDIA RTX 4090, noting that the distilled version is much faster (53 seconds for a 5-second clip) than the full model (2 minutes 27 seconds). The video concludes by showing the model's ability to animate static images, such as Edvard Munch's 'The Scream,' and provides the link to the Hugging Face repository for users to download the weights and test the model themselves.

### LTX-2 Release Highlights

- Sets a new standard for AI video generation
- Supports native 4K 50 FPS
- Generates clips up to 20 seconds long
- Open-source weights available for local use.

### Model Capabilities Showcase

- Features text-to-video, image-to-video, and audio-conditioned generation
- Demonstrates high-fidelity output across varied styles (e.g., puppet animation, hyper-realistic scenes, classic art).

### Camera Control LoRAs

- Introduces specialized LoRAs (e.g., Dolly-Left, Dolly-Right) to precisely control camera motion, structure, and parallax, offering creative control beyond simple prompting.

### Performance Benchmarks (RTX 4090)

- Distilled model generates a 5-second clip in 53 seconds, while the full model takes 2 minutes 27 seconds, showing significant speed improvements.

### ComfyUI Workflow Demonstration

- Shows the node-based workflow, highlighting where to load the base model and the specific Camera LoRA nodes, and how to explicitly prompt for camera movements to utilize the LoRAs effectively.

![Screenshot at 00:00: A dramatic shot of a person playing a grand piano submerged in ocean waves during a storm, illustrating the model's capability for high-concept visual scenes.](https://ss.rapidrecap.app/screens/hkfSfr-hMWs/00-00-00.jpg)
![Screenshot at 00:05: A close-up of a photorealistic AI-generated face blowing a large pink bubble, showcasing high fidelity in portrait generation.](https://ss.rapidrecap.app/screens/hkfSfr-hMWs/00-00-05.jpg)
![Screenshot at 00:16: The official LTX-2 webpage is displayed, highlighting the model's capabilities such as 'Audio to Video', '20 sec Clip', '50 FPS Performance', and 'Native 4K 50 FPS', alongside download and API key access buttons.](https://ss.rapidrecap.app/screens/hkfSfr-hMWs/00-00-16.jpg)
![Screenshot at 00:38: A fantasy scene featuring a woman with long purple hair and glowing blue eyes, holding a sword in a darkly lit, neon-tinged cavern.](https://ss.rapidrecap.app/screens/hkfSfr-hMWs/00-00-38.jpg)
![Screenshot at 03:48: The ComfyUI interface displaying a complex node graph used to run the LTX-2 model locally, showing connections between text input, model loading, and save nodes.](https://ss.rapidrecap.app/screens/hkfSfr-hMWs/00-03-48.jpg)
