Lightx2v is an advanced image-to-video generation model built upon the Wan2.2-I2V-A14B foundation. This approach allows the model to generate videos with significantly fewer inference steps (4 steps, 2 steps for high noise and 2 steps for low noise) and without classifier-free guidance, substantially reducing video generation time while maintaining high quality outputs.
This version has the following features:
- We found that the training challenge of Wan2.2 lies in high noise; therefore, we have focused on the two-step training for the high-noise model. Compared with the previous version, we have adopted several new strategies, which have improved the consistency and dynamics of the model.
- We found that the low-noise model can achieve good results by directly using the LoRA from Wan2.1; thus, the low-noise model still adopts the old Wan2.1 LoRA.
Description
Details
Downloads
25
Platform
SeaArt
Platform Status
Available
Created
10/14/2025
Updated
10/14/2025
Deleted
-
