
Transforms a single static image into dynamic video. By uploading a first-frame reference image and a prompt, the model generates motion while maintaining high consistency in style, character, and environment. The workflow uses LTXV preprocessing to ensure high fidelity to the source image, along with synchronized ambient audio. Output is at 24fps, and generating a 10-second video takes approximately 2 minutes and 30
No creation yet
