Skip to content
GenLovers

Stable Video Diffusion

Open-weight models you can run yourself
Visit Stable Video Diffusion →

About Stable Video Diffusion

An image-to-video model: it animates a single still. The XT model produces 25 frames at 576x1024, roughly four seconds or less. The card lists limits: no text control, weak text rendering and difficulty with realistic faces. The license allows commercial use below USD 1 million in annual revenue; above that, a separate license from Stability is needed.

Everything here is what the tool's own site states on the date shown. It is not a test result. Spotted a mistake? Tell us.

More in Open-weight models you can run yourself

Back to the full AI video generators list →