Stable Video Diffusion
Stability AIStable Video Diffusion is Stability AI's latent diffusion model that turns a single still image into a short video clip. The original SVD generates 14 frames and SVD-XT 25 frames at 576x1024; SVD-XT 1.1, released in February 2024, is a fine-tune of SVD-XT for more consistent output.
Stable Video 3D (SV3D) and Stable Video 4D (SV4D, SV4D 2.0 of May 2025) are built on SVD but generate orbital and multi-view video of an object rather than a clip from an image, so they are not part of this entry; the category already treats Stability AI's 3D models as separate products. The generative-models repository is shared with Stable Diffusion and publishes SVD inference only. Dormant: no SVD release since February 2024, but the two 2023 checkpoints still draw about 330K downloads a month.
Openness
3 medium confidence- weights
- open(SVD and SVD-XT download without a gate
- data
- described(the SVD paper and cards describe the curated training video set and its filtering
- code
- partial(Stability-AI/generative-models publishes SVD inference configs and sampling scripts under MIT
- license
- Stability-AI-Community-License(the LICENSE.md in every SVD repository: free below USD 1M annual revenue, after which an enterprise license is required
Stable Video Diffusion is free to use, commercially included, for organizations under USD 1M in annual revenue, with a paid license above that; the newest checkpoint's download form still asks for non-commercial use. Stability publishes inference code only, and the training videos are described but not released. Stability-AI-Community-License allows commercial use only within a bound, so the license tier is use_bounded.
- https://huggingface.co/api/models?author=stabilityai&search=stable-video&sort=downloads&direction=-1&limit=100 recorded 2026-09-27
Listing of the stable-video repositories: img2vid-xt, img2vid and img2vid-xt-1-1 are public.
- https://huggingface.co/api/models/stabilityai/stable-video-diffusion-img2vid-xt-1-1?expand[]=downloads&expand[]=likes&expand[]=cardData&expand[]=gated&expand[]=lastModified&expand[]=createdAt recorded 2026-09-27
Hugging Face API record for SVD-XT 1.1: gated "auto", license_name stable-video-diffusion-1-1-community, license_link LICENSE.md; the gate field reads "will use the Software Products and Derivative Works for non-commercial or research purposes only".
- https://huggingface.co/api/models/stabilityai/stable-video-diffusion-img2vid-xt?expand[]=gated&expand[]=cardData recorded 2026-09-27
Hugging Face API record for SVD-XT: gated false.
- https://huggingface.co/api/models/stabilityai/stable-video-diffusion-img2vid?expand[]=gated recorded 2026-09-27
Hugging Face API record for SVD: gated false.
- https://huggingface.co/stabilityai/stable-video-diffusion-img2vid-xt/raw/main/LICENSE.md recorded 2026-09-27
License body: "STABILITY AI COMMUNITY LICENSE AGREEMENT / Last Updated: July 5, 2024", with "free access to the Models for people or organizations generating annual revenue of less than US $1,000,000".
- https://huggingface.co/stabilityai/stable-video-diffusion-img2vid-xt/raw/main/README.md recorded 2026-09-27
Card: "The model is intended for both non-commercial and commercial usage" and "All considered potential data sources were included for final training"; no dataset is released.
- https://huggingface.co/stabilityai/stable-video-diffusion-img2vid/raw/main/LICENSE.md recorded 2026-09-27
The same Stability AI Community License body in the SVD repository.
- https://ungh.cc/repos/Stability-AI/generative-models/files/main recorded 2026-09-27
generative-models file list: inference configs for svd and sv3d; configs/example_training holds ImageNet, CIFAR-10, MNIST and txt2img-clipl configs, none for SVD.
- https://ungh.cc/repos/Stability-AI/generative-models/files/main/LICENSE-CODE recorded 2026-09-27
LICENSE-CODE: "MIT License".
Adoption
3 high confidenceAdoption is measured as Hugging Face downloads across the SVD checkpoints, almost all of them the original SVD-XT and SVD rather than the newer 1.1.
- https://huggingface.co/api/models?author=stabilityai&search=stable-video&sort=downloads&direction=-1&limit=100 recorded 2026-09-27
Listing sorted by downloads: stable-video-diffusion-img2vid-xt 241386, img2vid 86318, img2vid-xt-1-1 3223, img2vid-xt-1-1-tensorrt 35.
Capability
1 medium confidenceStable Video Diffusion generates only from an image and has no entry on either public video arena. It sits below Mochi 1, which places on the text-to-video board.
- https://artificialanalysis.ai/video/leaderboard/image-to-video recorded 2026-09-27
The image-to-video leaderboard lists no Stable Video Diffusion entry.
- https://artificialanalysis.ai/video/leaderboard/text-to-video recorded 2026-09-27
The text-to-video leaderboard lists no Stable Video Diffusion entry; "Mochi 1" (rank 73) is on it.
- https://huggingface.co/stabilityai/stable-video-diffusion-img2vid-xt/raw/main/README.md recorded 2026-09-27
Card: "trained to generate 25 frames at resolution 576x1024".
Verified 2026-09-27