Model reference · open weights

stable-video-diffusion-img2vid

stable-video-diffusion-img2vid is an open-weight video model from stabilityai, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Licence fee required Video stabilityai 1 variants 38k downloads/mo
Request a licence + hosting quote All served models Not on the shared API today — deployed on request.

About

What stable-video-diffusion-img2vid is

Stable Video Diffusion Image-to-Video Model Card Stable Video Diffusion (SVD) Image-to-Video is a diffusion model that takes in a still image as a conditioning frame, and generates a video from it. Please note: For commercial use of this model, please refer to https://stability.ai/license. Model Details Model Description (SVD) Image-to-Video is a latent diffusion model trained to generate short video clips from an image conditioning. This model was trained to generate 14 frames at resolution 576x1024 given a context frame of the same size. We also finetune the widely used f8-decoder for temporal consistency. For convenience, we additionally provide the model with the standard frame-wise decoder here. - Developed by: Stability AI - Funded by: Stability AI - Model type: Generative image-to-video model Model Sources For research purposes, we recommend our generative-models Github repository (https://github.com/Stability-AI/generative-models), which implements the most popular diffusion frameworks (both training and inference). - Repository: https://github.com/Stability-AI/generative-models - Paper: https://stability.ai/research/stable-video-diffusion-scaling-latent-video-diffusion-models-to-large-datasets Evaluation The chart above evaluates user preference for SVD-Image-to-Video over GEN-2 and PikaLabs. SVD-Image-to-Video is preferred by human voters in terms of video quality. For details on the user study, we refer to the research paper Uses Direct Use The model is intended for research purposes only. Possible research areas and tasks include - Research on generative models. - Safe deployment of models which have the potential to generate harmful content. - Probing and understanding the limitations and biases of generative models. - Generation of artworks and use in design and other artistic processes. - Applications in educational or creative tools. Excluded uses are described below. Out-of-Scope Use The model was not trained to be factual or true representations of people or events, and therefore using the model to generate such content is out-of-scope for the abilities of this model. The model should not be used in any way that violates Stability AI's Acceptab

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makerstabilityai
TypeVideo models
Parameters (lead)1.5B
Variants1
Runs withdiffusers
Released2023-11-20
Popularity38k downloads / month
Likes1,056
LicenceCommercial licence needed

How it works

How video models work

Prompt / imagestart pointTemporal diffusionframes over timeVideoMP4 clipA video model generates a sequence of coherent frames from your prompt or a starting image.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
stable-video-diffusion-img2vid1.5BBF16~3.5 GBWeights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys stable-video-diffusion-img2vid for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (stable-video-diffusion-img2vid below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/videos/generations \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"stable-video-diffusion-img2vid","prompt":"a drone shot over a forest"}'

Details

Languages, data & research

Tags

diffusers safetensors image-to-video diffusers:StableVideoDiffusionPipeline

Licence

Commercial licence needed

The weights are open but its licence needs a commercial agreement for business use. AxForge can arrange that licence and host the model for you — you pay AxForge, we settle with the model’s maker. Ask us for a quote. Read the licence ↗

Sources

Weights & code

Want stable-video-diffusion-img2vid on EU-owned hardware?

Request a licence + hosting quote See what’s served now

Explore

More video models

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms