28
Video
Vladimir Mandic edited this page 2026-08-22 09:23:54 +02:00

Video

SD.Next supports video creation using the top-level Video tab.

Important

Video support requires ffmpeg to be installed and available in the PATH

Important

Most video model are large and require a GPU with at least 16GB VRAM and a system with 64GB RAM
Running video models on lower-spec systems may be possible, but cannot be guaranteed to work

Important

All videos are automatically downloaded on first use and cached for future use
SD.Next does not support manually downloaded video models

Tip

Use aggressive offloading to reduce VRAM usage
Use pre-quantized models where available, and if not use quantization-on-the-fly

UI: Common Tabs

Parameters in common tabs are used by all video model

Prompt

Both positive and negative prompts

Tip

Video models typically require a very long, descriptive prompts

Output

Output video encoding settings and options

Save Enable saving of encoded video file, raw image frames, creation of video thumbnail or save raw video data as safetensors file for future processing

Encode Set target frames-per-second for encoded video, choose video codec and quality settings

Interpolate

Use RiFE to interpolate generated frames to increase frame-rate
For example, most video models generate at native 24 FPS, but you can add 1-frame interpolation to increase the output to 48 FPS

Upscale

Use standard or custom upscalers to increase the resolution of generated frames before video encoding
Currently compatible families of upscalers include ChaiNNer and Spandrel engines

Recommended upscalers are low-latency upscalers that can upscale frames in near-real-time, otherwise upscaling may take significant time and resources:

  • Spandrel SAFMN and RealSAFMN families of upscalers
  • ChaiNNer RealESRGAN Compact family of upscalers

Extras

Additional extensions that can run preprocessing or postprocessing steps

Examples:

  • NudeNet: blur nudity in generated video frames before video encoding
  • Prompt Enhance: use LLM to enhance your short prompts

UI: Model Tabs

  • Base Models
    This is the largest area and includes all models that are supported by SD.Next
    and do not have their own tab with optimized workflows
  • MiniMax: Used for MiniMax-H3 video generation
    See MiniMax wiki for more details
  • LTX: Used for LTXVideo generation
    See LTX wiki for more details
  • FramePack: Used for FramePack generation
    See FramePack wiki for more details

Base Models

Base models supported by SD.Next include the following model families:

  • Hunyuan
  • WAN
  • SkyReels
  • Mochi
  • Latte
  • Allegro
  • Cog
  • Cosmos
  • Sana
  • Kandinsky
  • Veo

And each family includes multiple model versions and variants

Each model is marked as either T2V (text-to-video), I2V (image-to-video), or FLF2V (frame-to-frame video)
Depending on which model you select and load, the UI will automatically switch to the appropriate workflow for that model type

Note

Video models are different than typical image models as they require separate model variant for each workflow type (T2V, I2V, FLF2V)

Tip

Each model may require specific resolutions and parameters for best results.
See each model's original notes for recommended settings.

Note

It is recommended to use Default sampler unless you need a model-specific setting.
For example, to change Sampler Shift, select the matching sampler for that model.

Reference list

Below is a reference list of some of the video models supported by SD.Next with their recommended settings

Engine Model Type Size Optimal Resolution Default Sampler Reference Values Special Notes License
Hunyuan HunyuanVideo T2V 40.9GB 1280x720 Euler FlowMatch Frames:129 CFG:6.0 Steps:50 Proprietary
Hunyuan HunyuanVideo I2V 59.2GB 1280x720 Euler FlowMatch Frames:129 CFG:1.0 Steps:50 Proprietary
HunyuanVideo FramePack T2V/I2V/FLF2V 25.0GB+15GB 608x640 UniPC FlowMatch Frames:73 Steps:25
Hunyuan FastHunyuan T2V 25.0GB+15GB 1280x720 Euler FlowMatch Frames:125 CFG:6.0 True:1.0 Shift:17 Steps:6
Hunyuan SkyReels v1 T2V 25.0GB+15GB 960x544 Euler FlowMatch Frames:97 CFG:1.0 True:6.0 Steps:50
Hunyuan SkyReels v1 I2V 25.0GB+15GB 960x544 Euler FlowMatch Frames:97 CFG:1.0 True:6.0 Steps:50
WAN21 WAN 2.1 1.3B T2V 28.2GB 832x480 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
WAN21 WAN 2.1 14B T2V 78.1GB 1280x720 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
WAN21 WAN 2.1 14B 480p I2V 832x480 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
WAN21 WAN 2.1 14B 720p I2V 1280x720 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
WAN21 WAN 2.1 14B 720p FLF2V 1280x720 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
WAN21 WAN 2.2 5B T2V/I2V 1280x720 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
WAN21 WAN 2.2 A14B T2V/I2V 1280x720 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
WAN21 WAN 2.2 14B VACE T2V/I2V 1280x720 UniPC Frames:81 CFG:5.0 Steps:50 Apache 2.0
LTXVideo LTXVideo 0.9.0 T2V 704x480 Euler FlowMatch Frames:161 Steps:50 Proprietary
LTXVideo LTXVideo 0.9.0 I2V 704x480 Euler FlowMatch Frames:161 Steps:50 Proprietary
LTXVideo LTXVideo 0.9.1 T2V 24.1GB 704x512 Euler FlowMatch Frames:161 CFG:3 Steps:50 Proprietary
LTXVideo LTXVideo 0.9.1 I2V 24.1GB 704x512 Euler FlowMatch Frames:161 CFG:3 Steps:50 Proprietary
LTXVideo LTXVideo 0.9.5 T2V 24.8GB 768x512 Euler FlowMatch Frames:161 Steps:40 Proprietary
LTXVideo LTXVideo 0.9.5 I2V 768x512 Euler FlowMatch Frames:161 Steps:40 Proprietary
LTXVideo LTXVideo 0.9.6 2B T2V 768x512 Euler FlowMatch Frames:161 Steps:50 Proprietary
LTXVideo LTXVideo 0.9.6 2B Distilled T2V 768x512 Euler FlowMatch Frames:161 Steps:8 Proprietary
LTXVideo LTXVideo 0.9.7 13B T2V/I2V/V2V 46.3GB 768x512 Euler FlowMatch Frames:161 Steps:50 Proprietary
LTXVideo LTXVideo 0.9.8 13B T2V/I2V/V2V 46.3GB 768x512 Euler FlowMatch Frames:161 Steps:50 Proprietary
LTXVideo LTXVideo 2.0 19B T2V/I2V/V2V 63.0GB 768x512 Euler FlowMatch Frames:161 Steps:50 Proprietary
LTXVideo LTXVideo 2.3 22B T2V/I2V/V2V 71.7GB 768x512 Euler FlowMatch Frames:161 Steps:50 Proprietary
CogVideoX CogVideoX 1.0 2B T2V 720x480 Cog DDIM Frames:49 CFG:6.0 Steps:50 Apache 2.0
CogVideoX CogVideoX 1.0 5B T2V 720x480 Cog DDIM Frames:49 CFG:6.0 Steps:50 Proprietary
CogVideoX CogVideoX 1.0 5B I2V 720x480 Cog DDIM Frames:49 CFG:6.0 Steps:50 Proprietary
CogVideoX CogVideoX 1.5 5B T2V 30.3GB 1360x768 Cog DDIM Frames:81 CFG:6.0 Steps:50 Issue: blank output Proprietary
CogVideoX CogVideoX 1.5 5B I2V 1360x768 Cog DDIM Frames:81 CFG:6.0 Steps:50 Issue: blank output Proprietary
Allegro Allegro T2V 24.7GB 1280x720 Euler a Frames:88 CFG:7.5 Steps=100 Issue: blank output Apache 2.0
Mochi Mochi1 T2V 23.4GB 512x512 Euler FlowMatch Frames:16 CFG:7.5 Steps:50 Apache 2.0
Latte Latte1 T2V 23.4GB 512x512 DDIM Frames:16 CFG:7.5 Steps:50 Apache 2.0
Kandinsky Kandinsky 5 Lite T2V 23.4GB 768x512 Euler FlowMatch Frames:121 Steps:50 CFG:5.0 Apache 2.0
Kandinsky Kandinsky 5 Lite CFG-Distilled T2V 23.4GB 768x512 Euler FlowMatch Frames:121 Steps:50 CFG:1.0 Apache 2.0
Kandinsky Kandinsky 5 Lite Steps-Distilled T2V 23.4GB 768x512 Euler FlowMatch Frames:121 Steps:16 CFG:1.0 Apache 2.0

Legacy models

Additional video models are available as individually selectable scripts in text or image interfaces.