Table of Contents
Video
SD.Next supports video creation using the top-level Video tab.
Important
Video support requires ffmpeg to be installed and available in the
PATH
Important
Most video model are large and require a GPU with at least 16GB VRAM and a system with 64GB RAM
Running video models on lower-spec systems may be possible, but cannot be guaranteed to work
Important
All videos are automatically downloaded on first use and cached for future use
SD.Next does not support manually downloaded video models
Tip
Use aggressive offloading to reduce VRAM usage
Use pre-quantized models where available, and if not use quantization-on-the-fly
UI: Common Tabs
Parameters in common tabs are used by all video model
Prompt
Both positive and negative prompts
Tip
Video models typically require a very long, descriptive prompts
Output
Output video encoding settings and options
Save Enable saving of encoded video file, raw image frames, creation of video thumbnail or save raw video data as safetensors file for future processing
Encode Set target frames-per-second for encoded video, choose video codec and quality settings
Interpolate
Use RiFE to interpolate generated frames to increase frame-rate
For example, most video models generate at native 24 FPS, but you can add 1-frame interpolation to increase the output to 48 FPS
Upscale
Use standard or custom upscalers to increase the resolution of generated frames before video encoding
Currently compatible families of upscalers include ChaiNNer and Spandrel engines
Recommended upscalers are low-latency upscalers that can upscale frames in near-real-time, otherwise upscaling may take significant time and resources:
- Spandrel SAFMN and RealSAFMN families of upscalers
- ChaiNNer RealESRGAN Compact family of upscalers
Extras
Additional extensions that can run preprocessing or postprocessing steps
Examples:
- NudeNet: blur nudity in generated video frames before video encoding
- Prompt Enhance: use LLM to enhance your short prompts
UI: Model Tabs
- Base Models
This is the largest area and includes all models that are supported by SD.Next
and do not have their own tab with optimized workflows - MiniMax: Used for MiniMax-H3 video generation
See MiniMax wiki for more details - LTX: Used for LTXVideo generation
See LTX wiki for more details - FramePack: Used for FramePack generation
See FramePack wiki for more details
Base Models
Base models supported by SD.Next include the following model families:
- Hunyuan
- WAN
- SkyReels
- Mochi
- Latte
- Allegro
- Cog
- Cosmos
- Sana
- Kandinsky
- Veo
And each family includes multiple model versions and variants
Each model is marked as either T2V (text-to-video), I2V (image-to-video), or FLF2V (frame-to-frame video)
Depending on which model you select and load, the UI will automatically switch to the appropriate workflow for that model type
Note
Video models are different than typical image models as they require separate model variant for each workflow type (T2V, I2V, FLF2V)
Tip
Each model may require specific resolutions and parameters for best results.
See each model's original notes for recommended settings.
Note
It is recommended to use Default sampler unless you need a model-specific setting.
For example, to change Sampler Shift, select the matching sampler for that model.
Reference list
Below is a reference list of some of the video models supported by SD.Next with their recommended settings
| Engine | Model | Type | Size | Optimal Resolution | Default Sampler | Reference Values | Special Notes | License |
|---|---|---|---|---|---|---|---|---|
| Hunyuan | HunyuanVideo | T2V | 40.9GB | 1280x720 | Euler FlowMatch | Frames:129 CFG:6.0 Steps:50 | Proprietary | |
| Hunyuan | HunyuanVideo | I2V | 59.2GB | 1280x720 | Euler FlowMatch | Frames:129 CFG:1.0 Steps:50 | Proprietary | |
| HunyuanVideo | FramePack | T2V/I2V/FLF2V | 25.0GB+15GB | 608x640 | UniPC FlowMatch | Frames:73 Steps:25 | ||
| Hunyuan | FastHunyuan | T2V | 25.0GB+15GB | 1280x720 | Euler FlowMatch | Frames:125 CFG:6.0 True:1.0 Shift:17 Steps:6 | ||
| Hunyuan | SkyReels v1 | T2V | 25.0GB+15GB | 960x544 | Euler FlowMatch | Frames:97 CFG:1.0 True:6.0 Steps:50 | ||
| Hunyuan | SkyReels v1 | I2V | 25.0GB+15GB | 960x544 | Euler FlowMatch | Frames:97 CFG:1.0 True:6.0 Steps:50 | ||
| WAN21 | WAN 2.1 1.3B | T2V | 28.2GB | 832x480 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | |
| WAN21 | WAN 2.1 14B | T2V | 78.1GB | 1280x720 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | |
| WAN21 | WAN 2.1 14B 480p | I2V | 832x480 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | ||
| WAN21 | WAN 2.1 14B 720p | I2V | 1280x720 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | ||
| WAN21 | WAN 2.1 14B 720p | FLF2V | 1280x720 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | ||
| WAN21 | WAN 2.2 5B | T2V/I2V | 1280x720 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | ||
| WAN21 | WAN 2.2 A14B | T2V/I2V | 1280x720 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | ||
| WAN21 | WAN 2.2 14B VACE | T2V/I2V | 1280x720 | UniPC | Frames:81 CFG:5.0 Steps:50 | Apache 2.0 | ||
| LTXVideo | LTXVideo 0.9.0 | T2V | 704x480 | Euler FlowMatch | Frames:161 Steps:50 | Proprietary | ||
| LTXVideo | LTXVideo 0.9.0 | I2V | 704x480 | Euler FlowMatch | Frames:161 Steps:50 | Proprietary | ||
| LTXVideo | LTXVideo 0.9.1 | T2V | 24.1GB | 704x512 | Euler FlowMatch | Frames:161 CFG:3 Steps:50 | Proprietary | |
| LTXVideo | LTXVideo 0.9.1 | I2V | 24.1GB | 704x512 | Euler FlowMatch | Frames:161 CFG:3 Steps:50 | Proprietary | |
| LTXVideo | LTXVideo 0.9.5 | T2V | 24.8GB | 768x512 | Euler FlowMatch | Frames:161 Steps:40 | Proprietary | |
| LTXVideo | LTXVideo 0.9.5 | I2V | 768x512 | Euler FlowMatch | Frames:161 Steps:40 | Proprietary | ||
| LTXVideo | LTXVideo 0.9.6 2B | T2V | 768x512 | Euler FlowMatch | Frames:161 Steps:50 | Proprietary | ||
| LTXVideo | LTXVideo 0.9.6 2B Distilled | T2V | 768x512 | Euler FlowMatch | Frames:161 Steps:8 | Proprietary | ||
| LTXVideo | LTXVideo 0.9.7 13B | T2V/I2V/V2V | 46.3GB | 768x512 | Euler FlowMatch | Frames:161 Steps:50 | Proprietary | |
| LTXVideo | LTXVideo 0.9.8 13B | T2V/I2V/V2V | 46.3GB | 768x512 | Euler FlowMatch | Frames:161 Steps:50 | Proprietary | |
| LTXVideo | LTXVideo 2.0 19B | T2V/I2V/V2V | 63.0GB | 768x512 | Euler FlowMatch | Frames:161 Steps:50 | Proprietary | |
| LTXVideo | LTXVideo 2.3 22B | T2V/I2V/V2V | 71.7GB | 768x512 | Euler FlowMatch | Frames:161 Steps:50 | Proprietary | |
| CogVideoX | CogVideoX 1.0 2B | T2V | 720x480 | Cog DDIM | Frames:49 CFG:6.0 Steps:50 | Apache 2.0 | ||
| CogVideoX | CogVideoX 1.0 5B | T2V | 720x480 | Cog DDIM | Frames:49 CFG:6.0 Steps:50 | Proprietary | ||
| CogVideoX | CogVideoX 1.0 5B | I2V | 720x480 | Cog DDIM | Frames:49 CFG:6.0 Steps:50 | Proprietary | ||
| CogVideoX | CogVideoX 1.5 5B | T2V | 30.3GB | 1360x768 | Cog DDIM | Frames:81 CFG:6.0 Steps:50 | Issue: blank output | Proprietary |
| CogVideoX | CogVideoX 1.5 5B | I2V | 1360x768 | Cog DDIM | Frames:81 CFG:6.0 Steps:50 | Issue: blank output | Proprietary | |
| Allegro | Allegro | T2V | 24.7GB | 1280x720 | Euler a | Frames:88 CFG:7.5 Steps=100 | Issue: blank output | Apache 2.0 |
| Mochi | Mochi1 | T2V | 23.4GB | 512x512 | Euler FlowMatch | Frames:16 CFG:7.5 Steps:50 | Apache 2.0 | |
| Latte | Latte1 | T2V | 23.4GB | 512x512 | DDIM | Frames:16 CFG:7.5 Steps:50 | Apache 2.0 | |
| Kandinsky | Kandinsky 5 Lite | T2V | 23.4GB | 768x512 | Euler FlowMatch | Frames:121 Steps:50 CFG:5.0 | Apache 2.0 | |
| Kandinsky | Kandinsky 5 Lite CFG-Distilled | T2V | 23.4GB | 768x512 | Euler FlowMatch | Frames:121 Steps:50 CFG:1.0 | Apache 2.0 | |
| Kandinsky | Kandinsky 5 Lite Steps-Distilled | T2V | 23.4GB | 768x512 | Euler FlowMatch | Frames:121 Steps:16 CFG:1.0 | Apache 2.0 |
Legacy models
Additional video models are available as individually selectable scripts in text or image interfaces.
- Stable Video Diffusion, Base, XY 1.0 and XT 1.1
- VGen
- AnimateDiff