图生视频:首帧、尾帧与模式选择 — MiniMax H3 教程
让一张图片动起来,或用首尾两张图约束开场与结尾。
原文: ComfyUI MiniMax H3 Text-to-Video, Image-to-Video, and Reference-to-Video Workflows
ComfyUI · 出处
MiniMax H3 Image to Video (I2V)
Generate videos from an input image, with optional first/last-frame keyframes.
Model downloads
Model storage
ComfyUI/
├── 📂 models/
│ ├── 📂 diffusion_models/
│ │ └── minimax_h3_fl2va_pruned_int8_convrot.safetensors
│ ├── 📂 text_encoders/
│ │ └── qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors
│ ├── 📂 loras/
│ │ └── minimax_h3_fl2v_turbo_8step_v1.0_comfyui_bf16.safetensors
│ ├── 📂 vae/
│ │ ├── minimax_h3_video_vae_fp16.safetensors
│ │ └── minimax_h3_audio_vae_fp32.safetensors
│ └── 📂 embeddings/
│ └── minimaxh3_art_is_explosion.safetensors
Prompting tips
- Keyframes: The first_frame and last_frame inputs are optional; the model generates the motion between them
- Prompt: Describe the shots, motion, and the accompanying audio (dialogue, SFX, music) in one block
- Resolution: H3's native canvas is a 768px short edge, which is 1344x768 at 16:9, and resolutions are rounded to a multiple of 32
- Duration: The duration input snaps to the model's 17-frame-per-block (17k+5) grid at 24fps
- Turbo mode: Enable turbo_mode on the MiniMax H3 node to switch to the turbo LoRA and generate in turbo_steps (default 8) instead of 20 steps. turbo_model_strength controls the LoRA strength (default 1.0).
The official base-mode prompt writing guide is summarized in the prompt guide.
适用人群: 第一次在 ComfyUI 中看到 H3 多种工作流的用户。
硬件要求: 已能运行 H3 的本地或云端 ComfyUI;图生视频复用 FL2VA 权重。
前置条件
- 一张首帧图;需要指定结尾时再准备尾帧图
- 先通过入门教程确认模型和两个 VAE 可用
执行步骤
- 打开官方 I2V JSON 或模板库中的 MiniMax H3 I2V,确认加载 FL2VA 模型。
- 上传首帧并连接 first_frame;只想让一张图动起来时,保留 last_frame 为空。
- 需要首尾帧控制时,再上传尾图并连接 last_frame;两张图应能通过一个清晰动作连接。
- 设置与构图相符的宽高比,先用模板预览尺寸;参照官方例子写从输入图开始如何运动。
- 运行后检查开头是否符合首图、结尾是否接近尾图、主体有没有不必要变化;一次调整构图或动作之一。
注意事项
- 不同输入模式对应的权重和接线不同,不要仅替换输入端口。
查看原始来源 · ComfyUI · 2026-09-06
深度精选 · 来源核对: 2026-09-06 · 本站实测: 未进行生成实测
把“首帧、尾帧和参考图”分清楚,再用原生模板连接输入,适合从静态图片开始创作。
本地生成占用自己的硬件、内存和磁盘;模型与工具的使用条件请看原始说明。
链接提供原作者工作流、模型或示例;自选参考素材需自行准备。本站没有另行实测这套生成流程。
适用版本
ComfyUI 0.30+ · native H3 templates
材料与演示
故障排查
图片被裁切
先把输入图裁成目标宽高比,再对照工作流尺寸预览。
把参考人物图当成固定首帧
需要复用身份而不是固定构图时,改看 Ref2VA 教程。
案例与技巧
下一步学习