通过文字指令编辑现有视频,可添加最多 10 张参考图片,并选择是否保留原声。
Minimax H3 Image to Video is MiniMax's H3-series model for turning one picture into moving footage. You hand it the opening frame and describe what should happen; the model resolves motion, camera behavior, and scene development, then renders at 768p or 2K.
Because the first frame is pinned by your upload, subject identity, wardrobe, palette, and framing carry through the clip. Minimax H3 Image to Video therefore fits work where the look is settled and only movement is missing.
768p for cheaper drafts or 2k when faces, packaging text, and fine texture need to survive a full-screen crop.These are the live fields for Minimax H3 Image to Video on this page.
| Parameter | Required | Type | Default | Range / Options | Description |
|---|---|---|---|---|---|
| image * | Yes (*) | string (URL) | Sample first frame | 256-5760 px per side, aspect ratio 0.4-2.5 | Opening frame that anchors subject, composition, and style. |
| prompt * | Yes (*) | string | Sample motion brief | 1-4000 characters | What should move, how the camera behaves, how the scene develops. |
| last_image | No | string (URL) | None | Same image constraints as image | Closing frame used to steer where the clip ends. |
| duration * | Yes (*) | integer | 5 | 4-15 | Clip length in whole seconds. |
| resolution * | Yes (*) | string | 768p | 768p, 2k | Output resolution tier. |
Minimax H3 Image to Video bills on resolution and the length of the finished clip:
| Resolution | Price per second | 5s | 10s | 15s |
|---|---|---|---|---|
| 768p | $0.11 | $0.55 | $1.10 | $1.65 |
| 2K | $0.16 | $0.80 | $1.60 | $2.40 |
1) Upload the opening frame — Pick an image where the subject reads clearly and the composition matches the shot you want.
2) Write the motion brief — Say what moves, how fast, and what the camera does. Minimax H3 Image to Video answers to verbs and camera language, not adjective lists.
3) Add a closing frame (optional) — Supply a last image when the ending matters, and keep it compatible with the first frame.
4) Set the length — Start at 5 seconds while tuning wording, then extend toward 15 once the motion behaves.
5) Choose the resolution — Use 768p for drafts; switch to 2k for final deliveries that need more detail.
6) Generate — Submit and review the Minimax H3 Image to Video result at full size.
7) Iterate one variable at a time — Swap the prompt, image, duration, or resolution separately so you can tell what changed the take.
When Minimax H3 Image to Video is not the right fit, these are worth a pass:
通过文字指令编辑现有视频,可添加最多 10 张参考图片,并选择是否保留原声。
支持风格控制与物体编辑的电影级AI视频生成工具
智能文本转视频工具,支持1080p高质量生成,轻松打造生动镜头与真实情感,激发创作灵感。
按每秒 $0.084,将图像制作成 3–15 秒视频。
使用 Pikadditions 将图像中的人物或物体添加到现有视频中。
Seedance 2.5 480p:用提示词快速、低成本生成文生视频草稿
Minimax H3 图像转视频会拍摄静态图像和书面运动简介,并返回 768p 或 2K 视频剪辑。它专为将艺术作品、产品摄影或渲染帧转变为移动镜头而无需在 3D 或合成工具中重建场景。上传的图像固定了开头框架,因此您已经批准的主题和框架会带入剪辑中。
除了所需的第一帧之外,Minimax H3 图像到视频还接受可选的最后一个图像,用于指导剪辑如何结束。该模型计算出两个静止图像之间的运动,而不是漂移到任意的最终状态。保持两个图像在灯光、调色板和取景方面在视觉上兼容,否则过渡看起来会很强制。
输出分辨率可选择“768p”或“2k”,持续时间可选择 4 至 15 秒(整秒)。当您仍在使用 Minimax H3 Image to Video 测试提示措辞时,较短的剪辑是实用的选择,而较长的剪辑则为场景提供了发展空间。检查此页面上的参数面板以获取当前公开的确切值。
源图像每边的尺寸应在 256 到 5760 像素之间,长宽比在 0.4 到 2.5 之间,提示接受 1 到 4000 个字符。 Minimax H3 图像转视频不会重新组合裁剪不当的主题,因此请提供已与您的预期取景相匹配的图像。限制可能会因模式或提供商设置而异,因此请根据实时面板进行确认。
首先描述什么物理移动,然后描述相机的行为,然后描述灯光和情绪。 Minimax H3 图像到视频对具体动词和命名相机移动(例如缓慢推入或轨道)的响应比对堆叠形容词的响应更好。避免在一个提示中混合两种相互竞争的视觉风格,因为结果往往会将它们平均化。
文本到视频从头开始决定主题、构图和风格,这使得很难两次获得认可的外观。 Minimax H3 图像转视频通过将开头框架锁定到上传内容来删除该变量,这适合艺术指导已签署的产品镜头、角色作品和活动材料。它在设计上是一个更窄的工具,具有较短的输入列表和快速的迭代循环。
是的。在 RunComfy AI Playground Web UI 中进行原型设计,确定图像、提示、分辨率和持续时间,然后通过 RunComfy API 使用相同的参数调用相同的模型。这使您可以在自己的应用程序中的手动探索和批处理或计划作业之间保持 Minimax H3 图像到视频设置的一致。
一代又一代地消耗您的 RunComfy 余额中的美元或积分。 768p 的速率为每秒 0.11 美元,2K 的速率为每秒 0.16 美元。因此,一段 5 秒的剪辑在 768p 下售价为 0.55 美元,在 2K 下售价为 0.80 美元; 10 秒费用为 1.10 美元或 1.60 美元; 15 秒的费用为 1.65 美元或 2.40 美元。新用户通常会收到免费试用金额进行测试;有关计费问题,请联系 hi@runcomfy.com。
RunComfy 是首选的 ComfyUI 平台,提供 ComfyUI 在线 环境和服务,以及 ComfyUI 工作流 具有惊艳的视觉效果。 RunComfy还提供 AI Models, 帮助艺术家利用最新的AI工具创作出令人惊叹的艺术作品。





