logo
RunComfy
  • ComfyUI
  • TrainerNew
  • Models
  • API
  • Pricing
discord logo
MODELS
Explore
All Models
LIBRARY
Generations
MODEL APIS
API Docs
API Keys
ACCOUNT
Usage

Minimax H3 Image to Video: 768p & 2K Image-to-Video with Last-Frame Control on Models and API | RunComfy

minimax/minimax-h3/image-to-video

Minimax H3 Image to Video animates a first-frame image into a 768p or 2K clip of 4–15 seconds, with optional last-frame control.

First-frame image. Each side must be 256-5760 pixels with an aspect ratio between 0.4 and 2.5.
Description of the motion and scene. 1-7000 characters.
Optional last-frame image used to guide the ending of the video.
Length of the generated video in seconds.
Output video resolution. Choose 768p or 2k.
Idle
The rate is $0.09 per second for 768p, and $0.145 per second for 2k.

Introduction To Minimax H3 Image to Video

MiniMax's Minimax H3 Image to Video turns a single still into a coherent 768p or 2K clip of 4 to 15 seconds, driven by a plain-language motion brief. Trading keyframe rigs, rotoscoping, and manual in-betweening for one prompt plus an optional closing frame, Minimax H3 Image to Video gives designers, ad teams, previz artists, and product marketers a direct path from one picture to a finished shot. For developers, the same parameters are available through the RunComfy API.

MiniMax / MiniMax H3 Image To Video#


Minimax H3 Image to Video is MiniMax's H3-series model for turning one picture into moving footage. You hand it the opening frame and describe what should happen; the model resolves motion, camera behavior, and scene development, then renders at 768p or 2K.


Because the first frame is pinned by your upload, subject identity, wardrobe, palette, and framing carry through the clip. Minimax H3 Image to Video therefore fits work where the look is settled and only movement is missing.


Highlights#


  • Frame-anchored motion: Your image sets the starting state, so Minimax H3 Image to Video extends what is on screen instead of inventing a subject.

  • Optional closing frame: Give Minimax H3 Image to Video a last image and it solves the path between two stills — good for reveals and shot hand-offs.

  • 768p or 2K output: Choose 768p for cheaper drafts or 2k when faces, packaging text, and fine texture need to survive a full-screen crop.

  • Adjustable length: Minimax H3 Image to Video runs any whole number of seconds from 4 to 15.

  • Plain-language direction: One prompt field carries subject action, camera move, lighting, and mood together, so Minimax H3 Image to Video reads intent instead of demanding preset menus.

Parameters#


These are the live fields for Minimax H3 Image to Video on this page.


ParameterRequiredTypeDefaultRange / OptionsDescription
image *Yes (*)string (URL)Sample first frame256-5760 px per side, aspect ratio 0.4-2.5Opening frame that anchors subject, composition, and style.
prompt *Yes (*)stringSample motion brief1-7000 charactersWhat should move, how the camera behaves, how the scene develops.
last_imageNostring (URL)NoneSame image constraints as imageClosing frame used to steer where the clip ends.
duration *Yes (*)integer54-15Clip length in whole seconds.
resolution *Yes (*)string768p768p, 2kOutput resolution tier.

Pricing#


Minimax H3 Image to Video bills on resolution and the length of the finished clip:


ResolutionPrice per second5s10s15s
768p$0.09$0.45$0.90$1.35
2K$0.145$0.725$1.45$2.175

How to Use#


1) Upload the opening frame — Pick an image where the subject reads clearly and the composition matches the shot you want.


2) Write the motion brief — Say what moves, how fast, and what the camera does. Minimax H3 Image to Video answers to verbs and camera language, not adjective lists.


3) Add a closing frame (optional) — Supply a last image when the ending matters, and keep it compatible with the first frame.


4) Set the length — Start at 5 seconds while tuning wording, then extend toward 15 once the motion behaves.


5) Choose the resolution — Use 768p for drafts; switch to 2k for final deliveries that need more detail.


6) Generate — Submit and review the Minimax H3 Image to Video result at full size.


7) Iterate one variable at a time — Swap the prompt, image, duration, or resolution separately so you can tell what changed the take.


Prompt & Reference Tips#


  • Lead with the subject and its action, then layer camera, light, and mood behind it, since Minimax H3 Image to Video weights the opening clause heavily.

  • Name the camera move outright: slow push-in, orbit, handheld follow.

  • Keep the source image close to your target framing; Minimax H3 Image to Video will not recompose a badly cropped subject.

  • Match lighting and palette across both stills when using a closing frame.

  • Describe progression across longer clips, such as mist thinning before sun breaks through.

  • Drop competing style cues; one visual direction per prompt reads cleaner.

How MiniMax H3 Image To Video compares to other models#


  • Against text-only video models: A fixed opening frame removes guesswork about who appears, so Minimax H3 Image to Video suits jobs where art direction is locked.

  • Against fixed-tier animators: Minimax H3 Image to Video lets you choose 768p or 2K, so you can draft cheaply and deliver at higher detail when needed.

  • Against first-frame-only tools: Last-frame guidance controls where a shot ends, so Minimax H3 Image to Video can carry a planned transition.

  • Against reference-heavy video models: Minimax H3 Image to Video keeps a short input list, trading broad multimodal conditioning for speed.

More Models to Try#


When Minimax H3 Image to Video is not the right fit, these are worth a pass:


  • Hailuo 02 — MiniMax's earlier image-to-video release for quick motion.

  • Hailuo 2.3 Pro — a production-tuned Hailuo option.

  • Kling 3.0 — structured multi-shot with native audio.

  • Seedance 2.0 Pro — reference-heavy video work.

Related Models

happyhorse-1.1/image-to-video

Animate a still photo into smooth 720P or 1080P video from one prompt.

hailuo-2-3/standard/image-to-video

Transform images into motion-rich clips with Hailuo 2.3's precise control and realistic visuals.

elevenlabs/music-generation

Prompt-driven song creation with 44.1 kHz WAV control and section editing

kling-video-o3/4K/reference-to-video

Cinematic 4K reference-to-video at $0.47 per second of output.

wan-2-6/video-to-video

Transforms reference clips into 1080p short videos with precise motion and voice alignment.

wan-2-5/image-to-video

Generate clips with fluid motion and audios for creatives

Frequently Asked Questions

What is Minimax H3 Image to Video used for?

Minimax H3 Image to Video takes a still image plus a written motion brief and returns a 768p or 2K video clip. It is built for turning artwork, product photography, or a rendered frame into a moving shot without rebuilding the scene in a 3D or compositing tool. The uploaded image fixes the opening frame, so the subject and framing you already approved carry into the clip.

How does the last image option work in Minimax H3 Image to Video?

Besides the required first frame, Minimax H3 Image to Video accepts an optional last image that guides how the clip ends. The model works out the movement between the two stills instead of drifting toward an arbitrary final state. Keep the two images visually compatible in lighting, palette, and framing, or the transition will look forced.

What video quality and length can Minimax H3 Image to Video produce?

Output resolution is selectable as 768p or 2k, and duration is selectable from 4 to 15 seconds in whole seconds. Shorter clips are the practical choice while you are still testing prompt wording with Minimax H3 Image to Video, and longer clips give a scene room to develop. Check the parameter panel on this page for the exact values currently exposed.

What input limits should I know before using Minimax H3 Image to Video?

The source image should measure between 256 and 5760 pixels on each side with an aspect ratio between 0.4 and 2.5, and the prompt accepts 1 to 4000 characters. Minimax H3 Image to Video does not recompose a poorly cropped subject, so supply an image that already matches your intended framing. Limits may vary by mode or provider settings, so confirm against the live panel.

How should I write prompts for Minimax H3 Image to Video?

Describe what physically moves first, then the camera behavior, then lighting and mood. Minimax H3 Image to Video responds better to concrete verbs and named camera moves such as a slow push-in or an orbit than to stacked adjectives. Avoid mixing two competing visual styles in one prompt, since the result tends to average them out.

How does Minimax H3 Image to Video compare to text-to-video generation?

Text-to-video decides the subject, composition, and style from scratch, which makes it hard to hit an approved look twice. Minimax H3 Image to Video removes that variable by locking the opening frame to your upload, which suits product shots, character work, and campaign material where the art direction is already signed off. It is a narrower tool by design, with a short input list and a fast iteration loop.

Can developers call Minimax H3 Image to Video through the RunComfy API?

Yes. Prototype in the RunComfy AI Playground Web UI to settle on your image, prompt, resolution, and duration, then call the same model through the RunComfy API with identical parameters. This lets you keep Minimax H3 Image to Video settings consistent between manual exploration and batch or scheduled jobs in your own application.

How much does it cost to generate with Minimax H3 Image to Video on RunComfy?

Generations consume usd or credits from your RunComfy balance. The rate is $0.09 per second at 768p and $0.145 per second at 2K. A 5-second clip therefore costs $0.45 at 768p or $0.725 at 2K; 10 seconds costs $0.90 or $1.45; 15 seconds costs $1.35 or $2.175. New users typically receive a free trial amount to test with; for billing questions, contact hi@runcomfy.com.

Follow us
  • LinkedIn
  • Facebook
  • Instagram
  • Twitter
Support
  • Discord
  • Email
  • System Status
  • Affiliate
Video Models
  • Seedance 2.5 Reference to Video 1080p
  • Seedance 2.5 1080p Text to video
  • Seedance 2.5 1080p
  • MiniMax H3 Open
  • Wan 2.6 Flash
  • Happy Horse 1.1 reference to video
  • View All Models →
Image Models
  • Qwen Image 3.0 Edit
  • Qwen Image 3.0 Pro Edit
  • Qwen Image 3.0
  • seedream 4.0
  • Flux 2 Flash Edit
  • Nano Banana Pro
  • View All Models →
Legal
  • Terms of Service
  • Privacy Policy
  • Cookie Policy
RunComfy
Copyright 2026 RunComfy. All Rights Reserved.

RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.

Examples Of Minimax H3 Image to Video

Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...