How to Use the MiniMax H3 Image to Video API: image to video integration guide (2026)
Learn how to call the MiniMax H3 Image to Video API on NamiFusion: parameters, pricing (from ≈$0.650) and runnable copy-paste code — try it in the playground first.
MiniMax H3 Image to Video is an image to video model by Minimax: MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — a typical run costs about $0.650.
TL;DR
· Task type: image to video, provided by Minimax
· Pricing: about $0.650 per typical run ($1 = 100 credits), pay-as-you-go
· Getting started: playground (exact quote before submit) → create an API key → one POST request
· Tunable parameters: 5 (see the table below)
· No platform watermark; commercial use allowed (per the Terms of Service)
What can you build with MiniMax H3 Image to Video?
Short-form video and ad clips: go from a prompt or a single image to publishable motion content.
Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.
A social content factory: batch vertical clips and ride trending topics fast.
Remixing existing assets: image-to-video brings static work to life and extends its lifespan.
How do you call this API?
1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).
2. Submit a task with the cURL example below, passing your key as a Bearer token.
3. Poll the task status with the returned task_uuid and read the output URLs when completed.
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/minimax/h3/image-to-video" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"input": {
"first_frame": "https://example.com/input.jpg",
"prompt": "A cinematic portrait, dramatic rim light, shallow depth of field",
"resolution": "2k",
"duration": 5
}
}'
# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
-H "Authorization: Bearer YOUR_API_KEY"What parameters does it take?
first_frame (Image): required
prompt (Prompt): required
last_frame (Last Image): optional
resolution (Resolution): optional, default "2k"
duration (Duration): optional, default 5
How do you write prompts for MiniMax H3 Image to Video?
1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").
2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".
3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.
4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".
How much does it cost?
Pay-per-use in credits. A typical run costs about $0.650 ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.
MiniMax H3 Image to Video
Turn a single first-frame image into a coherent 2K video with prompt-driven motion control.
MiniMax H3 Image to Video is an image-to-video model that transforms a starting image into a high-resolution 2K video using natural-language motion instructions. With optional last-frame guidance, it is well suited for workflows that require stronger scene continuity, controlled progression, and cinematic short-form video generation. It also stands out for no cold start behavior and cost-efficient pricing.
🚀 Key Features
- First-frame guided generation: Uses the input image to anchor subject identity, composition, style, and opening visual state.
- Prompt-based motion control: Define movement, camera behavior, lighting, mood, and scene progression with natural language.
- Optional last-frame guidance: Add a final reference image to better control the ending frame and improve visual continuity.
- Native 2K output: Generates high-resolution 2K video suitable for premium creative and commercial content.
- Flexible duration options: Supports video lengths from 4 to 15 seconds for both rapid iteration and more developed scenes.
- No cold start: Designed for responsive production usage with more predictable turnaround time.
🛠️ Technical Specifications
| Item | Details |
|---|---|
| Model Name | MiniMax H3 Image to Video |
| Model ID | minimax/h3/image-to-video |
| Task Type | image-to-video |
| Model Architecture | First-frame conditioned video generation model with optional last-frame control |
| Input Format | Image(URL or Base64) + Prompt; optional last-frame image(URL or Base64) |
| Output Format | Video |
| Input Image Requirements | 256–5760 pixels per side; supported aspect ratio from 0.4 to 2.5 |
| Prompt Length | 1–7000 characters |
| Resolution | 2K |
| Duration | 4–15 seconds |
| Fps | Not publicly specified |
| Last-frame Control | Supported(optional) |
| Latency | No cold start; actual inference time varies by workload and duration |
Sample Prompts
Slow cinematic push-in on a person standing in a rainy neon-lit street, subtle head movement, reflections on wet pavement, realistic lighting, moody atmosphere.Animate the product with a slow rotating motion on a clean studio background, slight camera orbit, premium commercial lighting, emphasize metallic texture and highlights.A white cat jumps down from a windowsill and walks toward the camera, warm sunlight filling the room, natural motion, soft and cozy mood.
💰 Pricing
| Option | Price |
|---|---|
| Base Price | $0.13 / second |
| 5s video | $0.65 |
| 10s video | $1.30 |
| 15s video | $1.95 |
FAQ
Can I try it for free first?+
Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.
Can I use the output commercially? Is there a watermark?+
Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.
What is the content policy?+
NamiFusion is an unfiltered engine: no arbitrary restrictions and NSFW-friendly, but illegal content and non-consensual use of real people are prohibited — see the Terms of Service.
How long does a MiniMax H3 Image to Video task take?+
It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.
What are the resolution and duration limits?+
See the parameter table above — each model lists its available resolution and duration options there and in the playground.