MiniMax H3 Reference to Video
minimax/h3/reference-to-video
minimax/h3/reference-to-video. Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et.
İlgili modeller
Tüm parametreler, çıktı şeması ve güncel fiyatlar
| Yapay zeka model API pazarı | Tek platformda her şey | En popüler | Koşullar | Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et. |
|---|---|---|---|---|
| prompt *Prompt | textarea | — | ≤ 7000 chars | |
| imagesReference Images | image_upload | — | image/* · 1–9 items | |
| audiosReference Audios | audio_upload_group | — | audio/* · 0–3 items | |
| aspect_ratioAspect Ratio | select | adaptive | 21:9 | 16:9 | 4:3 | 1:1 | 3:4 | 9:16 | adaptive | |
| resolutionResolution | select | 2k | 2k | |
| durationDuration | slider | 5 | 4 ~ 15 · step 1 |
Tüm parametreler, çıktı şeması ve güncel fiyatlar
| Daha fazla bilgi | Tek platformda her şey | Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et. |
|---|---|---|
| videos | array<string> |
API
Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et.
cURL
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/minimax/h3/reference-to-video" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"input": {
"prompt": "Use the reference subjects consistently in a cinematic scene with natural movement, coherent scale, and realistic lighting.",
"aspect_ratio": "adaptive",
"resolution": "2k",
"duration": 5
}
}'
# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
-H "Authorization: Bearer YOUR_API_KEY"Python
import time, requests
API_KEY = "YOUR_API_KEY"
HEADERS = {"Authorization": f"Bearer {API_KEY}"}
# 1) Submit
resp = requests.post(
"https://www.namifusion.com/api/v1/marketplace/run/minimax/h3/reference-to-video",
headers=HEADERS,
json={
"input": {
"prompt": "Use the reference subjects consistently in a cinematic scene with natural movement, coherent scale, and realistic lighting.",
"aspect_ratio": "adaptive",
"resolution": "2k",
"duration": 5
}
},
)
resp.raise_for_status() # 401/402/429/5xx stop here instead of polling a bad task
task = resp.json()
# 2) Poll until a terminal state (completed / failed / cancelled).
# This model is allowed up to 300s server-side.
deadline = time.time() + 360
while task.get("status") not in ("completed", "failed", "cancelled"):
if time.time() > deadline:
raise TimeoutError(f"still {task.get('status')} — keep the task_uuid and poll later")
time.sleep(3)
poll = requests.get(f"https://www.namifusion.com/api/v1/marketplace/run/tasks/{task['task_uuid']}", headers=HEADERS)
poll.raise_for_status()
task = poll.json()
print(task["status"], task.get("output"))JavaScript
const API_KEY = "YOUR_API_KEY";
const HEADERS = { Authorization: `Bearer ${API_KEY}` };
// 1) Submit
const resp = await fetch("https://www.namifusion.com/api/v1/marketplace/run/minimax/h3/reference-to-video", {
method: "POST",
headers: { ...HEADERS, "Content-Type": "application/json" },
body: JSON.stringify({
"input": {
"prompt": "Use the reference subjects consistently in a cinematic scene with natural movement, coherent scale, and realistic lighting.",
"aspect_ratio": "adaptive",
"resolution": "2k",
"duration": 5
}
}),
});
if (!resp.ok) throw new Error(`submit failed: ${resp.status} ${await resp.text()}`);
let task = await resp.json();
// 2) Poll until a terminal state (completed / failed / cancelled).
// This model is allowed up to 300s server-side.
const deadline = Date.now() + 360 * 1000;
while (!["completed", "failed", "cancelled"].includes(task.status)) {
if (Date.now() > deadline) throw new Error(`still ${task.status} — keep the task_uuid and poll later`);
await new Promise((r) => setTimeout(r, 3000));
const poll = await fetch(`https://www.namifusion.com/api/v1/marketplace/run/tasks/${task.task_uuid}`, { headers: HEADERS });
if (!poll.ok) throw new Error(`poll failed: ${poll.status}`);
task = await poll.json();
}
console.log(task.status, task.output);API belgeleri
MiniMax H3 Reference to Video
Reference-guided 2K video generation for stronger consistency, motion control, and scene continuity.
MiniMax H3 Reference to Video is a multimodal image-to-video model that generates coherent 2K videos from natural-language prompts plus reference media. By combining text with images, videos, and optional audio, it offers stronger control over subject identity, motion, timing, visual style, and overall scene progression—while also delivering no cold start behavior and cost-efficient production.
🚀 Key Features
- Multimodal reference guidance: Generate videos from a Prompt combined with reference images, reference videos, and optional reference audio for more controllable results.
- Stronger subject consistency: Use reference images to preserve characters, products, objects, or brand visuals across the generated video.
- Motion and camera guidance: Reference videos help steer movement, pacing, interactions, and camera behavior for more directed outputs.
- Native 2K output: Produces high-resolution 2K videos suitable for polished creative, commercial, and concept work.
- Flexible Aspect Ratio support: Supports 21:9, 16:9, 4:3, 1:1, 3:4, 9:16, and Adaptive for cinematic, social, square, and vertical formats.
- No cold start, production-friendly pricing: Well suited for reliable generation workflows, rapid iteration, and scalable content creation.
🛠️ Technical Specifications
| Item | Details |
|---|---|
| Model Name | MiniMax H3 Reference to Video |
| Model ID | minimax/h3/reference-to-video |
| Model Architecture | Reference-guided image-to-video generation model with multimodal conditioning |
| Task Type | reference-to-video |
| Input Format | Prompt + reference images and/or reference videos, with optional reference audio |
| Output Format | Video |
| Resolution | 2K(fixed) |
| Duration | 4–15 seconds |
| Aspect Ratio | 21:9, 16:9, 4:3, 1:1, 3:4, 9:16, Adaptive |
| Reference Images | Up to 9 |
| Reference Videos | Up to 3 |
| Reference Audio | Up to 3 files(cannot be used alone) |
| Input Requirement | At least one reference image or reference video is required |
| Latency | No cold start highlighted by the provider |
| Primary Strengths | Subject consistency, motion guidance, style control, timing, scene continuity |
Sample Prompts
A cinematic woman in a futuristic jacket walks through a neon-lit street after rain, slow camera push-in, reflective pavement, dramatic lighting, highly detailed.Create a premium product commercial using the reference product images, with smooth orbit camera movement, clean studio background, elegant lighting, and refined motion.Using the reference character image and motion video, generate a sunset beach scene where the character runs forward, turns back, and smiles, with gentle tracking camera movement.
💰 Pricing
| Pricing Item | Price |
|---|---|
| Base Price | $0.13 per output second |
| 5s output | $0.65 |
| 10s output | $1.30 |
| 15s output | $1.95 |
Billing Rules
| Item | Rule |
|---|---|
| Reference images | First 5 images are free; each additional image costs $0.04 |
| Reference videos | Billed at $0.13 per second based on combined duration |
| Reference audio | Supported audio inputs are free |
| Reference video limits | Up to 3 videos, 15 seconds total per request |
💡 Best Use Cases
- Commercial and e-commerce video creation: Turn product shots, brand assets, and reference footage into polished short-form marketing videos.
- Character and IP consistency workflows: Maintain a stable visual identity for people, avatars, mascots, or branded subjects across generated clips.
- Storyboarding and concept development: Rapidly prototype scenes, camera moves, and visual direction from written ideas plus reference media.
- Social and mobile-first content: Produce vertical, square, or widescreen videos tailored to different publishing channels.
🔗 Related Models
- MiniMax H3 Text-to-Video: Best for generating high-resolution video directly from text-only prompts.
- MiniMax H3 Image-to-Video: Best for generating video from a first-frame image plus a Prompt.
Best Practices
- Use clear, relevant reference images when subject identity, product detail, or style consistency matters most.
- Use reference videos when motion, pacing, interaction, or camera movement is the priority.
- Keep the Prompt focused on visible action, scene progression, camera language, lighting, and mood.
- Start with shorter Duration values for fast iteration, then extend length for more developed scenes.
- Match the Aspect Ratio to the target channel: 16:9 for standard widescreen, 9:16 for mobile vertical, and 1:1 for square social content.
İlgili modeller
xAI Grok Imagine Video v1.5 Reference to Video
x-ai/grok-imagine-video-v1.5/reference-to-video. Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et.
Gemini Omni Flash Reference to Video API
google/gemini-omni-flash/reference-to-video. Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et.
Seedance 2.0 Mini Reference to Video
doubao/seedance-2.0-mini/reference-to-video. Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et.
Seedance 2.5 Reference to Video
doubao/seedance-2-5/reference-to-video. Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et.
Seedance 2.0 Fast reference-to-video
doubao/seedance-2-0-fast/reference-to-video. Görsel, video ve yüz değiştirme model API’lerine göz at. Bir modeli dene ve birleşik REST API ile entegre et.