How to Use the xAI Grok Imagine Video v1.5 Reference to Video API: reference to video integration guide (2026)
Learn how to call the xAI Grok Imagine Video v1.5 Reference to Video API on NamiFusion: parameters, pricing (from ≈$0.840) and runnable copy-paste code — try it in the playground first.
xAI Grok Imagine Video v1.5 Reference to Video is a reference to video model by X Ai: Generate a 1-15 second video from a prompt and 1-7 reference images at 480p or 720p. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — a typical run costs about $0.840.
TL;DR
· Task type: reference to video, provided by X Ai
· Pricing: about $0.840 per typical run ($1 = 100 credits), pay-as-you-go
· Getting started: playground (exact quote before submit) → create an API key → one POST request
· Tunable parameters: 5 (see the table below)
· No platform watermark; commercial use allowed (per the Terms of Service)
What can you build with xAI Grok Imagine Video v1.5 Reference to Video?
Short-form video and ad clips: go from a prompt or a single image to publishable motion content.
Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.
A social content factory: batch vertical clips and ride trending topics fast.
Remixing existing assets: image-to-video brings static work to life and extends its lifespan.
How do you call this API?
1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).
2. Submit a task with the cURL example below, passing your key as a Bearer token.
3. Poll the task status with the returned task_uuid and read the output URLs when completed.
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/x-ai/grok-imagine-video-v1.5/reference-to-video" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"input": {
"prompt": "A cinematic portrait, dramatic rim light, shallow depth of field",
"images": [
"https://example.com/input.jpg"
],
"duration": 6,
"aspect_ratio": "16:9",
"resolution": "720p"
}
}'
# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
-H "Authorization: Bearer YOUR_API_KEY"What parameters does it take?
prompt (Prompt): required
images (Reference Images): required
duration (Duration): optional, default 6
aspect_ratio (Aspect Ratio): optional, default "16:9"
resolution (Resolution): optional, default "720p"
How do you write prompts for xAI Grok Imagine Video v1.5 Reference to Video?
1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").
2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".
3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.
4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".
How much does it cost?
Pay-per-use in credits. A typical run costs about $0.840 ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.
xAI Grok Imagine Video v1.5 Reference to Video
Turn up to seven reference images into identity-consistent, stylized short videos.
🚀 Key Features
- Reference-guided generation: Use 1 to 7 reference images to anchor subject identity, appearance, and visual style.
- Prompt-based motion control: Describe motion, camera behavior, atmosphere, and scene progression using natural language.
- Identity and style consistency: Multiple references help maintain a coherent character, brand look, or aesthetic across the clip.
- Flexible output settings: Choose between 480p and 720p, with multiple aspect ratios for landscape, square, and vertical delivery.
- Production-ready performance: Suitable for character content, concept visualization, ad creatives, and stylized storytelling.
🛠️ Technical Specifications
| Item | Details |
|---|---|
| Model architecture | Reference-guided image-to-video generation model |
| Input | Prompt + 1 to 7 reference images |
| Output Format | Video |
| Reference image count | Up to 7 images |
| Duration | 1 to 15 seconds(default: 6s) |
| Aspect Ratio | 16:9, 1:1, 9:16, 3:2, 2:3(default: 16:9) |
| Resolution | 720p, 480p(default: 720p) |
| Motion/style control | Jointly controlled by reference images and prompt |
| Latency | Optimized for fast inference; actual runtime varies by duration, resolution, and number of reference images |
| Ideal tasks | Character-driven clips, identity-consistent videos, creative prototyping, marketing content |
Sample Prompts
- A stylish woman walks slowly through a city street at dusk, hair moving in the wind, smooth tracking shot, cinematic lighting.
- The character from the reference images turns toward the camera under neon lights, subtle depth of field, dreamy modern atmosphere.
- Animate the subject with natural hand movement and a gentle smile, with the camera slowly pushing from medium shot to close-up for a social promo feel.
💰 Pricing
| Pricing Item | Price |
|---|---|
| Base Price | $0.09 |
| 480p | $0.08 / second |
| 720p | $0.14 / second |
| Additional reference image | +$0.01 each |
| Billing rule | Linear by duration, rounded up to the next whole second |
Example Pricing(1 reference image)
| Resolution | 1s | 5s | 10s | 15s |
|---|---|---|---|---|
| 480p | $0.09 | $0.41 | $0.81 | $1.21 |
| 720p | $0.15 | $0.71 | $1.41 | $2.11 |
FAQ
Can I try it for free first?+
Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.
Can I use the output commercially? Is there a watermark?+
Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.
What is the content policy?+
NamiFusion is an unfiltered engine: no arbitrary restrictions and NSFW-friendly, but illegal content and non-consensual use of real people are prohibited — see the Terms of Service.
How long does a xAI Grok Imagine Video v1.5 Reference to Video task take?+
It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.
What are the resolution and duration limits?+
See the parameter table above — each model lists its available resolution and duration options there and in the playground.