How to Use the alibaba/wan-2.2-spicy/image-to-video API: image to video integration guide (2026)
Learn how to call the alibaba/wan-2.2-spicy/image-to-video API on NamiFusion: parameters, pricing (from ≈$0.101) and runnable copy-paste code — try it in the playground first.
alibaba/wan-2.2-spicy/image-to-video is an image to video model by Alibaba: Turn a first frame and optional last frame into coherent motion video with 5/8-second, 480p/720p output and intelligent prompt expansion. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — about $0.101 per typical run.
TL;DR
· Task type: image to video, provided by Alibaba
· Pricing: about $0.101 per typical run ($1 = 100 credits), pay-as-you-go
· Getting started: playground (exact quote before submit) → create an API key → one POST request
· Tunable parameters: 7 (see the table below)
· No platform watermark; commercial use allowed (per the Terms of Service)
What can you build with alibaba/wan-2.2-spicy/image-to-video?
Short-form video and ad clips: go from a prompt or a single image to publishable motion content.
Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.
A social content factory: batch vertical clips and ride trending topics fast.
Remixing existing assets: image-to-video brings static work to life and extends its lifespan.
How do you call this API?
- Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).
- Submit a task with the cURL example below, passing your key as a Bearer token.
- Poll the task status with the returned task_uuid and read the output URLs when completed.
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/alibaba/wan-2.2-spicy/image-to-video" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"input": {
"image": [
"https://assets-public.namifusion.com/marketplace/images/2026-06-12/638e1ee18ac7.jpeg"
],
"prompt": "Animate the subject with natural motion, a slow cinematic camera push-in, soft wind, stable anatomy, and consistent lighting.",
"duration": 5,
"resolution": "480p",
"prompt_extend": true
}
}'
# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
-H "Authorization: Bearer YOUR_API_KEY"What parameters does it take?
image (Input Image): required
last_image (Last Frame): optional
prompt (Prompt): required
duration (Duration): optional, default 5
resolution (Resolution): optional, default "480p"
prompt_extend (Prompt Extend): optional, default true
seed (Seed): optional
How do you write prompts for alibaba/wan-2.2-spicy/image-to-video?
- Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").
- Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".
- Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.
- State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".
How much does it cost?
Pay-per-use in credits. About $0.101 per typical run ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.
alibaba/wan-2.2-spicy/image-to-video
Animate a still image into a controlled 5- or 8-second video with optional end-frame guidance.
Wan 2.2 Spicy Image-to-Video turns a source image and text prompt into a short video while preserving the subject and overall visual identity of the input. It supports two practical resolution tiers, optional last-frame control, prompt expansion, and a reproducible seed, making it suitable for both quick previews and polished content production.
Key Features
- Image-led motion generation: Uses the input image as the visual anchor and adds subject, camera, and environmental motion.
- Optional final-frame guidance: Supply
last_imagewhen the clip needs to converge toward a specific ending composition. - Two duration choices: Generate either 5-second or 8-second clips.
- 480p and 720p output: Balance turnaround time, detail, and credit usage for each job.
- Prompt expansion: Automatically enrich short prompts before generation when
prompt_extendis enabled. - Seed control: Reuse a seed when you need comparable variations from the same input.
Technical Specifications
| Item | Details |
|---|---|
| Task | Image-to-video generation |
| Required input | image, prompt |
| Optional input | last_image, seed |
| Resolution | 480p (default), 720p |
| Duration | 5 seconds (default), 8 seconds |
| Prompt expansion | Supported and enabled by default |
| Output | Generated video URL |
Sample Prompts
Animate the subject with natural breathing and subtle head movement, slow cinematic camera push-in, soft wind, stable anatomy, warm evening light.The product remains centered while the camera makes a gentle orbit; reflections move naturally across the surface, premium studio lighting.Clouds drift behind the character, fabric moves in a light breeze, shallow depth of field, smooth and coherent motion.
Pricing
| Resolution | 5 seconds | 8 seconds |
|---|---|---|
| 480p | 10 credits | 16 credits |
| 720p | 20 credits | 32 credits |
Prompt expansion adds 0.1 credit per request when enabled. These values reproduce the existing Marketplace pricing configuration; this documentation update does not alter pricing.
FAQ
Can I try it for free first?+
Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.
Can I use the output commercially? Is there a watermark?+
Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.
What is the content policy?+
Pornographic content, sexualized content involving a minor, unauthorized use of real people, deceptive impersonation, fraud and false endorsements are prohibited. Lawful, consensual, non-explicit adult themes may be permitted; see the Terms of Service.
How long does a alibaba/wan-2.2-spicy/image-to-video task take?+
It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.
What are the resolution and duration limits?+
See the parameter table above — each model lists its available resolution and duration options there and in the playground.