How to Use the alibaba/wan-2.2-spicy/image-to-video API: image to video integration guide (2026)

2 min · Updated 2026-07-11 · By the NamiFusion team

Learn how to call the alibaba/wan-2.2-spicy/image-to-video API on NamiFusion: parameters, pricing (from ≈$0.101) and runnable copy-paste code — try it in the playground first.

alibaba/wan-2.2-spicy/image-to-video is an image to video model by Alibaba: Turn a first frame and optional last frame into coherent motion video with 5/8-second, 480p/720p output and intelligent prompt expansion. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — about $0.101 per typical run.

TL;DR

· Task type: image to video, provided by Alibaba

· Pricing: about $0.101 per typical run ($1 = 100 credits), pay-as-you-go

· Getting started: playground (exact quote before submit) → create an API key → one POST request

· Tunable parameters: 7 (see the table below)

· No platform watermark; commercial use allowed (per the Terms of Service)

What can you build with alibaba/wan-2.2-spicy/image-to-video?

Short-form video and ad clips: go from a prompt or a single image to publishable motion content.

Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.

A social content factory: batch vertical clips and ride trending topics fast.

Remixing existing assets: image-to-video brings static work to life and extends its lifespan.

How do you call this API?

  1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).
  2. Submit a task with the cURL example below, passing your key as a Bearer token.
  3. Poll the task status with the returned task_uuid and read the output URLs when completed.
Submit a task
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/alibaba/wan-2.2-spicy/image-to-video" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "image": [
      "https://assets-public.namifusion.com/marketplace/images/2026-06-12/638e1ee18ac7.jpeg"
    ],
    "prompt": "Animate the subject with natural motion, a slow cinematic camera push-in, soft wind, stable anatomy, and consistent lighting.",
    "duration": 5,
    "resolution": "480p",
    "prompt_extend": true
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"

What parameters does it take?

image (Input Image): required

last_image (Last Frame): optional

prompt (Prompt): required

duration (Duration): optional, default 5

resolution (Resolution): optional, default "480p"

prompt_extend (Prompt Extend): optional, default true

seed (Seed): optional

How do you write prompts for alibaba/wan-2.2-spicy/image-to-video?

  1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").
  2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".
  3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.
  4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".

How much does it cost?

Pay-per-use in credits. About $0.101 per typical run ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.

alibaba/wan-2.2-spicy/image-to-video

Animate a still image into a controlled 5- or 8-second video with optional end-frame guidance.

Wan 2.2 Spicy Image-to-Video turns a source image and text prompt into a short video while preserving the subject and overall visual identity of the input. It supports two practical resolution tiers, optional last-frame control, prompt expansion, and a reproducible seed, making it suitable for both quick previews and polished content production.

Key Features

  • Image-led motion generation: Uses the input image as the visual anchor and adds subject, camera, and environmental motion.
  • Optional final-frame guidance: Supply last_image when the clip needs to converge toward a specific ending composition.
  • Two duration choices: Generate either 5-second or 8-second clips.
  • 480p and 720p output: Balance turnaround time, detail, and credit usage for each job.
  • Prompt expansion: Automatically enrich short prompts before generation when prompt_extend is enabled.
  • Seed control: Reuse a seed when you need comparable variations from the same input.

Technical Specifications

ItemDetails
TaskImage-to-video generation
Required inputimage, prompt
Optional inputlast_image, seed
Resolution480p (default), 720p
Duration5 seconds (default), 8 seconds
Prompt expansionSupported and enabled by default
OutputGenerated video URL

Sample Prompts

  • Animate the subject with natural breathing and subtle head movement, slow cinematic camera push-in, soft wind, stable anatomy, warm evening light.
  • The product remains centered while the camera makes a gentle orbit; reflections move naturally across the surface, premium studio lighting.
  • Clouds drift behind the character, fabric moves in a light breeze, shallow depth of field, smooth and coherent motion.

Pricing

Resolution5 seconds8 seconds
480p10 credits16 credits
720p20 credits32 credits

Prompt expansion adds 0.1 credit per request when enabled. These values reproduce the existing Marketplace pricing configuration; this documentation update does not alter pricing.

FAQ

Can I try it for free first?+

Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.

Can I use the output commercially? Is there a watermark?+

Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.

What is the content policy?+

Pornographic content, sexualized content involving a minor, unauthorized use of real people, deceptive impersonation, fraud and false endorsements are prohibited. Lawful, consensual, non-explicit adult themes may be permitted; see the Terms of Service.

How long does a alibaba/wan-2.2-spicy/image-to-video task take?+

It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.

What are the resolution and duration limits?+

See the parameter table above — each model lists its available resolution and duration options there and in the playground.

How to Use the alibaba/wan-2.2-spicy/image-to-video API: image to video integration guide (2026) | NamiFusion Blog