How to Use the seedance-2.0-mini/image-to-video-spicy API: image to video integration guide (2026)
Learn how to call the seedance-2.0-mini/image-to-video-spicy API on NamiFusion: parameters, pricing (from ≈$0.600) and runnable copy-paste code — try it in the playground first.
seedance-2.0-mini/image-to-video-spicy is an image to video model by Doubao: Seedance 2.0 Mini Spicy Image to Video is ByteDance's faster, lower-cost image-to-video model for cinematic multi-shot videos. It turns reference images and optional text prompts into narrative sequences with AI camera control, consistent characters, 480P / 720P / 1080P / 4K output, 4-15s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — about $0.600 per typical run.
TL;DR
· Task type: image to video, provided by Doubao
· Pricing: about $0.600 per typical run ($1 = 100 credits), pay-as-you-go
· Getting started: playground (exact quote before submit) → create an API key → one POST request
· Tunable parameters: 8 (see the table below)
· No platform watermark; commercial use allowed (per the Terms of Service)
What can you build with seedance-2.0-mini/image-to-video-spicy?
Short-form video and ad clips: go from a prompt or a single image to publishable motion content.
Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.
A social content factory: batch vertical clips and ride trending topics fast.
Remixing existing assets: image-to-video brings static work to life and extends its lifespan.
How do you call this API?
- Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).
- Submit a task with the cURL example below, passing your key as a Bearer token.
- Poll the task status with the returned task_uuid and read the output URLs when completed.
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2.0-mini/image-to-video-spicy" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"input": {
"first_frame": [
"https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/ae63692086b8.jpeg"
],
"prompt": "The woman slowly turns toward the camera and begins walking forward through the rain as passing headlights streak across the frame. The camera starts with a medium shot, then gently dollies in and arcs slightly around her for a cinematic reveal. Moody, sensual atmosphere with neon reflections, drifting steam, subtle wind in her hair, and realistic motion. High-end film look, rich contrast, smooth camera movement, consistent character.",
"last_frame": [
"https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/50b4fab428d3.jpeg"
],
"aspect_ratio": "16:9",
"resolution": "720p",
"duration": 6,
"generate_audio": true,
"seed": -1
}
}'
# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
-H "Authorization: Bearer YOUR_API_KEY"What parameters does it take?
first_frame (Image): required
prompt (Prompt): optional
last_frame (Last Frame): optional
aspect_ratio (Aspect Ratio): optional
resolution (Resolution): optional, default "720p"
duration (Duration): optional, default 5
generate_audio (Generate Audio): optional, default true
seed (Seed): optional
How do you write prompts for seedance-2.0-mini/image-to-video-spicy?
- Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").
- Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".
- Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.
- State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".
How much does it cost?
Pay-per-use in credits. About $0.600 per typical run ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.
seedance-2.0-mini/image-to-video-spicy
A faster, lower-cost image-to-video model for cinematic multi-shot storytelling with native audio.
doubao/seedance-2.0-mini/image-to-video-spicy is ByteDance’s lightweight image-to-video model designed for fast, affordable generation of narrative short-form videos. It transforms a reference image, optional Prompt, and optional ending-frame guidance into dynamic clips with camera-aware motion, character consistency, and flexible output settings from 480p to 4k.
This model is well suited for social content, creative prototyping, and cinematic short clips where speed and cost efficiency matter. Key advantages include no cold start, flexible Aspect Ratio support, native audio generation, and strong control over motion and camera language.
🚀 Key Features
- Image-to-video generation: Turn a single reference Image into a motion-driven Video sequence.
- Cinematic multi-shot output: Use Prompt guidance to describe action, mood, pacing, and camera behavior for more narrative results.
- Last-frame guidance: Provide
last_frameto influence the ending frame or continuation direction of the clip. - Native audio generation: Generate synchronized Audio together with the Video for more complete outputs.
- Flexible output options: Supports 16:9, 9:16, 4:3, 3:4, 1:1, and 21:9 Aspect Ratio settings, plus 480p, 720p, 1080p, and 4k Resolution.
- Fast and cost-efficient: Optimized for affordable production workloads with reliable performance and no cold starts.
🛠️ Technical Specifications
| Item | Details |
|---|---|
| Model architecture | Image-to-Video model for cinematic short-form and multi-shot generation |
| Input | Start Image URL, optional Prompt, optional last-frame Image URL |
| Output | Video with optional native Audio |
| Output Format | Video |
| Resolution | 480p, 720p(default), 1080p, 4k |
| Duration | 4–15 seconds(default: 5) |
| Aspect Ratio | 16:9, 9:16, 4:3, 3:4, 1:1, 21:9; adapts to the input Image if unspecified |
| Audio | Supported; synchronized native Audio generation |
| Seed | Supported; -1 uses a random Seed |
| Character consistency | Maintains subject continuity from the reference Image |
| Camera control | Prompt-based control for push-ins, pans, tracking, handheld motion, and more |
| Latency | Optimized for fast inference with no cold start |
Sample Prompts
A cinematic shot of the character slowly walking through a neon-lit street at night, soft rain falling, reflections on the wet pavement, gentle camera movement, atmospheric lighting, calm but dramatic mood.0-2s: The product slowly rotates on a studio table under soft lighting. 2-5s: The camera slides sideways to reveal metallic edges and material detail. No subtitles.The parked car in the image pulls out and drives down the wet street. 0-3s: headlights flick on, the car eases forward. 3-6s: the camera pans to follow from a low angle, neon reflections on the asphalt. Ambient and engine sounds only, no music.
💰 Pricing
| Billing | Resolution | Price |
|---|---|---|
| Per second | 480p | $0.06 |
| Per second | 720p | $0.12 |
| Per second | 1080p | $0.30 |
| Per second | 4k | $0.60 |
| Per 5 seconds | 480p | $0.30 |
| Per 5 seconds | 720p | $0.60 |
| Per 5 seconds | 1080p | $1.50 |
| Per 5 seconds | 4k | $3.00 |
Example Costs
FAQ
Can I try it for free first?+
Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.
Can I use the output commercially? Is there a watermark?+
Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.
What is the content policy?+
Pornographic content, sexualized content involving a minor, unauthorized use of real people, deceptive impersonation, fraud and false endorsements are prohibited. Lawful, consensual, non-explicit adult themes may be permitted; see the Terms of Service.
How long does a seedance-2.0-mini/image-to-video-spicy task take?+
It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.
What are the resolution and duration limits?+
See the parameter table above — each model lists its available resolution and duration options there and in the playground.