seedance-2.0-mini/image-to-video-spicy 图生视频模型

doubao/seedance-2.0-mini/image-to-video-spicy

Seedance 2.0 Mini Spicy Image to Video 是字节跳动推出的更快、更低成本的图生视频模型,适合生成电影感多镜头视频。它可将参考图片和可选文本提示词转换为叙事视频序列,支持 AI 镜头控制、角色一致性、480P / 720P / 1080P / 4K 输出、4-15 秒时长以及灵活的宽高比。提供开箱即用的 REST 推理 API,性能出色,无冷启动,价格实惠。

示例

seedance-2.0-mini/image-to-video-spicy 图生视频模型 示例 1

参数

名称类型默认约束说明
first_frame *图片image_upload—image/* · 0–1 items用于引导视频生成的起始图片 URL。
prompt提示词textarea—≤ 5000 chars描述视频中的场景、动作、镜头运动和氛围。
last_frame最后一张图片image_upload—image/* · 0–1 items用于续接视频的最后一帧图片 URL。
aspect_ratio宽高比select—16:9 | 9:16 | 4:3 | 3:4 | 1:1 | 21:9生成视频的宽高比。如未指定,将自动适配输入图片。
resolution分辨率select720p480p | 720p | 1080p | 4k输出视频的分辨率。
duration时长slider54 ~ 15 · step 1生成视频的时长(单位为秒,4-15 秒)。
generate_audio生成音频booleantrue—是否生成与输出视频同步的原生音频。默认为 true。
seed随机种子number——用于生成的随机种子。-1 表示将使用随机种子。

输出字段

字段类型说明
videosarray<string>

API

通过统一 REST API 调用本模型;在 API Keys 页获取密钥。

cURL
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2.0-mini/image-to-video-spicy" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "first_frame": [
      "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/ae63692086b8.jpeg"
    ],
    "prompt": "The woman slowly turns toward the camera and begins walking forward through the rain as passing headlights streak across the frame. The camera starts with a medium shot, then gently dollies in and arcs slightly around her for a cinematic reveal. Moody, sensual atmosphere with neon reflections, drifting steam, subtle wind in her hair, and realistic motion. High-end film look, rich contrast, smooth camera movement, consistent character.",
    "last_frame": [
      "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/50b4fab428d3.jpeg"
    ],
    "aspect_ratio": "16:9",
    "resolution": "720p",
    "duration": 6,
    "generate_audio": true,
    "seed": -1
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"
Python
import time, requests

API_KEY = "YOUR_API_KEY"
HEADERS = {"Authorization": f"Bearer {API_KEY}"}

# 1) Submit
resp = requests.post(
    "https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2.0-mini/image-to-video-spicy",
    headers=HEADERS,
    json={
        "input": {
            "first_frame": [
                "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/ae63692086b8.jpeg"
            ],
            "prompt": "The woman slowly turns toward the camera and begins walking forward through the rain as passing headlights streak across the frame. The camera starts with a medium shot, then gently dollies in and arcs slightly around her for a cinematic reveal. Moody, sensual atmosphere with neon reflections, drifting steam, subtle wind in her hair, and realistic motion. High-end film look, rich contrast, smooth camera movement, consistent character.",
            "last_frame": [
                "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/50b4fab428d3.jpeg"
            ],
            "aspect_ratio": "16:9",
            "resolution": "720p",
            "duration": 6,
            "generate_audio": True,
            "seed": -1
        }
    },
)
resp.raise_for_status()  # 401/402/429/5xx stop here instead of polling a bad task
task = resp.json()

# 2) Poll until a terminal state (completed / failed / cancelled).
#    This model is allowed up to 300s server-side.
deadline = time.time() + 360
while task.get("status") not in ("completed", "failed", "cancelled"):
    if time.time() > deadline:
        raise TimeoutError(f"still {task.get('status')} — keep the task_uuid and poll later")
    time.sleep(3)
    poll = requests.get(f"https://www.namifusion.com/api/v1/marketplace/run/tasks/{task['task_uuid']}", headers=HEADERS)
    poll.raise_for_status()
    task = poll.json()

print(task["status"], task.get("output"))
JavaScript
const API_KEY = "YOUR_API_KEY";
const HEADERS = { Authorization: `Bearer ${API_KEY}` };

// 1) Submit
const resp = await fetch("https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2.0-mini/image-to-video-spicy", {
  method: "POST",
  headers: { ...HEADERS, "Content-Type": "application/json" },
  body: JSON.stringify({
    "input": {
      "first_frame": [
        "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/ae63692086b8.jpeg"
      ],
      "prompt": "The woman slowly turns toward the camera and begins walking forward through the rain as passing headlights streak across the frame. The camera starts with a medium shot, then gently dollies in and arcs slightly around her for a cinematic reveal. Moody, sensual atmosphere with neon reflections, drifting steam, subtle wind in her hair, and realistic motion. High-end film look, rich contrast, smooth camera movement, consistent character.",
      "last_frame": [
        "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/50b4fab428d3.jpeg"
      ],
      "aspect_ratio": "16:9",
      "resolution": "720p",
      "duration": 6,
      "generate_audio": true,
      "seed": -1
    }
  }),
});
if (!resp.ok) throw new Error(`submit failed: ${resp.status} ${await resp.text()}`);
let task = await resp.json();

// 2) Poll until a terminal state (completed / failed / cancelled).
//    This model is allowed up to 300s server-side.
const deadline = Date.now() + 360 * 1000;
while (!["completed", "failed", "cancelled"].includes(task.status)) {
  if (Date.now() > deadline) throw new Error(`still ${task.status} — keep the task_uuid and poll later`);
  await new Promise((r) => setTimeout(r, 3000));
  const poll = await fetch(`https://www.namifusion.com/api/v1/marketplace/run/tasks/${task.task_uuid}`, { headers: HEADERS });
  if (!poll.ok) throw new Error(`poll failed: ${poll.status}`);
  task = await poll.json();
}

console.log(task.status, task.output);

文档

doubao/seedance-2.0-mini/image-to-video-spicy

更快、更低成本的电影级图片转视频模型,支持多镜头叙事与原生音频。

doubao/seedance-2.0-mini/image-to-video-spicy 是 ByteDance 推出的轻量级图片转视频模型,面向需要高性价比与快速交付的创作者和开发者。它可基于起始图片、可选提示词以及可选结束帧参考,生成具有镜头语言、角色一致性和叙事感的短视频,并支持 480p 到 4k 的多档分辨率输出。

该模型特别适合制作电影感短片、社交媒体内容和创意预演。其优势包括无冷启动、灵活宽高比、原生音频生成以及可控的镜头运动表达,在速度、成本与质量之间提供了出色平衡。

🚀 关键特性

  • 图片驱动视频生成:以单张起始图片为基础生成动态视频,适合角色、产品、场景等内容的动画化。
  • 多镜头叙事能力:支持通过提示词描述动作、镜头运动、氛围与节奏,生成更具故事性的连续画面。
  • 结束帧引导:可通过 last_frame 指定结尾方向或续接目标,提升视频首尾控制能力。
  • 原生音频同步生成:支持与视频同步生成音频,适合需要环境声或动作声的短视频内容。
  • 灵活输出规格:支持 16:9、9:16、4:3、3:4、1:1、21:9 等多种宽高比,以及 480p、720p、1080p、4k 分辨率。
  • 高性价比与无冷启动:面向高频调用场景优化,兼顾低成本、稳定响应与快速出图体验。

🛠️ 技术规格

项目说明
模型架构图片转视频模型(Image-to-Video),面向电影感多镜头短视频生成
输入内容图片 URL、可选提示词、可选结束帧图片 URL
输出内容视频(可选原生音频)
输出格式视频
分辨率480p、720p(默认)、1080p、4k
时长4–15 秒(默认 5 秒)
宽高比16:9、9:16、4:3、3:4、1:1、21:9;未指定时可自适应输入图片
音频支持,可生成与视频同步的原生音频
随机种子支持;-1 表示使用随机随机种子
角色一致性支持基于参考图片保持主体外观连续性
镜头控制支持通过提示词描述推镜、平移、跟拍、摇镜等镜头语言
延迟表现面向快速推理优化;无冷启动,适合稳定生产使用

示例提示词

  1. 画面中的角色缓慢走过夜晚霓虹街道,细雨落下,地面有湿润反光,镜头缓慢推进,氛围安静但富有戏剧感。
  2. 0-2 秒:产品从静止状态缓慢旋转,柔和棚拍光。2-5 秒:镜头侧向移动,突出金属边缘与材质细节。无字幕。
  3. 图中的汽车启动并驶离路边。0-3 秒:车灯亮起,车辆缓慢前进。3-6 秒:镜头低角度跟随,雨夜路面反射霓虹灯光。仅保留环境音,无背景音乐。

💰 价格

计费方式分辨率价格
每秒480p$0.06
每秒720p$0.12
每秒1080p$0.30
每秒4k$0.60
每 5 秒480p$0.30
每 5 秒720p$0.60
每 5 秒1080p$1.50
每 5 秒4k$3.00

示例成本

分辨率4 秒5 秒10 秒15 秒
480p$0.24$0.30$0.60$0.90
720p$0.48$0.60$1.20$1.80
1080p$1.20$1.50$3.00$4.50
4k$2.40$3.00$6.00$9.00

💡 最佳使用场景

  • 社交媒体短视频制作:支持横屏、竖屏、正方形与超宽屏输出,适合短视频平台内容生产。
  • 电商与产品展示:将静态产品图转为带镜头运动的展示视频,提升商品表现力。
  • 影视预演与创意分镜:快速验证镜头调度、氛围、动作节奏与场景方向。
  • 角色与场景动画化:基于角色设定图、插画或概念图生成具有连续性的动态片段。

🔗 相关模型

  • ByteDance Seedance 2.0 Mini Image-to-Video:标准版图片转视频模型,适合通用图生视频任务。
  • ByteDance Seedance 2.0 Mini Text-to-Video:从提示词直接生成视频,适合无参考图片的创作流程。

Nami API 调用

请求体为 JSON 对象 { "input": { ... } }。使用 X-API-Key 认证。图片参数是只含一个 URL 的字符串数组。下表的默认值仅在省略参数时生效;playground 和 API 示例采用下方示例值。

POST /api/v1/marketplace/run/doubao/seedance-2.0-mini/image-to-video-spicy

ParameterTypeRequiredDefaultConstraints
first_framestring[]yes—1 URL
promptstringno—max 5000 characters
last_framestring[]no—0–1 URL
aspect_ratiostringno—16:9, 9:16, 4:3, 3:4, 1:1, 21:9
resolutionstringno"720p"480p, 720p, 1080p, 4k
durationintegerno54–15; step 1
generate_audiobooleannotrue—
seedintegerno——
{
  "input": {
    "first_frame": [
      "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/ae63692086b8.jpeg"
    ],
    "prompt": "The woman slowly turns toward the camera and begins walking forward through the rain as passing headlights streak across the frame. The camera starts with a medium shot, then gently dollies in and arcs slightly around her for a cinematic reveal. Moody, sensual atmosphere with neon reflections, drifting steam, subtle wind in her hair, and realistic motion. High-end film look, rich contrast, smooth camera movement, consistent character.",
    "last_frame": [
      "https://d1s3annbwribom.cloudfront.net/marketplace/examples/2026/09/12/50b4fab428d3.jpeg"
    ],
    "aspect_ratio": "16:9",
    "resolution": "720p",
    "duration": 6,
    "generate_audio": true,
    "seed": -1
  }
}

创建请求返回 task_uuid 和任务状态。使用同一个 API Key 轮询:

GET /api/v1/marketplace/run/tasks/{task_uuid}

完成后从 output.videos 读取视频 URL 数组。音频(如启用)包含在视频中。

{
  "status": "completed",
  "output": {
    "videos": [
      "https://cdn.example.com/output/video.mp4"
    ]
  }
}

以上为响应字段片段;完整任务结构见 API tab。

价格按实际选择的分辨率和时长计算,1 美元 = 100 积分;以下为目录标价。

ResolutionUSD / secondCredits / secondUSD / 5 seconds
480p$0.066$0.30
720p$0.1212$0.60
1080p$0.3030$1.50
4k$0.6060$3.00

音频开关不影响标价。

相关模型

xAI Grok Imagine Video v1.5 图片转视频
图生视频X Ai

xAI Grok Imagine Video v1.5 图片转视频

使用文本提示词将单张输入图片生成 1-15 秒视频,支持 480p 和 720p。

起 $0.850 / 每次
vidu/q3/image-to-video-spicy 图片转视频
图生视频Vidu

vidu/q3/image-to-video-spicy 图片转视频

Vidu Q3 Image-to-Video Spicy 可从图片生成不限量的高质量视频,动画流畅、运动表现丰富,并针对大规模内容生成进行了优化。提供开箱即用的 REST 推理 API,性能出色,无冷启动,价格实惠。

起 $0.750 / 每次
Vidu Q3 图片转视频
图生视频Vidu

Vidu Q3 图片转视频

Vidu Q3 图片转视频将文本提示词转化为高质量视频,具有卓越的视觉保真度和多样的运动。即用型 REST 推理 API,最佳性能,无冷启动,价格实惠。

起 $0.750 / 每次
MiniMax H3 Spicy 图片转视频
图生视频Minimax

MiniMax H3 Spicy 图片转视频

通过首帧图片生成视频,支持可选提示词和末帧引导。可选 480p、540p、768p、1080p 分辨率,时长为 3–15 秒。

起 $0.200 / 每次
MiniMax H3 图片转视频
图生视频Minimax

MiniMax H3 图片转视频

MiniMax H3 图片转视频可将首帧图片动画化为连贯的 2K 视频,支持自然语言运动指令,并可选用末帧控制,以实现一致的运动、场景连贯性和电影感视频生成。提供开箱即用的 REST 推理 API,性能出色,无冷启动,价格实惠。

起 $0.650 / 每次
ltx-2.5/image-to-video 图片转视频
图生视频Ltx 2.5

ltx-2.5/image-to-video 图片转视频

LTX 2.5 图片转视频可将首帧图片动画化,生成高保真、音视频同步的内容,并支持可选的末帧引导以及 720P / 1080P / 2K / 4K 输出,适用于电影感视频、社交内容、广告和制作流程。提供开箱即用的 REST 推理 API,性能出色,无冷启动,价格实惠。

起 $0.500 / 每次
ai/ltx-2.3-spicy/image-to-video-lora 图片转视频
图生视频Lightricks

ai/ltx-2.3-spicy/image-to-video-lora 图片转视频

LTX 2.3 Spicy LoRA 图片转视频 API 可将参考图片和提示词转换为富有表现力的 AI 视频,并支持选择 LoRA 预设、可选的 LoRA 强度覆盖、时长和分辨率控制。在 MaaS 上可快速进行 REST 推理,无冷启动,价格实惠。

起 $0.150 / 每次
ai/ltx-2.3-spicy/image-to-video 图片转视频
图生视频Lightricks

ai/ltx-2.3-spicy/image-to-video 图片转视频

LTX 2.3 Spicy Image-to-Video API 可根据参考图片和提示词生成富有表现力的 AI 视频。可选择风格预设、时长和分辨率,依托 MaaS 提供快速 REST 推理、无冷启动和实惠价格。

起 $0.100 / 每次