Seedance 2.0 Mini Text to Video

doubao/seedance-2.0-mini/text-to-video

doubao/seedance-2.0-mini/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

โมเดลที่เกี่ยวข้อง

พารามิเตอร์ทั้งหมด โครงสร้างผลลัพธ์ และราคาปัจจุบัน

ตลาด API โมเดล AIแพลตฟอร์มเดียว ครบทุกอย่างยอดนิยมที่สุดข้อกำหนดเลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์
prompt *Prompttextarea—≤ 5000 chars
ratioAspect Ratioselect16:916:9 | 9:16 | 4:3 | 3:4 | 1:1 | 21:9
resolutionResolutionselect720p480p | 720p
durationDurationslider54 ~ 15 · step 1
generate_audioGenerate Audiobooleantrue—
enable_web_searchEnable Web Searchbooleanfalse—
return_last_framereturn the last frame boolean——

พารามิเตอร์ทั้งหมด โครงสร้างผลลัพธ์ และราคาปัจจุบัน

เรียนรู้เพิ่มเติมแพลตฟอร์มเดียว ครบทุกอย่างเลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์
idstring
videosarray<string>

API

เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

cURL
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2.0-mini/text-to-video" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "prompt": "A tranquil mountain lake at sunrise, golden light crossing the water, drifting mist, slow cinematic camera movement, photorealistic.",
    "ratio": "16:9",
    "resolution": "720p",
    "duration": 5,
    "generate_audio": true,
    "enable_web_search": false
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"
Python
import time, requests

API_KEY = "YOUR_API_KEY"
HEADERS = {"Authorization": f"Bearer {API_KEY}"}

# 1) Submit
resp = requests.post(
    "https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2.0-mini/text-to-video",
    headers=HEADERS,
    json={
        "input": {
            "prompt": "A tranquil mountain lake at sunrise, golden light crossing the water, drifting mist, slow cinematic camera movement, photorealistic.",
            "ratio": "16:9",
            "resolution": "720p",
            "duration": 5,
            "generate_audio": True,
            "enable_web_search": False
        }
    },
)
resp.raise_for_status()  # 401/402/429/5xx stop here instead of polling a bad task
task = resp.json()

# 2) Poll until a terminal state (completed / failed / cancelled).
#    This model is allowed up to 300s server-side.
deadline = time.time() + 360
while task.get("status") not in ("completed", "failed", "cancelled"):
    if time.time() > deadline:
        raise TimeoutError(f"still {task.get('status')} — keep the task_uuid and poll later")
    time.sleep(3)
    poll = requests.get(f"https://www.namifusion.com/api/v1/marketplace/run/tasks/{task['task_uuid']}", headers=HEADERS)
    poll.raise_for_status()
    task = poll.json()

print(task["status"], task.get("output"))
JavaScript
const API_KEY = "YOUR_API_KEY";
const HEADERS = { Authorization: `Bearer ${API_KEY}` };

// 1) Submit
const resp = await fetch("https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2.0-mini/text-to-video", {
  method: "POST",
  headers: { ...HEADERS, "Content-Type": "application/json" },
  body: JSON.stringify({
    "input": {
      "prompt": "A tranquil mountain lake at sunrise, golden light crossing the water, drifting mist, slow cinematic camera movement, photorealistic.",
      "ratio": "16:9",
      "resolution": "720p",
      "duration": 5,
      "generate_audio": true,
      "enable_web_search": false
    }
  }),
});
if (!resp.ok) throw new Error(`submit failed: ${resp.status} ${await resp.text()}`);
let task = await resp.json();

// 2) Poll until a terminal state (completed / failed / cancelled).
//    This model is allowed up to 300s server-side.
const deadline = Date.now() + 360 * 1000;
while (!["completed", "failed", "cancelled"].includes(task.status)) {
  if (Date.now() > deadline) throw new Error(`still ${task.status} — keep the task_uuid and poll later`);
  await new Promise((r) => setTimeout(r, 3000));
  const poll = await fetch(`https://www.namifusion.com/api/v1/marketplace/run/tasks/${task.task_uuid}`, { headers: HEADERS });
  if (!poll.ok) throw new Error(`poll failed: ${poll.status}`);
  task = await poll.json();
}

console.log(task.status, task.output);

เอกสาร API

Seedance 2.0 Mini Text to Video API

A faster, lower-cost cinematic text-to-video model with native audio and up to 4K output.

Seedance 2.0 Mini Text to Video is ByteDance’s efficient text-to-video model for generating short, cinematic multi-shot sequences from natural language. It combines strong prompt adherence, AI camera control, cross-scene character consistency, flexible output formats, and native audio generation, while delivering no cold start performance and cost-effective scaling for production workloads.

🚀 Key Features

  • Prompt-based video generation: Generate videos directly from natural-language prompts describing scene, action, camera movement, and mood.
  • Reference-guided control: Use reference images, videos, and audio to guide visual style, character consistency, composition, motion, or sound direction.
  • Cinematic multi-shot output: Optimized for narrative short-form video generation with stronger scene continuity and camera-aware storytelling.
  • Native audio generation: Generate synchronized audio together with the output video to streamline post-production.
  • Flexible output options: Supports multiple aspect ratios, 4–15 second durations, and 480p to 4K resolution tiers.
  • Fast and production-friendly: Designed for lower cost, strong performance, and no cold starts, making it suitable for iterative creation and scaled deployment.

🛠️ Technical Specifications

ItemSpecification
Model NameSeedance 2.0 Mini Text to Video
Model IDbytedance/seedance-2.0-mini/text-to-video
Model TypeText-to-video
Model ArchitectureMultimodal conditional video generation model using text prompts
Input FormatText prompt
Output FormatVideo(optionally with synchronized native audio)
Prompt GuidanceBest results come from describing scene, subject, action, camera movement, lighting, and mood
Aspect Ratio16:9, 9:16, 4:3, 3:4, 1:1, 21:9
Default Aspect Ratio16:9
Resolution480p, 720p, 1080p, 4k
Default Resolution720p
Duration4–15 seconds
Default Duration5 seconds
Audio GenerationSupported, synchronized native audio enabled by default
Real-time Information OptionOptional web search support for prompts requiring current information
Performance ProfileFaster inference, lower cost, no cold start

Sample Prompts

  1. A cinematic night market scene, warm lantern light, people walking through a narrow street, soft camera tracking movement, steam rising from food stalls, realistic atmosphere, gentle ambient motion, rich colors, detailed urban background.
  2. A lone traveler standing on a seaside cliff at dawn, wind moving the coat, camera slowly pushing in from behind to a side close-up, pale sky, restrained but epic mood.
  3. Inside a futuristic laboratory, a humanoid robot powers on and slowly raises its head, cool lighting reflecting on metallic surfaces, low-angle orbit camera, quiet and tense atmosphere.

💰 Pricing

Service TypeResolutionPrice
Video Gen (T2V/I2V)all$3.5 / 1M Tokens
Video Gen (V2V)all$2.1 / 1M Tokens

Estimate Costs by Duration

ResolutionPrice per 5 secondsPrice per second
480p$0.168$0.0336
720p$0.378$0.0756
1080p$1.50$0.30
4k$3.00$0.60

💡 Best Use Cases

  • Social media content production: Create vertical, square, widescreen, or ultrawide short-form videos for platform-specific campaigns.
  • Creative concepting and previs: Rapidly prototype scenes, camera moves, and mood for advertising, branded content, and film development.
  • Character and style consistency workflows: Use reference images to maintain visual identity across shots and scenes.
  • Audio-visual narrative clips: Generate short cinematic sequences with synchronized sound for immersive storytelling and promotional content.

🔗 Related Models

  • ByteDance Seedance 2.0 Mini Image-to-Video: Better suited for workflows that begin from a starting image plus prompt.

โมเดลที่เกี่ยวข้อง

xAI Grok Imagine Video v1.5 Text to Video
ข้อความเป็นวิดีโอX Ai

xAI Grok Imagine Video v1.5 Text to Video

x-ai/grok-imagine-video-v1.5/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $0.840 / ต่อการรัน
MiniMax H3 Text to Video
ข้อความเป็นวิดีโอMinimax

MiniMax H3 Text to Video

minimax/h3/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $0.650 / ต่อการรัน
ltx-2.5/text-to-video
ข้อความเป็นวิดีโอLtx 2.5

ltx-2.5/text-to-video

ltx-2.5/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $0.500 / ต่อการรัน
Kling Omni Video O3 Standard Text-To-Video
ข้อความเป็นวิดีโอKling

Kling Omni Video O3 Standard Text-To-Video

kwaivgi/kling-video-o3/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $0.420 / ต่อการรัน
Kling 3.0
ข้อความเป็นวิดีโอKling

Kling 3.0

kwaivgi/kling-v3.0/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $0.420 / ต่อการรัน
Kling V2.6 Text to Video API
ข้อความเป็นวิดีโอKling

Kling V2.6 Text to Video API

kwaivgi/kling-v2.6/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $0.210 / ต่อการรัน
Gemini Omni Flash Text to Video API
ข้อความเป็นวิดีโอGoogle

Gemini Omni Flash Text to Video API

google/gemini-omni-flash/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $1.04 / ต่อการรัน
Seedance 2.5 Text-to-Video
ข้อความเป็นวิดีโอDoubao

Seedance 2.5 Text-to-Video

doubao/seedance-2-5/text-to-video. เลือกดู API โมเดลภาพ วิดีโอ และสลับใบหน้า ทดลองโมเดลแล้วเชื่อมต่อผ่าน REST API แบบรวมศูนย์

เริ่มต้น $1.155 / ต่อการรัน