How to Use the Seedance 2.5 Reference to Video API: reference to video integration guide (2026)

2 min · Updated 2026-07-11 · By the NamiFusion team

Learn how to call the Seedance 2.5 Reference to Video API on NamiFusion: parameters, pricing (from ≈$1.16) and runnable copy-paste code — try it in the playground first.

Seedance 2.5 Reference to Video is a reference to video model by Doubao: The "Reference-to-Video" feature of Seedance 2.5 is the ultimate solution for visual stylistic unity. It precisely extracts artistic styles, lighting tones, or compositional intents from reference materials and seamlessly integrates them into newly generated videos, ensuring a highly consistent visual language for your creative series. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — a typical run costs about $1.16.

TL;DR

· Task type: reference to video, provided by Doubao

· Pricing: about $1.16 per typical run ($1 = 100 credits), pay-as-you-go

· Getting started: playground (exact quote before submit) → create an API key → one POST request

· Tunable parameters: 10 (see the table below)

· No platform watermark; commercial use allowed (per the Terms of Service)

What can you build with Seedance 2.5 Reference to Video?

Short-form video and ad clips: go from a prompt or a single image to publishable motion content.

Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.

A social content factory: batch vertical clips and ride trending topics fast.

Remixing existing assets: image-to-video brings static work to life and extends its lifespan.

How do you call this API?

1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).

2. Submit a task with the cURL example below, passing your key as a Bearer token.

3. Poll the task status with the returned task_uuid and read the output URLs when completed.

Submit a task
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/doubao/seedance-2-5/reference-to-video" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "images": [
      "https://assets-public.namifusion.com/uploads/images/2026-04-10/a9489888297a.png",
      "https://assets-public.namifusion.com/marketplace/images/2026-04-10/6abd3e6e8a4e.jpeg"
    ],
    "prompt": "The puppy in @Image2 is running on the beach in @Image1.",
    "generate_audio": true,
    "resolution": "720p",
    "ratio": "16:9",
    "duration": 7,
    "watermark": false,
    "seed": -1
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"

What parameters does it take?

images (Images): optional

audios (Audios): optional

prompt (Prompt): optional

generate_audio (Generate Audio): optional, default true

resolution (Resolution): optional, default "720p"

ratio (Aspect Ratio): optional, default "adaptive"

duration (Duration (Seconds)): optional, default 5

return_last_frame (return the last frame ): optional

watermark (Add Watermark): optional, default false

seed (Seed): optional, default -1

How do you write prompts for Seedance 2.5 Reference to Video?

1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").

2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".

3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.

4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".

How much does it cost?

Pay-per-use in credits. A typical run costs about $1.16 ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.

Seedance 2.5 Reference to Video

The "Film Studio in a Model": Precision Control with Multi-Modal Fusion.

Seedance 2.5 is ByteDance's next-generation video generation model, designed to revolutionize the "one-person film studio" workflow. Unlike traditional T2V models, it features a revolutionary Multi-modal Reference System, allowing users to blend up to 12 different inputs (images, videos, and audio) to achieve pixel-perfect control over characters, motion, and rhythm.

🚀 Key Features

  • Multi-modal Fusion: Supports simultaneous input of 30 images + 10 videos + 10 audio clips (up to 50 total) for precise character and scene consistency.
  • Cinematic Quality: High visual fidelity optimized for 480p and 720p resolutions, featuring realistic lighting and physics.
  • Audio-Visual Synchronicity: Native support for lip-sync and rhythm matching, automatically aligning visual motion with background music or speech.
  • Exceptional Physical Logic: Advanced simulation of fluid dynamics, gravity, and complex human movements with a 99.5% generation success rate.
  • Pro Camera Control: Direct language-based control over cinematic camera movements (pan, tilt, zoom) and lighting transitions.

🛠️ Technical Specifications

ParameterDetails
Model ArchitectureDiffusion-based Transformer (DiT) / Multi-modal
Max Resolution720p (Supported: 480p, 720p)
Video Duration4 - 30 Seconds (integer)
Aspect Ratios16:9, 9:16, 4:3, 3:4, 21:9, 1:1, adaptive
Input ModalityText, Image (up to 30), Video (up to 10), Audio (up to 10)
Generation Speed~30-90 seconds per clip

Sample Prompts

  • Cinematic: "Rainy city street at night, neon flickering. A man in a black trench coat walks with a red umbrella. Slow dolly-in to a facial close-up. Moody film noir aesthetic."
  • Character Consistency: "Use @Image1 as the character, reference @Video1 for the running motion, and match the tempo of @Audio1."

💰 Pricing (Volcengine API)

Service TypeResolutionPriceEstimated Cost
Video Gen (Input without video)480p/720p$10.70 / 1M Tokens~$0.103/0.231 per second
Video Gen (Input with video)480p/720p$6.40 / 1M Tokens~$0.079/$0.178 per second

FAQ

Can I try it for free first?+

Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.

Can I use the output commercially? Is there a watermark?+

Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.

What is the content policy?+

NamiFusion is an unfiltered engine: no arbitrary restrictions and NSFW-friendly, but illegal content and non-consensual use of real people are prohibited — see the Terms of Service.

How long does a Seedance 2.5 Reference to Video task take?+

It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.

What are the resolution and duration limits?+

See the parameter table above — each model lists its available resolution and duration options there and in the playground.

How to Use the Seedance 2.5 Reference to Video API: reference to video integration guide (2026) | NamiFusion Blog