Image to Video AI Models

All 15 Image to Video AI models on NamiFusion. Compare capabilities and price, try any model in the playground, then call it through one unified API — unfiltered, pay-as-you-go, no watermark.

15 models

xAI Grok Imagine Video v1.5 Image to Video
Image to VideoX Ai

xAI Grok Imagine Video v1.5 Image to Video

Animate one input image with a text prompt into a 1-15 second video at 480p or 720p.

from $0.840 / per run
Vidu Q3 Image To Video
Image to VideoVidu

Vidu Q3 Image To Video

Vidu Q3 Image-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

from $0.750 / per run
Nami Wan 2.7 I2V Spicy Prime
Image to Video

Nami Wan 2.7 I2V Spicy Prime

Create a 2–15 second video from a reference image and prompt, with 720p/1080p output and optional audio guidance.

from $1.00 / per run
MiniMax H3 Image to Video
Image to VideoMinimax

MiniMax H3 Image to Video

MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

from $0.650 / per run
Kling Omni Video O3 Image-To-Video
Image to VideoKling

Kling Omni Video O3 Image-To-Video

Kling Omni Video O3 Image-to-Video transforms static images into dynamic cinematic videos using MVL (Multi-modal Visual Language) technology. Maintains subject consistency while adding natural motion, physics simulation, and seamless scene dynamics. Supports audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

from $0.420 / per run
Kling Omni Video O1 Image-to-Video
Image to VideoKling

Kling Omni Video O1 Image-to-Video

Kling Omni Video O1 Image-to-Video transforms static images into dynamic cinematic videos using MVL (Multi-modal Visual Language) technology. Maintains subject consistency while adding natural motion, physics simulation, and seamless scene dynamics. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

from $0.560 / per run
Kling 3.0 Standard
Image to VideoKling

Kling 3.0 Standard

Kling 3.0 Standard delivers high-quality image-to-video generation with smooth motion, cinematic visuals, accurate prompt adherence, and native audio for ready-to-share clips. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

from $0.420 / per run
Kling V2.6 Image to Video API
Image to VideoKling

Kling V2.6 Image to Video API

Kling 2.6 delivers top-tier image-to-video generation with smooth motion, cinematic visuals, accurate prompt adherence, and native audio for ready-to-share clips. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

from $0.210 / per run
Image to Video
Image to VideoGoogle

Gemini Omni Flash Image to Video API

Gemini Omni Flash Image to Video animates input images into short AI videos with synchronized audio, adding motion and sound while following the source image. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

from $1.12 / per run
Image to Video
Image to VideoDoubao

Seedance 2.0 Mini Image to Video

Seedance 2.0 Mini Image to Video is ByteDance's faster, lower-cost image-to-video model for cinematic multi-shot videos. It turns reference images and optional text prompts into narrative sequences with AI camera control, consistent characters across scenes, 480P / 720P / 1080P / 4K output, 4-15s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

from $0.380 / per run
Seedance 2.5 Image to Video
Image to VideoDoubao

Seedance 2.5 Image to Video

Bring static moments to life. Seedance 2.5 features exceptional Image-to-Video capabilities, deeply analyzing spatial structures and textures to transform still images into nuanced, natural motion, making every photo the opening scene of a movie.

from $1.16 / per run
seedance-2-0 image-to-video
Image to VideoDoubao

seedance-2-0 image-to-video

Bring static moments to life. Seedance 2.0 features exceptional Image-to-Video capabilities, deeply analyzing spatial structures and textures to transform still images into nuanced, natural motion, making every photo the opening scene of a movie.

from $0.760 / per run
Seedance 2.0 Fast Image to Video
Image to VideoDoubao

Seedance 2.0 Fast Image to Video

Seedance 2.0 Fast (Image-to-Video) generates cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level control, and exceptional motion stability — optimized for faster generation at lower cost. Built on Seed's unified multimodal architecture.

from $0.610 / per run
Alibaba WAN 2.6
Image to VideoAlibaba

Alibaba WAN 2.6

Alibaba WAN 2.6 converts text or images into videos (720p/1080p) with synced audio, faster and more affordable than Google Veo3. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

from $0.750 / per run
Alibaba WAN 2.5
Image to VideoAlibaba

Alibaba WAN 2.5

Alibaba WAN 2.5 redefines AI video generation by transforming text or images into high-quality videos (480p/720p/1080p) with natively synced audio. Engineered to be faster and more cost-effective than Google Veo 3, it offers a ready-to-use REST API with industry-leading performance and zero cold starts. It is the ultimate solution for creators seeking professional-grade output at a fraction of the cost.

from $0.750 / per run