How to Use the kwaivgi/kling-v2-ai-avatar-pro API: avatar to video integration guide (2026)

2 min · Updated 2026-07-11 · By the NamiFusion team

Learn how to call the kwaivgi/kling-v2-ai-avatar-pro API on NamiFusion: parameters, pricing (from ≈$0.570) and runnable copy-paste code — try it in the playground first.

kwaivgi/kling-v2-ai-avatar-pro is an avatar to video model by Kling: Kling V2 AI Avatar Pro generates high-quality AI avatar videos with clean detail, stable motion, and strong identity consistency—ideal for profiles, intros, and social content. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — a typical run costs about $0.570.

TL;DR

· Task type: avatar to video, provided by Kling

· Pricing: about $0.570 per typical run ($1 = 100 credits), pay-as-you-go

· Getting started: playground (exact quote before submit) → create an API key → one POST request

· Tunable parameters: 2 (see the table below)

· No platform watermark; commercial use allowed (per the Terms of Service)

What can you build with kwaivgi/kling-v2-ai-avatar-pro?

Short-form video and ad clips: go from a prompt or a single image to publishable motion content.

Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.

A social content factory: batch vertical clips and ride trending topics fast.

Remixing existing assets: image-to-video brings static work to life and extends its lifespan.

How do you call this API?

1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).

2. Submit a task with the cURL example below, passing your key as a Bearer token.

3. Poll the task status with the returned task_uuid and read the output URLs when completed.

Submit a task
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/kwaivgi/kling-v2-ai-avatar-pro" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "image": "https://example.com/input.jpg"
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"

What parameters does it take?

image (Image): required

prompt (Prompt): optional

How do you write prompts for kwaivgi/kling-v2-ai-avatar-pro?

1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").

2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".

3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.

4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".

How much does it cost?

Pay-per-use in credits. A typical run costs about $0.570 ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.

kwaivgi/kling-v2-ai-avatar-pro

Create high-quality talking-head avatar videos from a single portrait and your own audio.

kwaivgi/kling-v2-ai-avatar-pro is an image-to-video model designed for AI avatar generation. It turns a single portrait image and an audio track into a clean, stable, lip-synced talking-head video with strong identity consistency. With no cold start, strong output quality, and affordable usage-based pricing, it is well suited for social content, profile videos, intros, and virtual presenter workflows.

🚀 Key Features

  • Audio-driven lip sync: Uses uploaded audio directly rather than synthetic speech, preserving timing, pauses, and emotional delivery.
  • Strong identity consistency: Maintains facial characteristics from the reference image while animating the face, eyes, and head naturally.
  • Stable talking-head motion: Optimized for clean, on-camera avatar performance with controlled movement and reliable visual coherence.
  • One-shot workflow: Requires only one Image and one Audio input, eliminating the need for video capture or motion recording.
  • Prompt-guided styling: Supports optional Prompt input to influence mood, expression, lighting feel, or subtle camera presence.
  • Social-ready vertical output: Ideal for short-form vertical content formats used across TikTok, Reels, Shorts, and Stories.

🛠️ Technical Specifications

ItemDetails
Model Namekwaivgi/kling-v2-ai-avatar-pro
Model TypeImage-to-video(AI avatar / talking-head generation)
Core FunctionGenerates a lip-synced avatar Video from a single portrait Image and an Audio track
InputsImage, Audio, optional Prompt
Output FormatVideo
Image RequirementsClear portrait, preferably front-facing or 3/4 view, visible eyes, minimal occlusion
Audio RequirementsClean mono or stereo speech audio with limited background noise
DurationAutomatically derived from the input audio length; minimum billed duration is 5 seconds
ResolutionHD vertical output(exact pixel dimensions not specified in source material)
Aspect RatioVertical / portrait-oriented social format
Motion DriverAudio-driven lip sync, timing, and facial performance
Style ControlOptional Prompt for expression, mood, lighting, and subtle motion cues
LatencyNo fixed public value provided; platform highlights no cold start
Output CharacteristicsClean detail, stable motion, strong identity preservation

Sample Prompts

  • soft studio lighting, subtle head movement, gentle smile
  • confident presenter in a tech promo, subtle head nods
  • friendly customer service tone, warm expression, clean professional framing

💰 Pricing

Audio Length(s)Billed SecondsPrice(USD)
0–550.56
10101.12
20202.24
30303.36
60606.72

Clips shorter than 5 seconds are still billed as 5 seconds.

FAQ

Can I try it for free first?+

Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.

Can I use the output commercially? Is there a watermark?+

Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.

What is the content policy?+

NamiFusion is an unfiltered engine: no arbitrary restrictions and NSFW-friendly, but illegal content and non-consensual use of real people are prohibited — see the Terms of Service.

How long does a kwaivgi/kling-v2-ai-avatar-pro task take?+

It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.

What are the resolution and duration limits?+

See the parameter table above — each model lists its available resolution and duration options there and in the playground.

How to Use the kwaivgi/kling-v2-ai-avatar-pro API: avatar to video integration guide (2026) | NamiFusion Blog