How to Use the kwaivgi/kling-v2-ai-avatar-standard API: avatar to video integration guide (2026)
Learn how to call the kwaivgi/kling-v2-ai-avatar-standard API on NamiFusion: parameters, pricing (from ≈$0.290) and runnable copy-paste code — try it in the playground first.
kwaivgi/kling-v2-ai-avatar-standard is an avatar to video model by Kling: Kling AI Avatar generates high-quality AI avatar videos for profiles, intros, and social content, delivering clean detail and cinematic motion with reliable prompt adherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — a typical run costs about $0.290.
TL;DR
· Task type: avatar to video, provided by Kling
· Pricing: about $0.290 per typical run ($1 = 100 credits), pay-as-you-go
· Getting started: playground (exact quote before submit) → create an API key → one POST request
· Tunable parameters: 2 (see the table below)
· No platform watermark; commercial use allowed (per the Terms of Service)
What can you build with kwaivgi/kling-v2-ai-avatar-standard?
Short-form video and ad clips: go from a prompt or a single image to publishable motion content.
Product demos and concept previews: validate storyboards and visual direction cheaply before any shoot.
A social content factory: batch vertical clips and ride trending topics fast.
Remixing existing assets: image-to-video brings static work to life and extends its lifespan.
How do you call this API?
1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).
2. Submit a task with the cURL example below, passing your key as a Bearer token.
3. Poll the task status with the returned task_uuid and read the output URLs when completed.
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/kwaivgi/kling-v2-ai-avatar-standard" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"input": {
"image": "https://example.com/input.jpg"
}
}'
# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
-H "Authorization: Bearer YOUR_API_KEY"What parameters does it take?
image (Image): required
prompt (Prompt): optional
How do you write prompts for kwaivgi/kling-v2-ai-avatar-standard?
1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").
2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".
3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.
4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".
How much does it cost?
Pay-per-use in credits. A typical run costs about $0.290 ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.
kwaivgi/kling-v2-ai-avatar-standard
Create natural talking-avatar videos from a single image and audio track, with reliable lip sync and no cold start.
kwaivgi/kling-v2-ai-avatar-standard is an image-to-video AI avatar model that turns one image plus one audio track into a realistic speaking or singing video. Built on the Kling V2 avatar stack, it combines strong prompt adherence, expressive facial animation, and affordable usage-based pricing, making it a practical choice for production-oriented avatar workflows.
🚀 Key Features
- Single-image avatar animation: Generate a talking-avatar video from just one image and one audio source, reducing asset requirements for fast content creation.
- Accurate lip synchronization: Mouth shapes and jaw motion are closely aligned to speech rhythm, pronunciation, and timing for more believable delivery.
- Expressive face and head motion: Beyond lip sync, the model animates blinks, eyebrow movement, subtle head turns, and micro-expressions to match vocal emotion.
- Strong identity preservation: Maintains facial identity, hairstyle, and overall visual style consistently across frames for stable avatar output.
- Prompt-guided performance control: Optional prompts can steer mood, energy, and behavior, such as a calm presenter or an energetic streamer.
- No cold start, cost-efficient deployment: Designed for responsive production usage with predictable startup behavior and accessible pricing.
🛠️ Technical Specifications
| Item | Details |
|---|---|
| Model architecture | Kling V2 AI Avatar Standard(image-to-video, audio-driven avatar generation) |
| Task type | image-to-video |
| Input format | Image + Audio + optional Prompt |
| Output format | Video |
| Image input | Single portrait, character image, or animal image; front-facing or slight 3/4 view recommended |
| Audio input | Single audio track; clean voice recordings or TTS work best |
| Prompt support | Yes; used to control mood, energy, and behavior |
| Core capabilities | Lip sync, facial expression animation, head motion, identity preservation |
| Supported subjects | Human portraits, stylized characters, pets/animals |
| Duration | Billed by audio length; up to 300 seconds(5 minutes)per job |
| Resolution | Not publicly specified; higher-resolution outputs generally require more render time |
| Frame rate | Not publicly specified |
| Latency | Variable; typically increases with clip length and output quality |
| Billing rules | Minimum billing of 5 seconds; billing capped at 300 seconds per job |
Sample Prompts
friendly teacher, gentle head nodsexcited host, big smiles and energetic motioncalm news anchor, steady eye contact, professional delivery
💰 Pricing
| Mode | Price |
|---|---|
| Base price(minimum 5 seconds) | $0.28 |
| 10-second audio | $0.56 |
| Minimum billed duration | 5 seconds |
| Billing cap per job | 300 seconds(5 minutes) |
Any clip shorter than 5 seconds is still billed as 5 seconds.
FAQ
Can I try it for free first?+
Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.
Can I use the output commercially? Is there a watermark?+
Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.
What is the content policy?+
NamiFusion is an unfiltered engine: no arbitrary restrictions and NSFW-friendly, but illegal content and non-consensual use of real people are prohibited — see the Terms of Service.
How long does a kwaivgi/kling-v2-ai-avatar-standard task take?+
It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.
What are the resolution and duration limits?+
See the parameter table above — each model lists its available resolution and duration options there and in the playground.