How to Use the alibaba/wan-2.7/text-to-image-pro API: text to image integration guide (2026)

2 min · Updated 2026-07-11 · By the NamiFusion team

Learn how to call the alibaba/wan-2.7/text-to-image-pro API on NamiFusion: parameters, pricing (from ≈$0.080) and runnable copy-paste code — try it in the playground first.

alibaba/wan-2.7/text-to-image-pro is a text to image model by Alibaba: WAN 2.7 Text-to-Image Pro generates high-quality images up to 4K from text prompts with thinking mode for enhanced image quality. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — a typical run costs about $0.080.

TL;DR

· Task type: text to image, provided by Alibaba

· Pricing: about $0.080 per typical run ($1 = 100 credits), pay-as-you-go

· Getting started: playground (exact quote before submit) → create an API key → one POST request

· Tunable parameters: 3 (see the table below)

· No platform watermark; commercial use allowed (per the Terms of Service)

What can you build with alibaba/wan-2.7/text-to-image-pro?

Social media visuals and covers: batch-generate on-theme images and schedule them straight away.

E-commerce and ad creatives: product scenes, banner backgrounds, and multi-variant visuals for A/B tests.

Concept art and ideation: characters, environments and style boards, iterated cheaply and fast.

Pipeline chaining: outputs feed directly into image-to-video or face swap as source material.

How do you call this API?

1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).

2. Submit a task with the cURL example below, passing your key as a Bearer token.

3. Poll the task status with the returned task_uuid and read the output URLs when completed.

Submit a task
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/alibaba/wan-2.7/text-to-image-pro" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "prompt": "A cinematic portrait, dramatic rim light, shallow depth of field",
    "size": "1024*1024",
    "thinking_mode": true
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"

What parameters does it take?

prompt (Prompt): required

size (Size): optional, default "1024*1024"

thinking_mode (Thinking Mode): optional, default true

How do you write prompts for alibaba/wan-2.7/text-to-image-pro?

1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").

2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".

3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.

4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".

How much does it cost?

Pay-per-use in credits. A typical run costs about $0.080 ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.

alibaba/wan-2.7/text-to-image-pro

Professional-grade 4K text-to-image generation with stronger reasoning, consistent quality, and no cold starts.

alibaba/wan-2.7/text-to-image-pro is a professional-tier text-to-image model designed for high-fidelity image generation from text prompts. It supports output up to 4K(4096×4096), includes a built-in thinking mode for improved prompt understanding, and offers flexible size control for production-ready creative workflows.

🚀 Key Features

  • Up to 4K output: Generate images at up to 4096×4096 resolution for print, large-format displays, and high-DPI use cases.
  • Thinking mode for better quality: Built-in reasoning improves prompt adherence, composition coherence, and overall visual fidelity.
  • Custom size control: Set width and height to match square, landscape, portrait, widescreen, or ultrawide formats.
  • No cold starts: Well suited for production environments that require responsive generation without startup delays.
  • Affordable professional tier: Delivers strong image quality and controllability at a competitive per-image price.

🛠️ Technical Specifications

ItemDetails
Model architectureProfessional text-to-image model
InputPrompt
OutputImage
Default Resolution1024×1024
Maximum effective output Resolution4096×4096(4K-class)
Size rangeWidth and Height each support 512–8192 pixels
Total pixel constraintMust fall between 768×768 and 4096×4096 total pixels
Aspect Ratio constraint1:8 to 8:1
Thinking modeEnabled by default; improves quality and reasoning, but increases latency
LatencyVariable; higher Resolution and thinking mode generally increase generation time
Output FormatSingle generated high-quality image

Sample Prompts

  • A modern tea shop interior, warm afternoon light, minimalist wood design, cinematic photography
  • A luxury skincare product on reflective glass, soft studio lighting, premium commercial photography, ultra detailed
  • A futuristic city skyline at sunrise, atmospheric haze, intricate architecture, highly detailed concept art

💰 Pricing

ItemPrice
Per generated image$0.075

FAQ

Can I try it for free first?+

Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.

Can I use the output commercially? Is there a watermark?+

Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.

What is the content policy?+

NamiFusion is an unfiltered engine: no arbitrary restrictions and NSFW-friendly, but illegal content and non-consensual use of real people are prohibited — see the Terms of Service.

How long does a alibaba/wan-2.7/text-to-image-pro task take?+

It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.

What are the resolution and duration limits?+

See the parameter table above — each model lists its available resolution and duration options there and in the playground.

How to Use the alibaba/wan-2.7/text-to-image-pro API: text to image integration guide (2026) | NamiFusion Blog