How to Use the z-image-spicy/text-to-image API: text to image integration guide (2026)

2 min · Updated 2026-07-11 · By the NamiFusion team

Learn how to call the z-image-spicy/text-to-image API on NamiFusion: parameters, pricing (from ≈$0.014) and runnable copy-paste code — try it in the playground first.

z-image-spicy/text-to-image is a text to image model by Z Image Spicy: Generate high-quality images from text with custom dimensions from 256 to 1536 pixels, intelligent prompt rewriting, and seed control for creative design and batch generation. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — about $0.014 per typical run.

TL;DR

· Task type: text to image, provided by Z Image Spicy

· Pricing: about $0.014 per typical run ($1 = 100 credits), pay-as-you-go

· Getting started: playground (exact quote before submit) → create an API key → one POST request

· Tunable parameters: 5 (see the table below)

· No platform watermark; commercial use allowed (per the Terms of Service)

What can you build with z-image-spicy/text-to-image?

Social media visuals and covers: batch-generate on-theme images and schedule them straight away.

E-commerce and ad creatives: product scenes, banner backgrounds, and multi-variant visuals for A/B tests.

Concept art and ideation: characters, environments and style boards, iterated cheaply and fast.

Pipeline chaining: outputs feed directly into image-to-video or face swap as source material.

How do you call this API?

  1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).
  2. Submit a task with the cURL example below, passing your key as a Bearer token.
  3. Poll the task status with the returned task_uuid and read the output URLs when completed.
Submit a task
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/z-image-spicy/text-to-image" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "prompt": "A cinematic portrait of an astronaut in a bioluminescent forest at dusk, intricate suit details, soft blue light, photorealistic.",
    "width": 1024,
    "height": 1536,
    "prompt_extend": true
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"

What parameters does it take?

prompt (Prompt): required

width (Width): optional, default 1024

height (Height): optional, default 1536

prompt_extend (Prompt Extend): optional, default true

seed (Seed): optional

How do you write prompts for z-image-spicy/text-to-image?

  1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").
  2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".
  3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.
  4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".

How much does it cost?

Pay-per-use in credits. About $0.014 per typical run ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.

z-image-spicy/text-to-image

Create custom-size images from concise or detailed prompts, with optional automatic prompt expansion.

z-image-spicy/text-to-image is a streamlined text-to-image model for illustrations, portraits, concept art, product visuals, and social creative. It accepts a natural-language prompt and exposes direct width and height controls from 256 to 1536 pixels in 8-pixel steps. Prompt expansion can enrich a short idea before generation, while an optional seed supports repeatable exploration.

Key Features

  • Direct text-to-image generation: Converts a written visual description into a finished image.
  • Custom dimensions: Set width and height independently instead of being limited to a short list of aspect ratios.
  • Broad 256-1536 pixel range: Covers square, portrait, and landscape layouts for common digital uses.
  • 8-pixel increments: Fine-grained dimensions help match downstream composition requirements.
  • Optional prompt expansion: Adds useful visual detail to brief prompts before inference.
  • Seed control: Reuse a seed to compare prompt or dimension changes more consistently.

Technical Specifications

ItemDetails
TaskText-to-image generation
Required inputprompt
Default dimensions1024 x 1536 pixels
Width range256-1536 pixels, step 8
Height range256-1536 pixels, step 8
Prompt expansionSupported and enabled by default
Optional inputseed
OutputGenerated image URL

Sample Prompts

  • A cinematic portrait of an astronaut in a bioluminescent forest at dusk, intricate suit details, soft blue rim light, photorealistic.
  • A minimalist perfume bottle on black reflective glass, controlled highlights, subtle mist, premium editorial product photography.
  • An isometric coastal eco-city with terraced gardens, electric transit, warm morning light, clean architectural illustration.

Pricing

Prompt expansionPrice per image
Disabled1.3 credits
Enabled1.4 credits

These values reproduce the existing Marketplace pricing configuration; this documentation update does not alter pricing.

FAQ

Can I try it for free first?+

Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.

Can I use the output commercially? Is there a watermark?+

Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.

What is the content policy?+

Pornographic content, sexualized content involving a minor, unauthorized use of real people, deceptive impersonation, fraud and false endorsements are prohibited. Lawful, consensual, non-explicit adult themes may be permitted; see the Terms of Service.

How long does a z-image-spicy/text-to-image task take?+

It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.

What are the resolution and duration limits?+

See the parameter table above — each model lists its available resolution and duration options there and in the playground.