How to Use the xAI Grok Imagine Image v2.0 Text to Image API: text to image integration guide (2026)

2 min · Updated 2026-07-11 · By the NamiFusion team

Learn how to call the xAI Grok Imagine Image v2.0 Text to Image API on NamiFusion: parameters, pricing (from ≈$0.080) and runnable copy-paste code — try it in the playground first.

xAI Grok Imagine Image v2.0 Text to Image is a text to image model by X Ai: xAI Grok Imagine Image V2.0 Text-to-Image generates high-quality images from text prompts, with configurable aspect ratio, resolution, and quality for creative visuals, social content, marketing assets, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing. On NamiFusion you can try it in the playground (exact quote shown before you submit), then call it through one unified REST API — a typical run costs about $0.080.

TL;DR

· Task type: text to image, provided by X Ai

· Pricing: about $0.080 per typical run ($1 = 100 credits), pay-as-you-go

· Getting started: playground (exact quote before submit) → create an API key → one POST request

· Tunable parameters: 5 (see the table below)

· No platform watermark; commercial use allowed (per the Terms of Service)

What can you build with xAI Grok Imagine Image v2.0 Text to Image?

Social media visuals and covers: batch-generate on-theme images and schedule them straight away.

E-commerce and ad creatives: product scenes, banner backgrounds, and multi-variant visuals for A/B tests.

Concept art and ideation: characters, environments and style boards, iterated cheaply and fast.

Pipeline chaining: outputs feed directly into image-to-video or face swap as source material.

How do you call this API?

1. Sign up on NamiFusion and create a key on the API Keys page (signup bonus credits can only be used on selected models).

2. Submit a task with the cURL example below, passing your key as a Bearer token.

3. Poll the task status with the returned task_uuid and read the output URLs when completed.

Submit a task
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/x-ai/grok-imagine-image-v2.0/text-to-image" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "input": {
    "prompt": "A cinematic futuristic city skyline at sunset, viewed from a rooftop garden, with glowing neon signs, flying vehicles, reflective glass towers, and soft golden light breaking through dramatic clouds; ultra-detailed, photorealistic, vibrant colors, sharp focus",
    "aspect_ratio": "16:9",
    "resolution": "2k",
    "quality": "medium",
    "n": 3
  }
}'

# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
  -H "Authorization: Bearer YOUR_API_KEY"

What parameters does it take?

prompt (Prompt): required

aspect_ratio (Aspect Ratio): optional, default "1:1"

resolution (Resolution): optional, default "2k"

quality (Quality): optional, default "medium"

n (Image Count): optional, default 1

How do you write prompts for xAI Grok Imagine Image v2.0 Text to Image?

1. Subject first: lead with the actor and action, then add environment, lighting and style ("an astronaut sprinting through a neon street in the rain, cinematic, shallow depth of field").

2. Concrete nouns beat stacked adjectives: "85mm portrait lens, golden-hour backlight" works far better than "very beautiful, ultra high quality".

3. Change one variable per iteration: pin the seed to reproduce a result, tweak a single phrase at a time so you know what actually moved the output.

4. State negatives positively: instead of "no blur", ask for "sharp focus, crisp details".

How much does it cost?

Pay-per-use in credits. A typical run costs about $0.080 ($1 = 100 credits); the exact price varies with parameters like resolution or duration, and the playground shows the exact quote before you submit.

xAI Grok Imagine Image v2.0 Text to Image

High-quality text-to-image generation with no coldstarts, built for fast iteration and production workflows.

xAI Grok Imagine Image V2.0 Text-to-Image generates high-quality visuals from natural language prompts. Configurable aspect ratio, resolution, and quality make it suitable for everything from creative exploration to production delivery, with stable latency and predictable cost.

🚀 Key Features

  • High-quality generation: Sharp, detailed images from natural language prompts for creative visuals, marketing assets, and concept design.
  • Flexible aspect ratios: 13 common and mobile-friendly ratios for social media, commerce, and content publishing.
  • 1k / 2k resolution: Iterate quickly at 1k, deliver finished assets at 2k.
  • Adjustable quality tier: low and medium let you trade speed and cost against visual fidelity.
  • Batch generation: Up to 10 images per request to speed up exploration and selection.
  • No coldstarts: Suited to high-frequency calls and stable production environments.

🛠️ Specifications

ItemDetail
ArchitectureText-to-image generation model
Task typeText to image
InputText prompt
OutputImage
Images per request1–10 (n, default 1)
Aspect ratios1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 2:1, 1:2, 19.5:9, 9:19.5, 20:9, 9:20
Default aspect ratio1:1
Resolution1k, 2k (default: 2k)
Qualitylow, medium (default: medium)
LatencyNot published; stable inference with no coldstarts
WorkflowsCreative generation, visual ideation, marketing assets, production image output

Example prompts

  • A cinematic sunrise over the ocean, towering waves, golden morning light, highly detailed, photorealistic
  • Minimalist premium skincare product shot, clean background, soft studio lighting, commercial advertising quality
  • Futuristic city nightscape concept art, neon reflections, rain-soaked streets, wide-angle lens, cyberpunk style

💰 Pricing

QualityResolutionPrice per image
low1k$0.04
low2k$0.06
medium1k$0.06
medium2k (default)$0.08

Billing notes

  • Priced by quality tier x resolution; all four combinations differ.
  • Total = price per image x n (number of images).
  • Aspect ratio does not affect price.
  • When quality and resolution are not specified, billing uses the defaults medium + 2k.

FAQ

Can I try it for free first?+

Signup bonus credits can only be used on selected models, and eligibility may change. The playground shows the exact quote before submission.

Can I use the output commercially? Is there a watermark?+

Yes, outputs are yours to use commercially, and NamiFusion adds no platform watermark — subject to the Terms of Service and applicable law.

What is the content policy?+

NamiFusion is an unfiltered engine: no arbitrary restrictions and NSFW-friendly, but illegal content and non-consensual use of real people are prohibited — see the Terms of Service.

How long does a xAI Grok Imagine Image v2.0 Text to Image task take?+

It depends on parameters — images usually take seconds, videos tens of seconds to minutes. The API returns a task_uuid for polling.

What are the resolution and duration limits?+

See the parameter table above — each model lists its available resolution and duration options there and in the playground.

How to Use the xAI Grok Imagine Image v2.0 Text to Image API: text to image integration guide (2026) | NamiFusion Blog