Nano Banana Text to Image
google/nano-banana/text-to-image
Gemini 2.5 Flash Image. Lightweight and fast. The most affordable option for instant, high-volume generation.
Examples



Parameters
| Name | Type | Default | Constraints | Description |
|---|---|---|---|---|
| prompt *Prompt | textarea | — | ≤ 5000 chars | Describe the image you want to generate. Be specific about the subject, style, colors, and composition. |
| aspect_ratioAspect Ratio | select | 16:9 | 1:1 | 3:2 | 2:3 | 3:4 | 4:3 | 4:5 | 5:4 | 9:16 | … | Choose the aspect ratio for the generated image. |
Output fields
| Field | Type | Description |
|---|---|---|
| images | array<string> | Generated image URLs |
API
Call this model through one unified REST API. Get a key on the API Keys page.
cURL
# 1) Submit — returns { "task_uuid": "..." }
curl -X POST "https://www.namifusion.com/api/v1/marketplace/run/google/nano-banana/text-to-image" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"input": {
"prompt": "A futuristic eco-city skyline at sunset, featuring organic spiraling skyscrapers covered in lush vertical gardens and cascading waterfalls. Sleek flying vehicles weaving through the towers, warm golden hour lighting reflecting off glass surfaces. Photorealistic, architectural visualization, clean composition, 8k resolution, wide angle lens.",
"aspect_ratio": "16:9"
}
}'
# 2) Poll until status is "completed", then read the output URLs
curl "https://www.namifusion.com/api/v1/marketplace/run/tasks/TASK_UUID" \
-H "Authorization: Bearer YOUR_API_KEY"Python
import time, requests
API_KEY = "YOUR_API_KEY"
HEADERS = {"Authorization": f"Bearer {API_KEY}"}
# 1) Submit
resp = requests.post(
"https://www.namifusion.com/api/v1/marketplace/run/google/nano-banana/text-to-image",
headers=HEADERS,
json={
"input": {
"prompt": "A futuristic eco-city skyline at sunset, featuring organic spiraling skyscrapers covered in lush vertical gardens and cascading waterfalls. Sleek flying vehicles weaving through the towers, warm golden hour lighting reflecting off glass surfaces. Photorealistic, architectural visualization, clean composition, 8k resolution, wide angle lens.",
"aspect_ratio": "16:9"
}
},
)
resp.raise_for_status() # 401/402/429/5xx stop here instead of polling a bad task
task = resp.json()
# 2) Poll until a terminal state (completed / failed / cancelled).
# This model is allowed up to 300s server-side.
deadline = time.time() + 360
while task.get("status") not in ("completed", "failed", "cancelled"):
if time.time() > deadline:
raise TimeoutError(f"still {task.get('status')} — keep the task_uuid and poll later")
time.sleep(3)
poll = requests.get(f"https://www.namifusion.com/api/v1/marketplace/run/tasks/{task['task_uuid']}", headers=HEADERS)
poll.raise_for_status()
task = poll.json()
print(task["status"], task.get("output"))JavaScript
const API_KEY = "YOUR_API_KEY";
const HEADERS = { Authorization: `Bearer ${API_KEY}` };
// 1) Submit
const resp = await fetch("https://www.namifusion.com/api/v1/marketplace/run/google/nano-banana/text-to-image", {
method: "POST",
headers: { ...HEADERS, "Content-Type": "application/json" },
body: JSON.stringify({
"input": {
"prompt": "A futuristic eco-city skyline at sunset, featuring organic spiraling skyscrapers covered in lush vertical gardens and cascading waterfalls. Sleek flying vehicles weaving through the towers, warm golden hour lighting reflecting off glass surfaces. Photorealistic, architectural visualization, clean composition, 8k resolution, wide angle lens.",
"aspect_ratio": "16:9"
}
}),
});
if (!resp.ok) throw new Error(`submit failed: ${resp.status} ${await resp.text()}`);
let task = await resp.json();
// 2) Poll until a terminal state (completed / failed / cancelled).
// This model is allowed up to 300s server-side.
const deadline = Date.now() + 360 * 1000;
while (!["completed", "failed", "cancelled"].includes(task.status)) {
if (Date.now() > deadline) throw new Error(`still ${task.status} — keep the task_uuid and poll later`);
await new Promise((r) => setTimeout(r, 3000));
const poll = await fetch(`https://www.namifusion.com/api/v1/marketplace/run/tasks/${task.task_uuid}`, { headers: HEADERS });
if (!poll.ok) throw new Error(`poll failed: ${poll.status}`);
task = await poll.json();
}
console.log(task.status, task.output);Documentation
Google Nano Banana (Gemini 2.5 Flash Image)
The Essential Text-to-Image Model: Fast, Affordable, and Lightweight.
Google’s Nano Banana (Standard) is a streamlined AI image generation model built for speed and efficiency. Powered by the Gemini 2.5 Flash architecture, it is designed for creators and developers who need instant visuals without the overhead of heavy compute. It transforms simple text prompts into clear, well-composed images in seconds—with zero cold-start latency.
🚀 Key Features
- Instant Image Creation: Optimized for low latency. Generate coherent visuals from short prompts instantly.
- Cost-Effective: The most affordable choice for high-volume generation and rapid prototyping.
- Versatile Styles: Adapts naturally to various artistic needs, including realistic photos, anime, illustrations, and painterly styles.
- Accurate Understanding: robustly interprets subjects, backgrounds, and object relationships to create contextually correct compositions.
- Clean Lighting: Produces balanced, visually appealing results without overexposure or unnatural shadows.
🛠️ Technical Specifications
| Parameter | Details |
|---|---|
| Model Architecture | Gemini 2.5 Flash Image |
| Input | Text Prompts |
| Output | JPEG / PNG / WEBP |
| Latency | Ultra-Low (No cold starts) |
Sample Prompts:
"A golden retriever playing in a field of sunflowers at sunset." "A futuristic city skyline with neon reflections on wet streets." "An elegant still-life photo of coffee and croissants by a window."
💰 Pricing
| Type | Cost |
|---|---|
| Standard Generation | $0.038 / image |
💡 Best Use Cases
- Rapid Prototyping: Quickly iterate on concepts and storyboards.
- Social Media: Generate daily content for feeds and stories at scale.
- E-commerce: Produce clean background assets and product visualizations.
- Education: Visualize simple concepts for teaching materials.
🔗 Model Family
Looking for more advanced features? Check out the Nano Banana Pro series:
- Nano Banana Pro: For 4K resolution and higher fidelity.
- Nano Banana Pro Edit: For in-painting and modification.
- Nano Banana Pro Ultra: For maximum detail.
Note: Please ensure all prompts comply with Google’s Safety Guidelines. Content violating safety policies may result in generation errors.
Related models
xAI Grok Imagine Image v2.0 Text to Image
xAI Grok Imagine Image V2.0 Text-to-Image generates high-quality images from text prompts, with configurable aspect ratio, resolution, and quality for creative visuals, social content, marketing assets, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Qwen Image 3.0 Pro Text to Image
Qwen Image 3.0 Pro is a professional-grade text-to-image model with superior quality and advanced prompt understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
OpenAI GPT Image 2 Text-to-Image
OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
OpenAI GPT Image 2 Edit
OpenAI's GPT Image 2 Edit enables image editing from natural-language instructions with one or more reference images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Nami Z-Image T2I Spicy
Generate an image from a text prompt with configurable width, height, prompt enhancement, and seed.
Nano Banana Pro Text to Image
Gemini 3.0 Pro Image. The high-fidelity choice for 4K visuals, multilingual text rendering, and pro camera controls.
Nano Banana Pro Text to Image
Gemini 3.0 Pro Image. The high-fidelity choice for 4K visuals, multilingual text rendering, and pro camera controls.
Google Nano Banana Lite Text to Image API
Google Nano Banana 2 Lite Text to Image generates high-quality images from text prompts with low latency, flexible aspect ratios, and fast image creation for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.