Fast bilingual image creation engine with depth and pose guidance for precise, photoreal visual design.
Nano Banana 2 (Gemini 3.1 Flash Image) is Google DeepMind’s Flash-tier image generation model designed for high-speed, instruction-following visual creation with strong typography rendering and real-world knowledge integration.
Nano Banana 2 text to image converts a single text prompt into a still image per request, with flexible resolution tiers from 0.5K to 4K. It is optimized for fast iteration, predictable framing, and production-ready outputs suitable for marketing visuals, product mockups, social media assets, and storyboards.
Outputs: still images (1 per request).
The following controls are exposed for Nano Banana 2 Text-to-Image.
| Parameter | Required | Type | Default | Range / Options | Description |
|---|---|---|---|---|---|
| prompt* | Yes (*) | string (str) | A cinematic close-up portrait of an American woman standing under neon lights in rainy Tokyo... | — | Text prompt describing subject, scene, lighting, style, and composition. Be clear and structured for best results. |
| aspect_ratio | No | string | auto | auto, 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16 | Output framing. "auto" preserves natural composition; select a ratio for specific layout needs. |
| resolution | No | string | 1K | 0.5K, 1K, 2K, 4K | Target resolution. Higher tiers increase detail and cost. |
Use lower resolution during ideation, then upscale to higher resolution once composition is finalized.
1) Write a structured prompt: Subject → action → environment → style → camera/lighting.
2) Set aspect_ratio to match your final deliverable (or keep "auto" for natural framing).
3) Choose resolution based on draft vs. production needs.
4) Generate and review composition, lighting, and typography.
5) Iterate by adjusting small variables (pose, color, mood) rather than rewriting the entire prompt.
Fast bilingual image creation engine with depth and pose guidance for precise, photoreal visual design.
Remix an image with a prompt while keeping the original style in Ideogram 3.
Precision visual editing tool for consistent, photorealistic brand assets
Fast, precise, iterative AI image editing model.
Create photorealistic, text-accurate visuals with precise prompt control.
Create photoreal visuals with multi-reference, color, and typography precision.
Nano Banana 2 text-to-image is designed for fast iteration and consistent prompt-following. It’s a great fit for rapid concept exploration, marketing drafts, thumbnails, and generating multiple variations quickly.
Nano Banana 2 text-to-image supports common aspect ratios including 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, and 9:16. Choose the ratio that matches your target layout (banner, square post, story, etc.).
If enabled, enhance_prompt tries to expand or refine your prompt to improve descriptiveness and coherence. If you prefer precise control, keep it off and write explicit constraints (subject, style, lighting, composition) yourself.
RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.









