Advanced open-weight model enabling refined image transformation and consistent visual editing.
Nano Banana 2 (Gemini 3.1 Flash Image) is Google DeepMind’s Flash-tier image generation model designed for high-speed, instruction-following visual creation with strong typography rendering and real-world knowledge integration.
Nano Banana 2 text to image converts a single text prompt into a still image per request, with flexible resolution tiers from 0.5K to 4K. It is optimized for fast iteration, predictable framing, and production-ready outputs suitable for marketing visuals, product mockups, social media assets, and storyboards.
Outputs: still images (1 per request).
The following controls are exposed for Nano Banana 2 Text-to-Image.
| Parameter | Required | Type | Default | Range / Options | Description |
|---|---|---|---|---|---|
| prompt* | Yes (*) | string (str) | A cinematic close-up portrait of an American woman standing under neon lights in rainy Tokyo... | — | Text prompt describing subject, scene, lighting, style, and composition. Be clear and structured for best results. |
| aspect_ratio | No | string | auto | auto, 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16 | Output framing. "auto" preserves natural composition; select a ratio for specific layout needs. |
| resolution | No | string | 1K | 0.5K, 1K, 2K, 4K | Target resolution. Higher tiers increase detail and cost. |
Use lower resolution during ideation, then upscale to higher resolution once composition is finalized.
1) Write a structured prompt: Subject → action → environment → style → camera/lighting.
2) Set aspect_ratio to match your final deliverable (or keep "auto" for natural framing).
3) Choose resolution based on draft vs. production needs.
4) Generate and review composition, lighting, and typography.
5) Iterate by adjusting small variables (pose, color, mood) rather than rewriting the entire prompt.
Advanced open-weight model enabling refined image transformation and consistent visual editing.
Generate accurate design visuals with refined control and repeatable detail.
Create consistent visual stories with advanced image editing and multi-scene control.
Generate refined visuals with accurate lighting and text control for design work.
Precise text rendering & multilingual edits for visual pros
Create cohesive story visuals with sequenced, style-stable image generation.
Nano Banana 2 text-to-image is designed for fast iteration and consistent prompt-following. It’s a great fit for rapid concept exploration, marketing drafts, thumbnails, and generating multiple variations quickly.
Nano Banana 2 text-to-image supports common aspect ratios including 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, and 9:16. Choose the ratio that matches your target layout (banner, square post, story, etc.).
If enabled, enhance_prompt tries to expand or refine your prompt to improve descriptiveness and coherence. If you prefer precise control, keep it off and write explicit constraints (subject, style, lighting, composition) yourself.
RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.









