
Nano Banana 2 Text-to-Image
Nano Banana 2 Text-to-Image (Gemini 3.1 Flash Image) is Google’s next-generation AI model for image editing and generation, making visual creation as simple and intuitive as describing it in words. Built on Google’s cutting-edge computer vision and generative AI technologies, it combines precise control, creative flexibility, and deep semantic understanding to deliver professional-grade image editing and generation.
Read Me
Google Nano Banana 2 Text-to-Image
Introduction
Nano Banana 2 Text-to-Image is Google’s prompt-only endpoint for high-frequency visual production. It is designed for creating marketing concepts, campaign key art, content covers, and multi-channel images directly from written requirements. Because the page does not include reference generation or image editing, it is best suited to users who can describe the intended result clearly in text.
On iCreat, users can enter prompts of up to 8,192 characters, choose from ten common aspect ratios, and generate at 1K, 2K, or 4K resolution. The endpoint uses fixed resolution-based pricing and is available through the Playground or the google/gemini-3-1-flash-image/text-to-image API.
Key Features
Prompt-only image generation
This endpoint accepts written prompts without requiring users to upload or organize reference images. Creative goals, subject relationships, composition, exact text, and publishing requirements can be combined in one brief, making it practical for rapid concept development, batch proposals, and visual production that does not depend on existing source assets.
Build complete compositions from written briefs
Nano Banana 2 can combine subjects, environments, camera direction, lighting, style, exact copy, and intended use inside one generated composition. When the prompt defines clear priorities, the model has more context for arranging subject placement, text areas, negative space, and supporting elements around the primary communication goal.
Plan layouts around the aspect ratio
The endpoint provides ten common aspect ratios, allowing users to choose the final placement before generation. Separate instructions can be written for vertical ads, square posts, landscape banners, or ultrawide assets, reducing the composition loss that often occurs when one finished image is forced into several different crops afterward.
Generate from 1K to 4K
Users can select 1K, 2K, or 4K output for early concepts, web publishing, presentations, or final delivery. The same prompt and ratio controls apply across the three tiers, so teams can validate a direction at a smaller size before generating a larger final version for production.
Fixed pricing for text-to-image work
The text-to-image endpoint is priced only by resolution: $0.07 for 1K, $0.105 for 2K, and $0.14 for 4K, with no additional quality tier. For teams that only need prompt-based generation, this structure is easier to budget than a complete endpoint that also includes reference inputs and editing workflows.
How to Use
1. Enter the prompt
This endpoint is dedicated to text-to-image generation. Enter a written brief to create a new image from scratch without uploading references or selecting an editing mode.
2. Define the composition
Describe the subject, environment, camera, lighting, style, exact copy, and intended use, then identify the most important content and any space that must remain open.
3. Select ratio and resolution
Choose the aspect ratio for the final placement, then select 1K, 2K, or 4K according to drafting, web publishing, presentation, cropping, or delivery requirements.
4. Submit the generation task
Generate directly in the Playground or call google/gemini-3-1-flash-image/text-to-image, poll the task status, and retrieve the completed image when processing succeeds.
Pricing
Nano Banana 2 Text-to-Image uses fixed per-image pricing based on output resolution, with no additional quality tier.
| Resolution | Unit Price |
|---|---|
| 1K | $0.07/image |
| 2K | $0.105/image |
| 4K | $0.14/image |
The 4K tier costs only $0.035 more than 2K, making it practical after the composition is approved and the final asset needs more cropping room or a larger delivery size.
Model Comparison
Nano Banana 2 Text-to-Image vs Gemini 3 Pro Image Text-to-Image
Nano Banana 2 targets high-frequency text-to-image production, while Gemini 3 Pro Image is positioned for complex professional visuals and finer creative control.
| What matters | Nano Banana 2 Text-to-Image | Gemini 3 Pro Image Text-to-Image |
|---|---|---|
| Best for | High-frequency text-to-image | Professional text-to-image |
| Production priority | Speed and scale | Precision and control |
| Text and layout | Reliable text generation | Professional text + complex design |
| Input method | Prompt only | Prompt only |
| Editing workflow | Not supported | Not supported |
| Reference limit | Not supported | Not supported |
| Prompt limit | 8,192 characters | 8,192 characters |
| Output control | Ratio + resolution | Ratio + resolution |
| Quality tiers | None | None |
| 1K price | $0.07 | $0.07 |
| 2K price | $0.105 | $0.105 |
| 4K price | $0.14 | $0.14 |
| Budget planning | Fixed by resolution | Fixed by resolution |
Nano Banana 2 Text-to-Image vs Seedream 5.0 Lite
Nano Banana 2 adds more aspect ratios and a 1K tier, while Seedream 5.0 Lite provides lower 2K and 4K pricing plus reference-based editing.
| What matters | Nano Banana 2 Text-to-Image | Seedream 5.0 Lite |
|---|---|---|
| Best for | High-frequency text-to-image | Low-cost image production |
| Input method | Prompt only | Prompt + references |
| Editing workflow | Not supported | Prompt + references |
| Reference limit | Not supported | Up to 5 |
| Prompt limit | 8,192 characters | 10,000 characters |
| Output control | Ratio + resolution | Resolution + format |
| Quality tiers | None | None |
| 1K price | $0.07 | Not supported |
| 2K price | $0.105 | $0.035 |
| 4K price | $0.14 | $0.035 |
| Budget planning | Fixed by resolution | Same price at 2K / 4K |
Parameters
| Parameter | Type | Required | Description |
|---|---|---|---|
prompt |
string | Yes | Image prompt with a maximum length of 8,192 characters |
aspect_ratio |
string | Yes | Selects 1:1, 2:3, 3:2, 3:4, 4:3, 4:5, 5:4, 9:16, 16:9, or 21:9 |
image_size |
string | Yes | Selects 1K, 2K, or 4K output |
| Specification | Value |
|---|---|
| Model ID | google/gemini-3-1-flash-image/text-to-image |
| Workflows | Text-to-image |
| Prompt limit | 8,192 characters |
| Resolution tiers | 1K, 2K, 4K |
| Aspect ratios | 1:1, 2:3, 3:2, 3:4, 4:3, 4:5, 5:4, 9:16, 16:9, 21:9 |
| Quality tiers | None |
| Billing unit | USD per image |
| API pattern | Asynchronous task submission |
Use Cases
Written-brief concept exploration
Turn campaign ideas, visual directions, or content requirements into several early images for internal discussion, direction selection, and proposal preparation without collecting reference assets first.
Headline-led campaign key art
Start from a headline, central subject, and placement requirements to create launch graphics, event covers, and promotional key art during the early creative stage.
Fictional product concepts
Describe a product’s function, materials, colors, and environment to create concept visuals for positioning discussions and early exploration before a real product asset library exists.
Presentation and editorial opening visuals
Generate presentation covers, feature headers, report openers, and editorial concepts with deliberate title areas, information space, visual focus, and an appropriate narrative mood.
Information graphics from written requirements
Describe steps, categories, labels, and hierarchy in the prompt to create simple infographic drafts, process visuals, and annotated explainers for concept validation.
Multi-channel composition planning
Generate separate square, vertical, landscape, and ultrawide compositions to test how the subject, headline, and negative space adapt across different publishing placements.
Prompt Tips
Define the visual hierarchy first. State the priority of the headline, primary subject, supporting elements, and negative space before adding style details.
Write exact image text. Put required copy in quotation marks and specify the language, position, size, alignment, and surrounding space.
Describe composition for the ratio. Select the final aspect ratio first, then define subject placement, text areas, and edge spacing instead of relying on later cropping.
State the image’s purpose. Tell the model whether the result is campaign key art, a presentation cover, an advertisement, or social content to guide information density and layout.
Notes
aspect_ratiois required in the current schema, but its description mentions inheriting the first uploaded image’s ratio; this text-to-image endpoint has no image input, so always pass the ratio explicitly.- The schema and Playground use
1K,2K, and4K, while the request example uses lowercase1k; confirm whether the iCreat endpoint normalizes casing before production integration. - Google’s native 0.5K option, extreme aspect ratios, Image Search grounding, and Thinking controls are not exposed in the current iCreat parameters and should not be presented as available endpoint features.
- The current endpoint has no visible watermark setting, but images generated by the Nano Banana family include an invisible SynthID watermark.
Related Models
Gemini 3 Pro Image Text-to-Image
Choose it for more demanding professional visuals, finer creative control, and higher-complexity text-to-image work.
Choose it when you need reference-based generation, image editing, and up to 11 reference images in the complete Nano Banana 2 workflow.
Choose it when lower unit pricing, equal 2K and 4K costs, reference inputs, and output-format control matter more.
