For the complete documentation index, see llms.txt. This page is also available as Markdown.

Gemini 3 Image Generation Tool

Generate high-quality images from text prompts using Gemini 3 Pro Image

Feature Overview

The Gemini 3 Image Generation Tool is an AI image generation feature powered by MaiAgent's integration with the Google Gemini 3 Pro Image model. This tool enables AI Assistants to generate high-quality images based on users' text descriptions, supports both Traditional Chinese and English prompts, and allows style transfer and image editing through reference images.

Gemini 3 Pro Image is Google's latest multimodal AI model with powerful image generation capabilities. It can generate high-quality, contextually appropriate images based on contextual understanding and world knowledge reasoning.


Key Features

Feature Type
Description
Use Cases

Text-to-Image

Generate images based on text descriptions

Marketing materials, social media posts, product visualization

Reference Image Editing

Perform style transfer or content modification based on reference images

Brand visual consistency, image style adjustment

Conversational Image Editing

Continuously adjust and refine images across multi-turn conversations

Iterative design, detail fine-tuning

Chinese Prompt Support

Directly describe desired images in Traditional Chinese

Chinese-speaking users can use it without translation


Enable the Gemini 3 Image Generation Tool

Prerequisites

  1. Confirm your organization plan supports image generation

  2. Confirm AI feature permissions are enabled

  3. Prepare the AI Assistant that will use this tool

Enable Steps

  1. Go to the AI Assistant Settings Page

    • Click "AI Assistant" in the left menu

    • Select the AI Assistant to configure

    • Click "Settings"

  2. Configure Tool Settings

    • Switch to the "Tools" tab

    • Click the "Select Tools" button

    • Find "Nano Banana Pro Image Generation" in the available tools list and check it

    • Click "Confirm"

  3. Save Settings

    • Click the "Save" button in the bottom right corner

    • Wait for the system to confirm the settings are complete


How to Use

Use via Chat Interface

  1. Enter an Image Description

    • Type your desired image description in the chat box

    • Examples:

      • "Generate a landscape image of a sunset beach for me"

      • "Draw a cat wearing a spacesuit"

      • "Design a minimalist-style coffee shop logo"

      • "Create a watercolor painting of a mountain landscape"

  2. Use Reference Images

    • Upload an image in the conversation as a reference

    • Combine with text instructions to describe the desired modifications

    • Examples:

      • Upload a photo and say: "Convert this photo to a watercolor painting style"

      • Upload a product image and say: "Design a promotional poster for this product"

  3. Conversational Editing

    • After generating an image, request modifications in subsequent messages

    • Examples:

      • "Change the background to blue"

      • "Add some stars"

      • "Make the overall atmosphere warmer"

The more detailed the description, the better the generated image quality. It is recommended to include details such as style, color, composition, and atmosphere.


Use Case Examples

Scenario 1: Marketing Material Creation

Requirement: Quickly generate visual materials for social media posts

Conversation Example:

Scenario 2: Product Visualization

Requirement: Generate visual representations for product concepts not yet in production

Conversation Example:

Scenario 3: Style Transfer

Requirement: Convert existing images to different artistic styles

Conversation Example:


Prompt Writing Tips

Effective Prompt Writing

Include specific descriptions:

  • Subject: What content you want to generate

  • Style: Realistic, cartoon, watercolor, oil painting, graphic design, etc.

  • Color: Main color tones and color preferences

  • Composition: Close-up, panoramic, bird's-eye view, etc.

  • Atmosphere: Warm, cool, lively, calm, etc.

Good examples:

  • "An orange Shiba Inu sitting on a park bench in autumn, with golden maple leaves in the background, warm afternoon light, realistic photography style"

  • "Minimalist-style tech company logo, using blue and white, geometric design"

Avoid being too vague:

  • "Draw a picture" (lacks specific description)

  • "A nice image" (no clear direction)


FAQ

Q: What languages does the Gemini 3 Image Generation Tool support for prompts?

A: It supports Traditional Chinese and English prompts. You can directly describe the desired image content in Chinese.

Q: How many tokens does it cost to generate one image?

A: Image generation consumes more tokens than text-only conversations. The actual consumption depends on the prompt complexity and image quality settings. It is recommended to monitor your organization's token usage to avoid exceeding plan limits.

Q: Can I use reference images?

A: Yes. Upload an image in the conversation as a reference, and the AI Assistant will incorporate it into the generation process. This is especially useful for style transfer, image editing, or creating based on existing materials.

Q: What is the resolution of generated images?

A: Images are generated in 4K high quality by default, suitable for most use cases.

Q: Can I continuously modify images across multi-turn conversations?

A: Yes. The AI Assistant remembers the conversation context. You can request detail adjustments, style modifications, or element additions in subsequent messages to progressively refine the image.

Q: What should I do if image generation fails?

A: Possible causes and solutions:

  • Prompt violates content policy: Adjust the description to avoid inappropriate content

  • System is busy: Try again later

  • Insufficient token quota: Check the remaining quota of your organization plan


Last updated

Was this helpful?