gemini-3-pro-image-preview — Image Edit AI Model

Nano Banana generates and edits images through Gemini's multimodal reasoning, so it follows a complicated instruction the way a language model does rather than the way diffusion guesses. Needs a Google AI key.

Details

  • CategoryImage Edit
  • Year2025
  • LicenseGoogle Cloud Terms

Compliance & Provenance

  • ProviderGoogle · GPAI
  • EU AI Act RiskLimited Risk
  • Art. 50 TransparencyRequired — AI-generated outputs are marked

Inputs & Outputs

  • TextInput · string

    Text prompt for image generation or editing instruction

  • ImageInput · image · optional

    Input image for editing

  • ImageOutput · image

    Generated or edited image

  • Generated TextOutput · string

    Text response from the model

Tags

  • image-generation
  • image-editing
  • multimodal
  • text-to-image

Alternatives in Image/Video Models

  • GPT Image

    OpenAI GPT-Image text-to-image / edit. Needs API key.

  • Omni

    Google Gemini Omni: text/image/video in, video out, with conversational editing. Which input ports apply depends on Task — Text to Video: Text only. Image to Video: + Image 1. Reference to Video: + Image 1–3. Edit: + Video (aspect ratio is ignored). Unused ports are dropped even if wired. Needs API key.

  • Sora 2

    OpenAI Sora 2 text-to-video / img2video. Needs API key.

  • VEO

    Google VEO 3.1 video gen (up to 4K, with native audio). Needs API key.

  • Video Analysis

    Gemini video analysis. Feed a video and get a transcript, timecodes, or highlight analysis as text. Needs API key.

Resources