gemini-3-pro-image-preview — Image Edit AI Model
Nano Banana generates and edits images through Gemini's multimodal reasoning, so it follows a complicated instruction the way a language model does rather than the way diffusion guesses. Needs a Google AI key.
Details
- CategoryImage Edit
- Year2025
- LicenseGoogle Cloud Terms
Compliance & Provenance
- ProviderGoogle · GPAI
- EU AI Act RiskLimited Risk
- Art. 50 TransparencyRequired — AI-generated outputs are marked
Inputs & Outputs
- TextInput · string
Text prompt for image generation or editing instruction
- ImageInput · image · optional
Input image for editing
- ImageOutput · image
Generated or edited image
- Generated TextOutput · string
Text response from the model
Tags
Alternatives in Image/Video Models
- GPT Image
OpenAI GPT-Image text-to-image / edit. Needs API key.
- Omni
Google Gemini Omni: text/image/video in, video out, with conversational editing. Which input ports apply depends on Task — Text to Video: Text only. Image to Video: + Image 1. Reference to Video: + Image 1–3. Edit: + Video (aspect ratio is ignored). Unused ports are dropped even if wired. Needs API key.
- Sora 2
OpenAI Sora 2 text-to-video / img2video. Needs API key.
- VEO
Google VEO 3.1 video gen (up to 4K, with native audio). Needs API key.
- Video Analysis
Gemini video analysis. Feed a video and get a transcript, timecodes, or highlight analysis as text. Needs API key.