SandBase is live — $1 in free credits on signupStart free ›

Google modelsimage generation api

google/gemini-3.1-flash-image-preview/edit

Gemini 3.1 Flash Image Preview Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

Input
The prompt for image editing.

PNG, JPEG, WebP, or GIF · 20 MiB maximum each

The URLs of the images to use for image-to-image generation or image editing.
The resolution of the image to generate. Allowed values: 0.5K, 1K, 2K, 4K.
The format of the generated image. Allowed values: jpeg, png.
Optional system instruction that steers the model's persona and output style across the request. Leave blank to omit; when provided, it is sent as the system instruction to Gemini (or as a system message on OpenAI-compatible providers).
When set, enables model thinking with the given level ('minimal' or 'high') and includes thoughts in the generation. Omit to disable. Allowed values: minimal, high.
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
The seed for the random number generator.
Idle

Example output — click Run to generate your own

API README

Gemini 3.1 Flash Image Preview Edit

google/gemini-3.1-flash-image-preview/edit revises one or more supplied images from natural-language direction, combining visual understanding with high-quality image generation. It can preserve important source details while changing objects, layout, style, lighting, or typography, and its reasoning modes help it plan more involved edits. This combination makes the model a practical choice when the creative outcome depends on those qualities rather than on a generic media conversion.

For production work, multiple references can be combined into one composition, ten aspect ratios cover common placements, and output from 0.5K through 4K supports drafts as well as polished masters. The result is most reliable when the source material and creative brief clearly describe the intended subject, progression, visual or sonic character, and the qualities that must remain unchanged.

Highlights

Multi-image comprehension combines information from several supplied references.

Instruction-led editing changes selected visual attributes while retaining requested source details.

Strong text rendering places legible, well-integrated typography inside generated images.

High-resolution generation delivers detailed results through 4K for demanding creative work.

Pricing

Output resolutionPrice per image
0.5K$0.060
1K$0.080
2K$0.120
4K$0.160

When to Use

✅ Good fit❌ Consider alternatives
The model's named workflow matches the source material and intended outputA different input modality or model route is required
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "aspect_ratio": "21:9",
  "images": [
    "https://static.sandbase.ai/examples/google/gemini-3.1-flash-image-preview/edit/input_images_0.png"
  ],
  "output_format": "png",
  "prompt": "make a photo of the man driving the car down the california coastline",
  "resolution": "1K"
}

Technical Specs

SpecValue
Model IDgoogle/gemini-3.1-flash-image-preview/edit
Inputsaspect_ratio, images, output_format, prompt, resolution, seed, system_prompt, thinking_level
Required inputsprompt
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Resolution0.5K / 1K / 2K / 4K
Aspect Ratio21:9 / 16:9 / 3:2 / 4:3 / 5:4 / 1:1 / 4:5 / 3:4 / 2:3 / 9:16
Output Formatjpeg / png

Related Models

Related Models

google/gemini-2.5-flash-image/editGemini 2.5 Flash Image Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.google/imagen-3Imagen 3 by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.google/imagen-3/fastImagen 3 Fast is Google's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.google/imagen-4/previewImagen 4 Preview by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.google/imagen-4/preview/fastImagen 4 Preview Fast is Google's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.google/imagen-4/preview/ultraImagen 4 Preview Ultra by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.google/nano-bananaGoogle's famous original image generation and editing model.google/nano-banana-2Nano Banana 2 is Google's new state-of-the-art fast image generation and editing model