GPT Image 2

OpenAIimage

About

GPT Image 2, OpenAI's latest image model, is capable of making fine-grained, detailed edits to images.

Documentation

OpenAI GPT Image 2 Text-to-Image

OpenAI's most capable image generation model. Excels at complex scenes with accurate text rendering, strong prompt adherence, and photorealistic output across Latin and CJK scripts.

Highlights

Near-perfect text rendering — Generates readable text inside images including signage, UI labels, handwritten notes, and posters. Supports Chinese, Japanese, Korean, and Latin scripts with correct spelling and consistent spacing.

Extreme prompt adherence — Follows complex multi-part instructions faithfully. Composition, lighting, material properties, and fine-grained details described in long prompts are preserved without drift.

Neutral, accurate color — Eliminates the warm color cast present in earlier models. Colors render true-to-intent across all scene types, critical for brand assets and product photography.

Flexible resolution up to 4K — Supports preset sizes and custom dimensions. Both edges must be multiples of 16, max single edge 3840px, total pixels between 655K and 8.3M.

Pricing

QualityStandard SizeLarge (21:9 / 9:21)
Low$0.015$0.03
Medium$0.05$0.08
High$0.20$0.30

Standard covers most aspect ratios (1:1, 2:3, 3:2, 4:3, 16:9, etc). Large applies to ultra-wide/tall formats only.

When to Use

✅ Good fit❌ Consider alternatives
Text-heavy images (posters, ads, packaging)Simple icons or flat illustrations
Complex multi-object scenesReal-time / sub-second generation
Brand-accurate color requirementsIterative mask-based editing (use Edit model)
Multilingual text in imagesVideo or animation
Product mockups and concept artExtremely large batch jobs on budget

Prompt Guide

Structure your prompts for best results:

Scene: [environment, time of day, background]
Subject: [main focus, position, action]
Details: [materials, lighting, camera angle, mood]
Text: [exact text to render, font style, placement]
Constraints: [no watermark, specific aspect ratio, color palette]

Example:

A minimalist product photo of a ceramic coffee mug on a marble countertop, morning sunlight from the left, soft shadows. The mug has "Good Morning" written in a clean serif font. White background, 3:2 aspect ratio, no other objects.

Technical Specs

SpecValue
InputText prompt (max 32,000 characters)
OutputPNG, JPEG, or WebP
Resolution655,360 – 8,294,400 total pixels
Max edge3,840 px
Aspect ratioUp to 3:1
Edge constraintMultiples of 16
ExecutionAsync (submit → poll for result)

Related

Try GPT Image 2

Test this model in the Sandbase Playground with your own prompts.

Open in Playground

Related Models