GPT Image 2
About
GPT Image 2, OpenAI's latest image model, is capable of making fine-grained, detailed edits to images.
Documentation
OpenAI GPT Image 2 Text-to-Image
OpenAI's most capable image generation model. Excels at complex scenes with accurate text rendering, strong prompt adherence, and photorealistic output across Latin and CJK scripts.
- Need to edit an existing image? Try OpenAI GPT Image 2 Edit.
Highlights
Near-perfect text rendering — Generates readable text inside images including signage, UI labels, handwritten notes, and posters. Supports Chinese, Japanese, Korean, and Latin scripts with correct spelling and consistent spacing.
Extreme prompt adherence — Follows complex multi-part instructions faithfully. Composition, lighting, material properties, and fine-grained details described in long prompts are preserved without drift.
Neutral, accurate color — Eliminates the warm color cast present in earlier models. Colors render true-to-intent across all scene types, critical for brand assets and product photography.
Flexible resolution up to 4K — Supports preset sizes and custom dimensions. Both edges must be multiples of 16, max single edge 3840px, total pixels between 655K and 8.3M.
Pricing
| Quality | Standard Size | Large (21:9 / 9:21) |
|---|---|---|
| Low | $0.015 | $0.03 |
| Medium | $0.05 | $0.08 |
| High | $0.20 | $0.30 |
Standard covers most aspect ratios (1:1, 2:3, 3:2, 4:3, 16:9, etc). Large applies to ultra-wide/tall formats only.
When to Use
| ✅ Good fit | ❌ Consider alternatives |
|---|---|
| Text-heavy images (posters, ads, packaging) | Simple icons or flat illustrations |
| Complex multi-object scenes | Real-time / sub-second generation |
| Brand-accurate color requirements | Iterative mask-based editing (use Edit model) |
| Multilingual text in images | Video or animation |
| Product mockups and concept art | Extremely large batch jobs on budget |
Prompt Guide
Structure your prompts for best results:
Scene: [environment, time of day, background]
Subject: [main focus, position, action]
Details: [materials, lighting, camera angle, mood]
Text: [exact text to render, font style, placement]
Constraints: [no watermark, specific aspect ratio, color palette]
Example:
A minimalist product photo of a ceramic coffee mug on a marble countertop, morning sunlight from the left, soft shadows. The mug has "Good Morning" written in a clean serif font. White background, 3:2 aspect ratio, no other objects.
Technical Specs
| Spec | Value |
|---|---|
| Input | Text prompt (max 32,000 characters) |
| Output | PNG, JPEG, or WebP |
| Resolution | 655,360 – 8,294,400 total pixels |
| Max edge | 3,840 px |
| Aspect ratio | Up to 3:1 |
| Edge constraint | Multiples of 16 |
| Execution | Async (submit → poll for result) |
Related
- OpenAI GPT Image 2 Edit — Edit existing images with multi-image input support
Try GPT Image 2
Test this model in the Sandbase Playground with your own prompts.
Open in Playground