SandBase is live — $1 in free credits on signupStart free ›

Google modelsimage generation api

google/imagen-4/preview/fast

Imagen 4 Preview Fast is Google's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Input
The text prompt to generate an image from.
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
The format of the generated image. Allowed values: jpeg, png.
The seed for the random number generator.
Idle

Example output — click Run to generate your own

API README

Imagen 4 Preview Fast

google/imagen-4/preview/fast creates original still imagery from detailed written art direction, resolving subject, environment, composition, lighting, material, and style without a source image. Imagen 4 improves fine detail, spelling and typography, prompt interpretation, and overall visual quality beyond the previous generation; this Fast variant reduces turnaround so teams can compare more concepts and compositions in the same creative window. This combination makes the model a practical choice when the creative outcome depends on those qualities rather than on a generic media conversion.

For production work, its speed-oriented generation is best used for rapid ideation, high-volume variants, and early art-direction review before choosing a slower quality-focused route. The result is most reliable when the source material and creative brief clearly describe the intended subject, progression, visual or sonic character, and the qualities that must remain unchanged.

Highlights

Fine-detail rendering resolves textures, small structures, and complex surfaces with greater precision.

Improved typography produces more accurate spelling and better integrated graphic text.

Stronger prompt interpretation follows nuanced composition, style, and subject relationships.

Rapid generation accelerates concept comparison and high-volume creative exploration.

Pricing

Billing unitPrice
Per request$0.02

When to Use

✅ Good fit❌ Consider alternatives
The model's named workflow matches the source material and intended outputA different input modality or model route is required
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "aspect_ratio": "21:9",
  "output_format": "png",
  "prompt": "Atmospheric narrative illustration depicting a young woman with dark hair styled with a single star clip, eating dumplings at a small round table in a bustling, late-night eatery reminiscent of a vintage Hong Kong diner. The style blends clean linework with textured color fields, evoking a sense of place and story. The mood is intimate contentment amidst vibrant surroundings. Soft, warm overhead lighting from unseen hanging lamps casts gentle highlights on her face and the porcelain plate of dumplings, creating soft-edged shadows on the tiled tabletop and floor. The background features detailed elements like wall menus with stylized illustrations, a retro wall clock, steam rising from a soup bowl, and glimpses of other patrons blurred slightly for depth. The woman, viewed from a slightly high angle, crouches slightly on her chair, intensely focused on her food, rendered with expressive linework defining her pose and features. The color palette mixes muted teal wall tiles and green chairs with pops of warm yellow in her top, pink trousers, red chili oil dish, and ambient light, creating a cozy yet lively feel. Subtle paper texture or digital grain is visible throughout. Focus is sharp on the character and her immediate table setting"
}

Technical Specs

SpecValue
Model IDgoogle/imagen-4/preview/fast
Inputsaspect_ratio, output_format, prompt, seed
Required inputsprompt
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Aspect Ratio21:9 / 16:9 / 3:2 / 4:3 / 5:4 / 1:1 / 4:5 / 3:4 / 2:3 / 9:16
Output Formatjpeg / png

Related Models

Related Models

google/imagen-4/previewImagen 4 Preview by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.google/imagen-4/preview/ultraImagen 4 Preview Ultra by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.google/imagen-3/fastImagen 3 Fast is Google's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.google/gemini-2.5-flash-image/editGemini 2.5 Flash Image Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.google/imagen-3Imagen 3 by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.google/nano-bananaGoogle's famous original image generation and editing model.google/nano-banana-2Nano Banana 2 is Google's new state-of-the-art fast image generation and editing modelgoogle/nano-banana-2-liteNano Banana 2 Lite by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.