SandBase is live — $1 in free credits on signupStart free ›

Alibaba modelsimage generation api

alibaba/z-image/turbo/inpaint

Z Image Turbo Inpaint is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

Input
The prompt to generate an image from.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of Image for Inpaint generation.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of Mask for Inpaint generation.
The strength of the inpaint conditioning. Maximum: 1.
The format of the generated image. Allowed values: jpeg, png.
01
The end of the controlnet conditioning. Range: 0 to 1.
01
The scale of the controlnet conditioning. Range: 0 to 1.
01
The start of the controlnet conditioning. Range: 0 to 1.
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
The same seed and the same prompt given to the same version of the model will output the same image every time.
Idle

Example output — click Run to generate your own

API README

Z-Image Turbo Inpaint

Z-Image Turbo Inpaint Z-Image Turbo Inpaint reconstructs selected regions with the fast Z-Image Turbo model. Its modeling balances prompt interpretation, composition, and visual detail for creative use.

Mask only the intended area and describe the content, perspective, material, and lighting that should fill it. Inspect difficult transitions such as hair, transparent material, repeated texture, shadow, and reflection at full size; these boundaries reveal whether the reconstructed patch truly belongs to the original image.

Highlights

  • Z-Image Turbo Inpaint Masked content reconstruction. Generates new content only inside selected image regions.
  • Z-Image Turbo Inpaint Boundary-consistent blending. Matches texture, perspective, color, and lighting around the mask.
  • Z-Image Turbo Inpaint Prompt-directed replacement. Adds, removes, or transforms localized objects through language.
  • Z-Image Turbo Inpaint Outside-mask preservation. Retains the source image beyond the requested edit area.

Pricing

Billing unitPrice
Per request$0.006

When to Use

✅ Good fit❌ Consider alternatives
A selected time region in existing audio needs regenerationThe entire track should be recomposed
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For image-conditioned generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "aspect_ratio": "21:9",
  "image": "https://example.com/start-frame.png",
  "output_format": "png",
  "prompt": "A young Asian woman with long, vibrant purple hair stands on a sunlit sandy beach, posing confidently with her left hand resting on her hip. She gazes directly at the camera with a neutral expression. A sleek black ribbon bow is tied neatly on the right side of her head, just above her ear. She wears a flowing white cotton dress with a fitted bodice and a flared skirt that reaches mid-calf, slightly lifted by a gentle sea breeze. The beach behind her features fine, pale golden sand with subtle footprints, leading to calm turquoise waves under a clear blue sky with soft, wispy clouds. The lighting is natural daylight, casting soft shadows to her left, indicating late afternoon sun. The horizon line is visible in the background, with a faint silhouette of distant dunes. Her skin tone is fair with a natural glow, and her facial features are delicately defined. The composition is centered on her figure, framed from mid-thigh up, with shallow depth of field blurring the distant waves slightly."
}

Technical Specs

SpecValue
Model IDalibaba/z-image/turbo/inpaint
Inputsaspect_ratio, control_end, control_scale, control_start, image, mask, output_format, prompt, seed, strength
Required inputsprompt, image
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Aspect Ratio21:9 / 16:9 / 3:2 / 4:3 / 5:4 / 1:1 / 4:5 / 3:4 / 2:3 / 9:16
Output Formatjpeg / png

Related Models

Related Models

alibaba/z-image/turbo/inpaint/loraZ Image Turbo Inpaint by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.alibaba/z-image/turbo/editZ Image Turbo Edit by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.alibaba/z-image/turbo/edit/loraZ Image Turbo Edit is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.alibaba/z-image/turboZ Image Turbo is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.alibaba/z-image/turbo/loraZ Image Turbo Lora by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.alibaba/z-image/turbo/tilingZ Image Turbo Tiling by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.alibaba/z-image/turbo/tiling/loraZ Image Turbo Tiling is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.alibaba/z-image/base/loraZ Image Base Lora is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.