SandBase is live — $1 in free credits on signupStart free ›

Alibaba modelsimage generation api

alibaba/z-image/turbo/edit

Z Image Turbo Edit by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

Input
The prompt to generate an image from.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of Image for Image-to-Image generation.
The strength of the image-to-image conditioning. Maximum: 1.
The format of the generated image. Allowed values: jpeg, png.
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
The same seed and the same prompt given to the same version of the model will output the same image every time.
Idle

Example output — click Run to generate your own

API README

Z-Image Turbo Edit

Z-Image Turbo Edit Z-Image Turbo Edit uses Z-Image Turbo to revise supplied imagery from natural-language direction. Its modeling balances prompt interpretation, composition, and visual detail for creative use.

State the requested change and separately list the source identity, composition, and details that must remain stable. For reliable iteration, make one coherent transformation at a time and review fine boundaries, reflected light, object scale, and source identity before adding the next requested change.

Highlights

  • Z-Image Turbo Edit Fast semantic image editing. Applies natural-language changes to objects, setting, style, and composition.
  • Z-Image Turbo Edit Source-detail retention. Keeps unrequested identity and layout connected to the original image.
  • Z-Image Turbo Edit Context-aware reconstruction. Matches new content to existing perspective, texture, and illumination.
  • Z-Image Turbo Edit Rapid revision cycles. Enables multiple creative edits through the Turbo generation path.

Pricing

Billing unitPrice
Per request$0.006

When to Use

✅ Good fit❌ Consider alternatives
The model's named workflow matches the source material and intended outputA different input modality or model route is required
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For image-conditioned generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "aspect_ratio": "21:9",
  "image": "https://example.com/start-frame.png",
  "output_format": "png",
  "prompt": "A young Asian woman with long, vibrant purple hair stands on a sunlit sandy beach, posing confidently with her left hand resting on her hip. She gazes directly at the camera with a neutral expression. A sleek black ribbon bow is tied neatly on the right side of her head, just above her ear. She wears a flowing white cotton dress with a fitted bodice and a flared skirt that reaches mid-calf, slightly lifted by a gentle sea breeze. The beach behind her features fine, pale golden sand with subtle footprints, leading to calm turquoise waves under a clear blue sky with soft, wispy clouds. The lighting is natural daylight, casting soft shadows to her left, indicating late afternoon sun. The horizon line is visible in the background, with a faint silhouette of distant dunes. Her skin tone is fair with a natural glow, and her facial features are delicately defined. The composition is centered on her figure, framed from mid-thigh up, with shallow depth of field blurring the distant waves slightly."
}

Technical Specs

SpecValue
Model IDalibaba/z-image/turbo/edit
Inputsaspect_ratio, image, output_format, prompt, seed, strength
Required inputsprompt, image
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Aspect Ratio21:9 / 16:9 / 3:2 / 4:3 / 5:4 / 1:1 / 4:5 / 3:4 / 2:3 / 9:16
Output Formatjpeg / png

Related Models

Related Models

alibaba/z-image/turbo/edit/loraZ Image Turbo Edit is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.alibaba/z-image/turboZ Image Turbo is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.alibaba/z-image/turbo/inpaintZ Image Turbo Inpaint is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.alibaba/z-image/turbo/inpaint/loraZ Image Turbo Inpaint by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.alibaba/z-image/turbo/loraZ Image Turbo Lora by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.alibaba/z-image/turbo/tilingZ Image Turbo Tiling by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.alibaba/z-image/turbo/tiling/loraZ Image Turbo Tiling is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.alibaba/z-image/base/loraZ Image Base Lora is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.