SandBase is live — $1 in free credits on signupStart free ›

Alibaba modelsimage generation api

alibaba/qwen-image/max/edit

Qwen Image Max Edit is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

Input
Text prompt describing the desired image. Supports Chinese and English. Max 800 characters.

PNG, JPEG, WebP, or GIF · 20 MiB maximum each

Reference images for editing (1-3 images required). Order matters: reference as 'image 1', 'image 2', 'image 3' in prompt. Resolution: 384-5000px each dimension. Max size: 10MB each. Formats: JPEG, JPG, PNG (no alpha), WEBP.
The format of the generated image. Allowed values: jpeg, png.
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
Random seed for reproducibility (0-2147483647).
Idle

Example output — click Run to generate your own

API README

Qwen Image Max Edit

Qwen Image Max Edit is the Qwen Image route for instruction-driven image revision. It uses Qwen’s combined visual-semantic and appearance modeling to interpret subjects, relationships, typography, spatial structure, and surface detail as a connected composition, giving this exact route a clearer production purpose than a generic image request.

Use Qwen Image Max Edit when the project specifically requires a controlled change to supplied imagery. State the creative goal, then define composition, viewpoint, wording, materials, lighting, and finish; for edits, explicitly separate the intended transformation from identity, layout, or regions that must remain stable.

Highlights

  • Qwen Image Semantic content editing. Changes object identity, scene meaning, viewpoint, or style through natural language.
  • Qwen Image Appearance-level revision. Modifies local color, texture, text, or material while preserving surrounding context.
  • Qwen Image Chinese and English text editing. Adds, removes, or replaces bilingual wording in designed imagery.
  • Maximum-fidelity Qwen revision. Prioritizes complex edit interpretation and preservation of small visual and textual details.

Pricing

ConfigurationBilling unitPrice
Base generationPer request$0.075

When to Use

✅ Good fit❌ Consider alternatives
The project needs this exact named workflowThe intended task belongs to another media route
All required reference or control media is availableNecessary assets or rights are unavailable
The brief can state transformation and preservation goalsOutput must be deterministic at pixel or frame level
Supported duration, resolution, and format fit deliveryFinal placement requires unsupported specifications
An asynchronous generated result fits productionA live frame-synchronous response is mandatory

Prompt Guide

Identify the primary subject and every source or condition, state the intended transformation or action, then describe composition, camera or viewpoint, lighting, materials, pacing, atmosphere, and exact preservation requirements. Refer to multiple inputs in their schema order.

{
  "prompt": "A precisely directed composition with explicit subject, transformation, camera or viewpoint, lighting, material, and preservation requirements",
  "seed": 1,
  "images": [
    "https://example.com/reference.png"
  ],
  "aspect_ratio": "21:9"
}

Technical Specs

SpecValue
Model IDalibaba/qwen-image/max/edit
Input fieldsseed (integer)<br>images (array)<br>prompt (string)<br>aspect_ratio (string; 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16)<br>output_format (string; jpeg, png)
Required inputprompt
Output fieldsurl, content_type
ExecutionAsynchronous job

Related Models

  • alibaba/qwen-image
  • alibaba/qwen-image/edit
  • alibaba/qwen-image/max

Related Models

alibaba/qwen-image/maxQwen Image Max by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.alibaba/qwen-image/2512Qwen Image 2512 is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.alibaba/qwen-image/2512/loraQwen Image 2512 Lora by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.alibaba/qwen-image/editQwen Image Edit is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.alibaba/qwen-image/layeredQwen Image Layered by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.alibaba/qwen-image/layered/loraQwen Image Layered Lora is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.alibaba/qwen-imageQwen Image by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.alibaba/qwen-image-2/pro/text-to-imageQwen Image 2 Pro by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.