Use in agentzhipu models

zhipu modelsimage generation api

zhipu/glm-image/edit

Glm Image Edit by zhipu - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

Input
Text prompt for image generation.

PNG, JPEG, WebP, or GIF · 20 MiB maximum each

URL(s) of the condition image(s) for image-to-image generation. Supports up to 4 URLs for multi-image references.
110
Classifier-free guidance scale. Higher values make the model follow the prompt more closely. Range: 1 to 10.
10100
Number of diffusion denoising steps. More steps generally produce higher quality images. Range: 10 to 100.
Output image format. Allowed values: jpeg, png.
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
Random seed for reproducibility. The same seed with the same prompt will produce the same image.
Idle

Example output — click Run to generate your own

API README

GLM Image Edit

GLM Image Edit is a specialized SandBase endpoint built to revise one or more condition images through text while retaining direct control over generation strength and refinement depth. It is most useful when teams need multi-image conditioning, deterministic iteration, adjustable prompt guidance, and selectable denoising effort, with the route’s inputs defining a repeatable production contract instead of leaving critical delivery choices to an ad-hoc manual workflow.

Use this exact route when its input mode matches the creative asset already in hand: describe the visual or spoken result clearly, set only the controls that support the intended delivery, and keep the subject, action, environment, camera or performance direction internally consistent. The result is returned asynchronously as a downloadable media URL suitable for review, automation, or downstream finishing.

Highlights

  • Multi-image conditioning. Supply as many as four visual references when an edit must reconcile subject, style, product, or layout cues from several sources.
  • Reproducible iterations. Reuse a seed with the same inputs to revisit a promising result and compare controlled prompt adjustments.
  • Adjustable prompt adherence. Guidance scale from 1 to 10 lets teams decide how strongly the written instruction should steer the revised image.
  • Refinement-depth control. Choose 10–100 inference steps to trade faster exploration for a more thoroughly resolved generation pass.

Pricing

Each image-edit request costs $0.0500.

OperationPrice
GLM Image edit$0.0500

When to Use

ScenarioWhy this route fits
Concept developmentChoose GLM Image Edit when its edit workflow matches the starting material and you need several clearly directed variations.
Production iterationUse explicit duration, resolution, ratio, seed, or quality controls to compare versions without changing the core creative brief.
Channel adaptationGenerate directly in the landscape, square, portrait, or vertical format required by the destination whenever that control is available.
Automated pipelinesIntegrate the asynchronous media URL into review queues, asset libraries, publishing tools, or a later finishing stage.
Alternative routePick a related text-, image-, reference-, edit-, or turbo route when the available source media or required degree of control is different.

Prompt Guide

Lead with the main subject or source asset, then describe the intended action or transformation, environment, composition, camera or vocal delivery, lighting and mood. Keep instructions concrete and compatible; use the route’s explicit fields for duration, resolution, ratio, quality, voice, or reproducibility instead of burying those settings in prose.

{
  "prompt": "A cinematic product reveal with deliberate subject motion, coherent lighting, and a slow camera push.",
  "aspect_ratio": "21:9",
  "output_format": "jpeg",
  "guidance_scale": 1.5,
  "num_inference_steps": 30,
  "images": [
    "https://example.com/reference-1.jpg"
  ]
}

Technical Specs

PropertyDetails
Model IDzhipu/glm-image/edit
Required inputsprompt
ExecutionAsynchronous; poll the returned generation ID
OutputDownloadable media URL
seedinteger; optional
imagesarray; optional
promptstring; required
aspect_ratiostring; optional; choices: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16
output_formatstring; optional; choices: jpeg, png; default: jpeg
guidance_scalenumber; optional; range: 1–10; default: 1.5
num_inference_stepsinteger; optional; range: 10–100; default: 30

Related Models

ModelBest for
zhipu/glm-imageAlternative glm image workflow
bfl/flux-2/devAlternative image generation or editing workflow
google/nano-banana-proAlternative image generation or editing workflow
bytedance/seedream/4.5/editAlternative image generation or editing workflow

Related Models