SandBase is live — $1 in free credits on signupStart free ›

stability-ai modelsimage generation api

stability-ai/sd/3.5-medium

Sd 3.5 Medium is stability-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Input
The prompt to generate an image from.
The aspect ratio of the generated image. Allowed values: 21:9, 16:9, 3:2, 4:3, 5:4, 1:1, 4:5, 3:4, 2:3, 9:16.
The format of the generated image. Allowed values: jpeg, png.
Idle

Example output — click Run to generate your own

API README

Stable Diffusion 3.5 Medium

Stable Diffusion 3.5 Medium is a text-to-image model built on a Multimodal Diffusion Transformer architecture. It turns a written scene into a finished raster image. The route uses one prompt as the creative input.

A request defines the desired subject and scene, then chooses a canvas ratio and raster format. The completed result is returned as an image URL. This keeps the workflow focused on generating a new image rather than editing supplied media.

Highlights

  • Improved image quality. The exact model page identifies stronger visual quality as a defining improvement. This improvement applies to the generated image.
  • Better typography. Stable Diffusion 3.5 Medium explicitly improves rendering for text-bearing images. Typography is identified separately from general image quality.
  • Complex prompt understanding. The model is designed to interpret prompts with multiple visual requirements. Complex prompt understanding is an explicit model capability.
  • Efficient MMDiT design. The exact route highlights computational efficiency alongside its multimodal diffusion-transformer architecture. Computational efficiency is listed alongside the model's other improvements.

Pricing

Billing unitPrice
Per generated image$0.02

When to Use

✅ Good fit❌ Consider alternatives
Producing concept imagery from a detailed written briefEditing a supplied image or mask
Exploring compositions that include visible letteringDelivering editable vector typography
Creating varied styles from one text workflowRequiring video, animation, or audio
Generating campaign drafts across portrait and landscape layoutsNeeding a transparent layered design file
Choosing the balanced 3.5 tier for routine generationChoosing maximum family capacity regardless of compute profile

Prompt Guide

Describe the subject first, then its setting, composition, visual style, lighting, and any text that must appear. Put exact visible wording in quotation marks and explain where it belongs.

Subject: [main subject and action]
Setting: [environment and context]
Composition: [camera, framing, placement]
Style: [medium and visual treatment]
Lighting/color: [light direction and palette]
Visible text: "[exact wording]" placed [location]
{
  "prompt": "Editorial poster of a glass greenhouse at dusk, centered symmetrical composition, lush botanical photography, warm interior light against a deep blue sky, the title \"NIGHT GARDEN\" in clean white lettering at the top",
  "aspect_ratio": "4:5",
  "output_format": "png"
}

Technical Specs

SpecValue
Model IDstability-ai/sd/3.5-medium
Required inputprompt
Aspect ratio10 presets from 21:9 through 9:16
Output formatjpeg or png; default jpeg
OutputImage URL and content type
ExecutionAsynchronous job

Related

Related Models

stability-ai/sd/3.5-largeSd 3.5 Large by stability-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.stability-ai/fast-sdxl-controlnet-cannyFast Sdxl Controlnet Canny by Stability AI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.stability-ai/fast-sdxl-controlnet-canny/image-to-imageFast Sdxl Controlnet Canny is Stability AI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.stability-ai/fast-sdxl/image-to-imageFast Sdxl by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.stability-ai/fast-sdxl/inpaintingFast Sdxl Inpainting by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.stability-ai/fast-sdxlFast Sdxl is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.stability-ai/sdxl-controlnet-unionSdxl Controlnet Union is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.stability-ai/sdxl-controlnet-union/image-to-imageSdxl Controlnet Union by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.