Skip to content

Image Generation ​

Browse enabled image generation models by provider in the left navigation. Open an entry for its exact model identifier, supported capabilities, and a working request.

Image Generation models use the SandBase generation protocol declared in each model registry file. Most are asynchronous: submit a request, receive an opaque run ID, then poll GET /v1/run/{id} until the generation is completed, failed, or timed out. Check the selected model page's execution mode because synchronous models return their result in the initial response.

Providers ​

OpenAI ​

  • GPT Image 2 Official API — GPT Image 2 through the OpenAI Images API contract.
  • GPT Image 2 Official Edit API — GPT Image 2 editing through the OpenAI Images API contract.
  • GPT Image 2 — GPT Image 2, OpenAI's available image model, is capable of making fine-grained, detailed edits to images.
  • GPT Image 2 Editing — GPT Image 2 Editing supports image editing and multi-image synthesis with high-quality results.
  • GPT-Image 1.5 — Gpt Image 1.5 Edit by OpenAI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • GPT Image 1.5 — Gpt Image 1.5 is OpenAI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • GPT Image 1 Mini Edit — Gpt Image 1 Mini Edit is OpenAI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • GPT Image 1 Mini — Gpt Image 1 Mini by OpenAI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • GPT Image 1 — Gpt Image 1 is OpenAI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • GPT Image 1 Edit — Gpt Image 1 Edit by OpenAI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

Google ​

  • Nano Banana 2 Lite — Nano Banana 2 Lite by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Nano Banana Lite — Nano Banana Lite by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Nano Banana Lite Edit — Nano Banana Lite Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Nano Banana 2 Image Editing — Nano Banana 2 is Google's new state-of-the-art image generation and editing model
  • Nano Banana 2 — Nano Banana 2 is Google's new state-of-the-art fast image generation and editing model
  • Gemini 3.1 Flash Image Preview — Gemini 3.1 Flash Image Preview Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Nano Banana Pro Image Editing — Nano Banana Pro is Google's new state-of-the-art image generation and editing model
  • Nano Banana Pro — Nano Banana Pro is Google's new state-of-the-art image generation and editing model
  • Gemini 2.5 Flash Image Edit — Gemini 2.5 Flash Image Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Nano Banana Image Editing — Nano Banana Pro is Google's new state-of-the-art image generation and editing model
  • Nano Banana — Google's famous original image generation and editing model.
  • Imagen 4 (Google: imagen-4 / preview / fast) — Imagen 4 Preview Fast is Google's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • …and 4 more models in the sidebar.

Ideogram ​

  • Ideogram Object Removal — Object Removal by Ideogram - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • V4.0q [instant] — 4.0 Instant is Ideogram's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • V4.0q [fast] — 4.0 Fast by Ideogram - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Ideogram V4.0q Tiling — 4.0 Tiling by Ideogram - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram V4.0q Image to Image — 4.0 by Ideogram - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram (Ideogram: custom-models) — Custom Models by Ideogram - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

ByteDance ​

  • Seedream 5.0 Pro Image Editing — Seedream 5.0 Pro is ByteDance's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Seedream 5.0 Pro Text to Image — Seedream 5.0 Pro by ByteDance - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • SeedVR2 (ByteDance: seedvr / upscale / image / seamless) — Seedvr Upscale Image is ByteDance's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • ByteDance Seed 2.0 Mini — Seed V2 Mini by ByteDance - advanced AI model for llm. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Seedream v5.0 Lite — Seedream 5.0 Lite — Fast Text-to-Image API The lightweight version of Seedream 5.0, delivering high-quality, low-latency AI image generation from text prompts. Ideal for real-time creative tools, e-commerce visuals, and high-volume AIGC pipelines.
  • Seedream v4.5 — A new-generation image creation model from ByteDance, Seedream 4.5 integrates text-to-image generation and image editing into a single unified architecture, delivering high-fidelity visuals, precise prompt control, and seamless creative workflows for professional AIGC applications.
  • Seedream v4.5 Image Editing — A new-generation image creation model from ByteDance, Seedream 4.5 integrates text-to-image generation and image editing into a single unified architecture, delivering high-fidelity visuals, precise prompt control, and seamless creative workflows for professional AIGC applications.
  • SeedVR2 (ByteDance: seedvr / upscale / video) — Seedvr Upscale Video by ByteDance - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • SeedVR2 (ByteDance: seedvr / upscale / image) — Seedvr Upscale Image by ByteDance - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • ByteDance Seedream v4 Edit — Seedream 4.0 Edit is ByteDance's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • ByteDance Seedream v4 — Seedream 4.0 by ByteDance - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Dreamina 3.1 — Dreamina 3.1 by ByteDance - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • …and 1 more models in the sidebar.

Recraft ​

  • Recraft V4.1 Utility — Recraft V4.1 Utility by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4.1 Utility Pro — Recraft V4.1 Utility by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4.1 Vector — Recraft V4.1 Vector is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4.1 — Recraft V4.1 by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4.1 Pro Vector — Recraft V4.1 Pro is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4.1 Pro — Recraft V4.1 Pro by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4 Pro (Vector) — Recraft V4 Pro is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4 (Vector) — Recraft V4 Vector is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4 Pro — Recraft V4 Pro by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4 — Recraft V4 by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft Vectorize — Recraft Vectorize is Recraft's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Recraft Creative Upscale — Recraft Upscale Creative by Recraft - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • …and 5 more models in the sidebar.

Luma ​

  • Luma Uni-1 Edit — Agent Uni 1 1.0 by Luma - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Luma Uni-1 Text to Image Max — Agent Uni 1 1.0 is Luma's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Luma Uni-1 Edit Max — Agent Uni 1 1.0 by Luma - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Luma Uni-1 Text to Image — Agent Uni 1 1.0 is Luma's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Luma Photon Flash Edit — Photon Flash 1 Edit is Luma's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Luma Photon Edit — Photon 1 Edit is Luma's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Luma Photon Flash Reframe — Photon Flash 1 Reframe by Luma - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Luma Photon Reframe — Photon 1 Reframe by Luma - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Luma Photon Flash — Photon Flash 1 by Luma - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Luma Photon — Photon 1 by Luma - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

Z-Image ​

  • Z-Image Turbo (Z-Image: turbo) — Turbo is Z-Image's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Alibaba ​

  • Qwen Image 3 — Qwen Image 3 by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image 3 Edit — Alibaba Qwen Image 3 image editing model with support for up to three reference images.
  • Qwen Image 3 Text to Image — Qwen Image 3 by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Wan 2.7 Pro Edit — Alibaba Wan 2.7 Pro image editing model with multi-reference support and configurable output format.
  • Wan 2.7 Edit — Alibaba Wan 2.7 image editing model with multi-reference support and configurable output format.
  • Wan 2.7 — Alibaba Wan 2.7 text-to-image model with high-quality generation and configurable output format.
  • Wan 2.7 Pro — Alibaba Wan 2.7 Pro text-to-image model with enhanced quality and configurable output format.
  • Z-Image Turbo Seamless Tiling (Alibaba: z-image / turbo / tiling / lora) — Z Image Turbo Tiling is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Z-Image Turbo Seamless Tiling (Alibaba: z-image / turbo / tiling) — Z Image Turbo Tiling by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image 2 (Alibaba: qwen-image-2 / pro / text-to-image) — Qwen Image 2 Pro by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image 2 Pro — Alibaba Qwen Image 2 Pro text-to-image model with enhanced quality and configurable output format.
  • Qwen Image 2 (Alibaba: qwen-image-2) — Alibaba Qwen Image 2 text-to-image model with high-quality generation and configurable output format.
  • …and 57 more models in the sidebar.

Baidu ​

  • ERNIE-Image Trainer — Ernie Image Trainer by Baidu - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Ernie Image (Baidu: ernie-image) — Ernie Image is Baidu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ernie Image (Baidu: ernie-image / turbo) — Ernie Image Turbo is Baidu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

BFL ​

Bria ​

  • Extract Object — Extract Object by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Embed Product — Embed Product is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Upscale Creative — Upscale Creative by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Replace Background — Replace Background by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Fibo Edit — Fibo Edit is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Fibo Lite — Fibo Lite is Bria's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Bria Fibo — Fibo by Bria - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Bria Reimagine 3.2 — Reimagine 3.2 is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Reimagine — Reimagine is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • BRIA RMBG 2.0 — Background Remove is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Expand Image — Expand by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Eraser — Eraser by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • …and 6 more models in the sidebar.

ClarityAI ​

  • Clarity Upscaler — Clarity Upscaler by ClarityAI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

ElevenLabs ​

  • ElevenLabs Speech to Text — Speech To Text by ElevenLabs - accurate speech-to-text transcription with AI. Convert audio and video to text with high accuracy, multilingual support, and speaker identification.

FASHN ​

  • FASHN Virtual Try-On V1.6 — Tryon V1.6 by FASHN - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • FASHN Virtual Try-On V1.5 — Tryon V1.5 by FASHN - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

HiDream ​

  • HiDream O1 Dev Edit — Hidream O1 Dev Edit is hidream-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • HiDream O1 Edit — Hidream O1 Edit is hidream-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • HiDream O1 — Hidream O1 by hidream-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • HiDream O1 Dev — Hidream O1 Dev by hidream-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • HiDream I1 Full Edit — Hidream I1 Full Edit by hidream-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • HiDream I1 Full — Hidream I1 Full is hidream-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • HiDream I1 Dev — Hidream I1 Dev by hidream-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • HiDream I1 Fast — Hidream I1 Fast is hidream-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • HiDream E1 — Hidream E1 1 by hidream-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • HiDream E1 Full — Hidream E1 Full is hidream-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

hyper3d ​

  • Hyper3d — Rodin V2 by hyper3d - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Hyper3D Rodin — Rodin is hyper3d's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.

Ideogram ​

  • Ideogram Remove Background — Ideogram Remove Background by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram (ideogram-ai: ideogram / custom-models / generate) — Ideogram Custom Models Generate is ideogram-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ideogram V3 Layerize Text — Ideogram V3 Layerize Text is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram V3 Transparent — Ideogram V3 Transparent is ideogram-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ideogram V3 Character Edit — Ideogram V3 Character Edit by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram V3 Character — Ideogram V3 Character is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram V3 Character Remix — Ideogram V3 Character Remix is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram V3 Reframe — Ideogram V3 Reframe is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram 3.0 — 3.0 is Ideogram's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ideogram V3 Replace Background — Ideogram V3 Replace Background by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram V3 Remix — Ideogram V3 Remix is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram V3 Edit — Ideogram V3 Edit by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • …and 11 more models in the sidebar.

ImagineArt ​

  • Imagineart 2.0 Edit Preview — Imagineart 2.0 Edit Preview by ImagineArt - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Imagineart 2.0 Preview — Imagineart 2.0 Preview by ImagineArt - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

KwaiVGI ​

  • Kling Video V3 Standard — Kling Video V3 Standard by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Kling Video V3 Pro — Kling Video V3 Pro is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Kling Video O3 Pro (KwaiVGI: kling-video / o3 / pro / edit) — Kling Video O3 Pro is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Kling Video O3 Pro (KwaiVGI: kling-video / o3 / pro / video-to-video) — Kling Video O3 Pro by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Kling Video O3 Standard (KwaiVGI: kling-video / o3 / standard / edit) — Kling Video O3 Standard by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Kling Video O3 Standard (KwaiVGI: kling-video / o3 / standard / video-to-video) — Kling Video O3 Standard is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Kling Image V3 — Kling Image V3 by KwaiVGI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Kling Image V3 Edit — Kling Image V3 Edit is KwaiVGI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Kling Image (KwaiVGI: kling-image / o3 / edit) — Kling Image O3 Edit is KwaiVGI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Kling Image (KwaiVGI: kling-image / o3) — Kling Image O3 by KwaiVGI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Kling Video V2.6 Standard — Kling Video V2.6 Standard by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Kling Video V2.6 Pro — Kling Video V2.6 Pro is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • …and 9 more models in the sidebar.

Lightricks ​

  • Ltx 2.3 Quality — Ltx 2.3 Quality Audio To Video by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2.3 22B Video to Video Trainer — Ltx23 V2v Trainer by Lightricks - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2.3 22B Video Trainer — Ltx23 Video Trainer by Lightricks - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2 19B Distilled — Ltx 2 19b Distilled Audio To Video by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2 19B (Lightricks: ltx-2-19b / audio-to-video) — Ltx 2 19b Audio To Video by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2 19B (Lightricks: ltx-2-19b / video-to-video) — Ltx 2 19b Video To Video by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • LTX-2 19B (Lightricks: ltx-2-19b / extend-video) — Ltx 2 19b Extend Video by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

meituan ​

  • Longcat Multi Avatar — Longcat Multi Avatar Image Audio To Video by meituan - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Longcat Single Avatar — Longcat Single Avatar Audio To Video by meituan - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Longcat Image (meituan: longcat-image / edit) — Longcat Image Edit by sandbase-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Longcat Image (meituan: longcat-image) — Longcat Image is meituan's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Meshy ​

  • Meshy Rigging Multi Animation — Rigging Multi Animation by Meshy - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Meshy Rigging — Rigging by Meshy - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Meshy 6 - Multi Image To 3D — Meshy V6 Multi Image To 3d by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Meshy 6 (Meshy: meshy-v6) — Meshy V6 by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Meshy 6 (Meshy: meshy-v6 / text-to-3d) — Meshy V6 is Meshy's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Meshy 5 Retexture — Meshy V5 Retexture by Meshy - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Meshy 5 Remesh — Meshy V5 Remesh by Meshy - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Meshy 6 Preview (Meshy: meshy-v6-preview / text-to-3d) — Meshy V6 Preview is Meshy's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Meshy 5 Multi — Meshy V5 Multi Image To 3d by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Meshy 6 Preview (Meshy: meshy-v6-preview) — Meshy V6 Preview by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.

Meta ​

  • Sam 3 1 (Meta: sam-3-1 / video) — Sam 3 1 Video is Meta's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Sam 3 1 (Meta: sam-3-1 / image-rle) — Sam 3 1 Image Rle is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Sam 3 1 (Meta: sam-3-1 / image) — Sam 3 1 Image is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • SAM 3 3D Align — Sam 3 3d Align by Meta - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Sam 3 (Meta: sam-3 / 3d-body) — Sam 3 3d Body is Meta's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Sam 3 (Meta: sam-3 / 3d-objects) — Sam 3 3d Objects by Meta - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Sam 3 (Meta: sam-3 / image-rle) — Sam 3 Image Rle is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • SAM 3 Embed — Sam 3 Image Embed by Meta - advanced AI model for vision. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Sam 3 (Meta: sam-3 / video-rle) — Sam 3 Video Rle is Meta's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • SAM 3 Image — Sam 3 Image is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Demucs — Demucs by Meta - advanced AI model for audio-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

Microsoft ​

  • MAI Image 2.5 Pro (Edit) — Mai Image 2.5 Pro Edit by Microsoft - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Mai Image 2.5 — Mai Image 2.5 Edit by Microsoft - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

MiniMax ​

  • Minimax Image Subject Reference — Image 01 Subject Reference by MiniMax - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • MiniMax (Hailuo AI) Text to Image — Image 01 by MiniMax - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

Mirelo ​

NVIDIA ​

patina ​

  • PATINA (patina: material / extract) — Material Extract by patina - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • PATINA (patina: material) — Material by patina - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

pixal3d ​

  • Pixal3d — Pixal3d is pixal3d's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.

PixVerse ​

  • PixVerse V6 Extend — V6 Extend is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

Reve ​

  • Reve 2.1 (Reve: 2.1 / remix) — 2.1 Remix is Reve's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Reve 2.1 (Reve: 2.1 / edit) — 2.1 Edit by Reve - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Reve 2.1 (Reve: 2.1 / text-to-image) — 2.1 is Reve's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

sandbase-ai ​

  • Phota Text to Image — Phota is sandbase-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Hy Wu Edit — Hy Wu Edit by sandbase-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Firered Image Edit V1.1 — Firered Image Edit V1.1 is sandbase-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • try-on — Cat VTON by sandbase-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Sana — Sana by sandbase-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Creative Upscaler — Creative Upscaler is sandbase-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

Sonilo ​

  • V1.1 Video to Sound Effects — 1.1 Video To Sound Effects by Sonilo - advanced AI model for video-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • V1.1 — 1.1 Video To Music by Sonilo - advanced AI model for video-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

Stability AI ​

  • Stable Diffusion 3.5 Large — Sd 3.5 Large by stability-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Stable Diffusion 3.5 Medium — Sd 3.5 Medium is stability-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Stable Diffusion V3 (Stability AI: stable-diffusion-v3-medium) — Stable Diffusion V3 Medium by Stability AI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • SDXL ControlNet Union (Stability AI: sdxl-controlnet-union / image-to-image) — Sdxl Controlnet Union by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • SDXL ControlNet Union (Stability AI: sdxl-controlnet-union) — Sdxl Controlnet Union is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • SDXL ControlNet Union (Stability AI: sdxl-controlnet-union / inpainting) — Sdxl Controlnet Union Inpainting by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Stable Cascade — Stable Cascade by Stability AI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Stable Diffusion XL (Stability AI: fast-sdxl) — Fast Sdxl is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Stable Diffusion V3 (Stability AI: stable-diffusion-v3-medium / image-to-image) — Stable Diffusion V3 Medium is Stability AI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • SoteDiffusion — Stable Cascade Sote Diffusion is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Stable Diffusion XL (Stability AI: fast-sdxl / image-to-image) — Fast Sdxl by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Stable Diffusion v1.5 — Stable Diffusion V15 by Stability AI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • …and 3 more models in the sidebar.

Sync Labs ​

  • sync-3 Lipsync — Lipsync V3 by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Sync React-1 — Lipsync React 1 is Sync Labs's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Sync Lipsync — Lipsync V2 Pro by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Sync Lipsync 2.0 — Lipsync V2 by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • sync.so -- lipsync 1.9.0-beta — Sync Lipsync by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

Tencent ​

  • Hunyuan 3D 3.1 Rapid Text to 3D — Hunyuan 3d 3.1 Rapid is Tencent's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Hunyuan Image 3.0 Edit — Hunyuan Image 3.0 Edit by Tencent - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Hunyuan Image 3.0 Instruct — Hunyuan Image 3.0 Instruct by Tencent - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Hunyuan 3D Smart Topology — Hunyuan 3d 3.1 Smart Topology by Tencent - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Hunyuan 3D 3.1 Rapid Image to 3D — Hunyuan 3d 3.1 Rapid by Tencent - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Hunyuan 3D 3.1 Pro Text to 3D — Hunyuan 3d 3.1 Pro is Tencent's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Hunyuan 3D 3.1 Pro Image to 3D — Hunyuan 3d 3.1 Pro by Tencent - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Hunyuan 3D Part Splitter — Hunyuan 3d 3.1 Part by Tencent - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Hunyuan Motion Fast — Hunyuan Motion Fast is Tencent's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Hunyuan Motion — Hunyuan Motion by Tencent - generate 3D models from text descriptions with AI. Create detailed 3D assets from natural language for games, visualization, and digital production.
  • Hunyuan3d V3 (Tencent: hunyuan-3d / v3 / text-to-3d) — Hunyuan 3d V3 by Tencent - generate 3D models from text descriptions with AI. Create detailed 3D assets from natural language for games, visualization, and digital production.
  • Hunyuan3d V3 (Tencent: hunyuan-3d / v3 / sketch-to-3d) — Hunyuan 3d V3 Sketch To 3d by Tencent - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • …and 11 more models in the sidebar.

Topaz Labs ​

  • Topaz Upscale — Upscale Image is Topaz Labs's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

trellis ​

  • Trellis 2 (trellis: trellis-2 / retexture) — Trellis 2 Retexture is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Trellis 2 (trellis: trellis-2) — Trellis 2 is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Trellis (trellis: trellis / multi) — Trellis Multi is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Trellis (trellis: trellis) — Trellis is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.

Tripo3D ​

  • Triposplat — Triposplat by Tripo3D - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Tripo H3.1 Multiview to 3D — Tripo H3.1 Multiview To 3d by Tripo3D - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Tripo H3.1 Text to 3D — Tripo H3.1 is Tripo3D's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Tripo H3.1 Image to 3D — Tripo H3.1 by Tripo3D - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Tripo P1 Text to 3D — Tripo P1 is Tripo3D's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Tripo P1 Image to 3D — Tripo P1 by Tripo3D - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.

VEED ​

  • Subtitles — Subtitles is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

Vidu ​

  • Vidu Q2 Reference to Image — Vidu Q2 Reference To Image by Shengshu - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Vidu Q2 Text to Image — Vidu Q2 is Shengshu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Vidu 2.0 Reference to Image — Vidu 2.0 Reference To Image is Shengshu's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

xAI ​

  • Grok Imagine Image Quality — Grok Imagine Image Quality by xAI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Grok Imagine Image Quality Edit — Grok Imagine Image Quality Edit is xAI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Grok Imagine Video Extend — Grok Imagine Video Extend is xAI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Grok Imagine Image Edit — Grok Imagine Image Edit is xAI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Grok Imagine Video Edit — Grok Imagine Video Edit is xAI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

zhipu ​

  • Glm Image (zhipu: glm-image / edit) — Glm Image Edit by zhipu - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Glm Image (zhipu: glm-image) — Glm Image is zhipu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Capability coverage ​

3d-to-3d, audio-to-audio, audio-to-text, audio-to-video, commercial, image-editing, image-to-3d, image-to-image, image-to-text, llm, mcp_exposable, protocol-ingress-only, speech-to-text, text-to-3d, text-to-image, training, video-to-audio, video-to-text, video-to-video, vision