Skip to content

Image Generation ​

SandBase currently publishes API reference pages for 416 enabled image generation models across 40 providers. Choose a provider in the left navigation, then open a model page for its exact API identifier, supported capabilities, and a working request.

Image Generation models use the async SandBase generation protocol declared in each model registry file. Submit a request, receive a task id, then poll the result endpoint until the generation is completed, failed, or timed out.

Providers ​

OpenAI ​

  • GPT Image 2 Editing — GPT Image 2 Editing supports image editing and multi-image synthesis with high-quality results.
  • GPT-Image 1.5 — Gpt Image 1.5 Edit by OpenAI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • GPT Image 1.5 — Gpt Image 1.5 is OpenAI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • GPT Image 1 Mini Edit — Gpt Image 1 Mini Edit is OpenAI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • GPT Image 1 Mini — Gpt Image 1 Mini by OpenAI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • GPT Image 1 — Gpt Image 1 is OpenAI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • GPT Image 1 Edit — Gpt Image 1 Edit by OpenAI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • GPT Image 2 — GPT Image 2, OpenAI's latest image model, is capable of making fine-grained, detailed edits to images.

Google ​

  • Nano Banana 2 Lite — Nano Banana 2 Lite by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Nano Banana Lite — Nano Banana Lite by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Nano Banana Lite Edit — Nano Banana Lite Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Nano Banana 2 Image Editing — Nano Banana 2 is Google's new state-of-the-art image generation and editing model
  • Nano Banana 2 — Nano Banana 2 is Google's new state-of-the-art fast image generation and editing model
  • Gemini 3.1 Flash Image Preview — Gemini 3.1 Flash Image Preview Edit is Google's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Nano Banana Image Editing — Nano Banana Pro is Google's new state-of-the-art image generation and editing model
  • Imagen 4 — Imagen 4 Preview Fast is Google's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Imagen 4 Ultra — Imagen 4 Preview Ultra by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Imagen 4 — Imagen 4 Preview by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Imagen3 — Imagen 3 by Google - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Imagen3 Fast — Imagen 3 Fast is Google's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • …and 4 more models in the sidebar.

Ideogram ​

  • Ideogram — Custom Models by Ideogram - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

Bytedance ​

  • Seedream 5.0 Pro Image Editing — Seedream 5.0 Pro is Bytedance's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Seedream 5.0 Pro Text to Image — Seedream 5.0 Pro by Bytedance - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • SeedVR2 — Seedvr Upscale Image is Bytedance's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bytedance Seed 2.0 Mini — Seed V2 Mini by Bytedance - advanced AI model for llm. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Seedream v5.0 Lite — Seedream 5.0 Lite — Fast Text-to-Image API The lightweight version of Seedream 5.0, delivering high-quality, low-latency AI image generation from text prompts. Ideal for real-time creative tools, e-commerce visuals, and high-volume AIGC pipelines.
  • SeedVR2 — Seedvr Upscale Video by Bytedance - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • SeedVR2 — Seedvr Upscale Image by Bytedance - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bytedance Seedream v4 Edit — Seedream 4.0 Edit is Bytedance's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bytedance Seedream v4 — Seedream 4.0 by Bytedance - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Dreamina 3.1 — Dreamina 3.1 by Bytedance - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Seedream v4.5 Image Editing — A new-generation image creation model from ByteDance, Seedream 4.5 integrates text-to-image generation and image editing into a single unified architecture, delivering high-fidelity visuals, precise prompt control, and seamless creative workflows for professional AIGC applications.
  • Seedream v5.0 Lite Editing — Seedream 5.0 Lite — Fast Text-to-Image API The lightweight version of Seedream 5.0, delivering high-quality, low-latency AI image generation from text prompts. Ideal for real-time creative tools, e-commerce visuals, and high-volume AIGC pipelines.
  • …and 1 more models in the sidebar.

Recraft ​

  • Recraft V4.1 Utility — Recraft V4.1 Utility by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4.1 Utility Pro — Recraft V4.1 Utility by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4.1 Vector — Recraft V4.1 Vector is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4.1 — Recraft V4.1 by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4.1 Pro Vector — Recraft V4.1 Pro is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4.1 Pro — Recraft V4.1 Pro by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4 Pro (Vector) — Recraft V4 Pro is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4 (Vector) — Recraft V4 Vector is Recraft's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Recraft V4 Pro — Recraft V4 Pro by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft V4 — Recraft V4 by Recraft - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Recraft Vectorize — Recraft Vectorize is Recraft's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Recraft Creative Upscale — Recraft Upscale Creative by Recraft - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • …and 5 more models in the sidebar.

Luma ​

  • Luma Photon Flash Edit — Photon Flash 1 Edit is Luma's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Luma Photon Edit — Photon 1 Edit is Luma's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Luma Photon Flash Reframe — Photon Flash 1 Reframe by Luma - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Luma Photon Reframe — Photon 1 Reframe by Luma - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Luma Photon Flash — Photon Flash 1 by Luma - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Luma Photon — Photon 1 by Luma - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

Z-Image ​

  • Z-Image Turbo — Turbo is Z-Image's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Alibaba ​

  • Qwen Image 3 — Qwen Image 3 by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image 3 Edit — Alibaba Qwen Image 3 image editing model with support for up to three reference images.
  • Wan 2.7 Edit — Alibaba Wan 2.7 image editing model with multi-reference support and configurable output format.
  • Wan 2.7 — Alibaba Wan 2.7 text-to-image model with high-quality generation and configurable output format.
  • Z-Image Turbo Seamless Tiling — Z Image Turbo Tiling is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Z-Image Turbo Seamless Tiling — Z Image Turbo Tiling by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image 2 — Qwen Image 2 Pro by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image 2 — Qwen Image 2 by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image Max — Qwen Image Max by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Qwen Image Max Edit — Qwen Image Max Edit is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Z Image Base (LoRA) — Z Image Base Lora is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Z Image Base — Z Image Base by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • …and 56 more models in the sidebar.

Baidu ​

  • ERNIE-Image Trainer — Ernie Image Trainer by Baidu - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Ernie Image — Ernie Image is Baidu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ernie Image — Ernie Image Turbo is Baidu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

BFL ​

  • FLUX Virtual Try-On — Flux Pro 1.0 Vto by BFL - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Flux Pro Erase — Flux Pro 1.0 Erase by BFL - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Flux 2 Pro — Flux 2 Pro Outpaint is BFL's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • FLUX.2 [klein] 9B LoRA — Flux 2 Klein 9b is BFL's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • FLUX.2 [klein] 9B LoRA — Flux 2 Klein 9b by BFL - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • FLUX.2 [klein] 4B LoRA — Flux 2 Klein 4b is BFL's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • FLUX.2 [klein] 4B LoRA — Flux 2 Klein 4b by BFL - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Flux 2 [klein] Realtime — Flux 2 Klein Realtime is BFL's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • FLUX.2 [klein] 9B Base LoRA — Flux 2 Klein 9b by BFL - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • FLUX.2 [klein] 9B Base LoRA — Flux 2 Klein 9b is BFL's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • FLUX.2 [klein] 4B Base LoRA — Flux 2 Klein 4b by BFL - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • FLUX.2 [klein] 4B Base LoRA — Flux 2 Klein 4b is BFL's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • …and 80 more models in the sidebar.

Bria ​

  • Bria Embed Product — Embed Product is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Upscale Creative — Upscale Creative by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Replace Background — Replace Background by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Fibo Edit — Fibo Edit is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Fibo Lite — Fibo Lite is Bria's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Bria Fibo — Fibo by Bria - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Bria Reimagine 3.2 — Reimagine 3.2 is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Reimagine — Reimagine is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • BRIA RMBG 2.0 — Background Remove is Bria's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Bria Expand Image — Expand by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Eraser — Eraser by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Bria Product Shot — Product Shot by Bria - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • …and 5 more models in the sidebar.

ClarityAI ​

  • Clarity Upscaler — Clarity Upscaler by ClarityAI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

ElevenLabs ​

  • ElevenLabs Speech to Text — Speech To Text by ElevenLabs - accurate speech-to-text transcription with AI. Convert audio and video to text with high accuracy, multilingual support, and speaker identification.

FASHN ​

  • FASHN Virtual Try-On V1.6 — Tryon V1.6 by FASHN - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • FASHN Virtual Try-On V1.5 — Tryon V1.5 by FASHN - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.

HiDream ​

  • HiDream O1 Dev Edit — Hidream O1 Dev Edit is hidream-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • HiDream O1 Edit — Hidream O1 Edit is hidream-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • HiDream O1 — Hidream O1 by hidream-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • HiDream O1 Dev — Hidream O1 Dev by hidream-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • HiDream I1 Full Edit — Hidream I1 Full Edit by hidream-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • HiDream I1 Full — Hidream I1 Full is hidream-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • HiDream I1 Dev — Hidream I1 Dev by hidream-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • HiDream I1 Fast — Hidream I1 Fast is hidream-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • HiDream E1 — Hidream E1 1 by hidream-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • HiDream E1 Full — Hidream E1 Full is hidream-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

hyper3d ​

  • Hyper3d — Rodin V2 by hyper3d - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Hyper3D Rodin — Rodin is hyper3d's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.

Ideogram ​

  • Ideogram — Ideogram Custom Models Generate is ideogram-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ideogram V3 Layerize Text — Ideogram V3 Layerize Text is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram V3 Transparent — Ideogram V3 Transparent is ideogram-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ideogram V3 Character — Ideogram V3 Character is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram V3 Character Remix — Ideogram V3 Character Remix is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram V3 Reframe — Ideogram V3 Reframe is ideogram-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Ideogram 3.0 — 3.0 is Ideogram's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Ideogram V3 Replace Background — Ideogram V3 Replace Background by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram V3 Edit — Ideogram V3 Edit by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram V2A Turbo — Ideogram V2a Turbo by ideogram-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Ideogram V2A Remix — Ideogram V2a Remix by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Ideogram V2A Turbo Remix — Ideogram V2a Turbo Remix by ideogram-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • …and 11 more models in the sidebar.

ImagineArt ​

  • Imagineart 2.0 Edit Preview — Imagineart 2.0 Edit Preview by ImagineArt - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Imagineart 2.0 Preview — Imagineart 2.0 Preview by ImagineArt - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

KwaiVGI ​

  • Kling Video V3 Standard — Kling Video V3 Standard by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Kling Video V3 Pro — Kling Video V3 Pro is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Kling Video O3 Pro — Kling Video O3 Pro is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Kling Video O3 Standard — Kling Video O3 Standard by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Kling Video O3 Standard — Kling Video O3 Standard is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Kling Image V3 — Kling Image V3 by KwaiVGI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Kling Image — Kling Image O3 Edit is KwaiVGI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Kling Image — Kling Image O3 by KwaiVGI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Kling Video V2.6 Standard — Kling Video V2.6 Standard by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Kling Video V2.6 Pro — Kling Video V2.6 Pro is KwaiVGI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Kling Video Create Voice — Kling Video Create Voice by KwaiVGI - advanced AI model for audio-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Kling Video O1 Standard — Kling Video O1 Standard by KwaiVGI - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • …and 9 more models in the sidebar.

Lightricks ​

  • LTX-2.3 22B Video to Video Trainer — Ltx23 V2v Trainer by Lightricks - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2.3 22B Video Trainer — Ltx23 Video Trainer by Lightricks - advanced AI model for training. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2 19B Distilled — Ltx 2 19b Distilled Audio To Video by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2 19B — Ltx 2 19b Audio To Video by Lightricks - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • LTX-2 19B — Ltx 2 19b Video To Video by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • LTX-2 19B — Ltx 2 19b Extend Video by Lightricks - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

meituan ​

  • Longcat Multi Avatar — Longcat Multi Avatar Image Audio To Video by meituan - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Longcat Single Avatar — Longcat Single Avatar Audio To Video by meituan - advanced AI model for audio-to-video. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Longcat Image — Longcat Image Edit by sandbase-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Longcat Image — Longcat Image is meituan's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Meshy ​

  • Meshy Rigging — Rigging by Meshy - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Meshy 6 - Multi Image To 3D — Meshy V6 Multi Image To 3d by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Meshy 6 — Meshy V6 by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Meshy 6 — Meshy V6 is Meshy's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Meshy 5 Retexture — Meshy V5 Retexture by Meshy - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Meshy 5 Remesh — Meshy V5 Remesh by Meshy - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Meshy 6 Preview — Meshy V6 Preview is Meshy's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Meshy 5 Multi — Meshy V5 Multi Image To 3d by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Meshy 6 Preview — Meshy V6 Preview by Meshy - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.

Meta ​

  • Sam 3 1 — Sam 3 1 Video is Meta's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Sam 3 1 — Sam 3 1 Image Rle is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Sam 3 1 — Sam 3 1 Image is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • SAM 3 3D Align — Sam 3 3d Align by Meta - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Sam 3 — Sam 3 3d Body is Meta's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Sam 3 — Sam 3 3d Objects by Meta - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Sam 3 — Sam 3 Image Rle is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • SAM 3 Embed — Sam 3 Image Embed by Meta - advanced AI model for vision. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Sam 3 — Sam 3 Video Rle is Meta's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • SAM 3 Image — Sam 3 Image is Meta's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Demucs — Demucs by Meta - advanced AI model for audio-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

MiniMax ​

  • Minimax Image Subject Reference — Image 01 Subject Reference by MiniMax - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • MiniMax (Hailuo AI) Text to Image — Image 01 by MiniMax - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

Mirelo ​

  • Mirelo SFX1.6 — Sfx1.6 Video To Video is Mirelo's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Mirelo SFX1.6 — Sfx1.6 Inpaint Audio by Mirelo - advanced AI model for audio-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Mirelo SFX1.6 — Sfx1.6 Extend Audio by Mirelo - advanced AI model for audio-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

NVIDIA ​

  • Nemotron 3 Nano Omni — Nemotron 3 Nano Omni Vision by NVIDIA - advanced AI model for image-to-text. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Nemotron 3 Nano Omni — Nemotron 3 Nano Omni Video by NVIDIA - advanced AI model for video-to-text. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Nemotron 3 Nano Omni — Nemotron 3 Nano Omni Audio by NVIDIA - advanced AI model for audio-to-text. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

patina ​

  • PATINA — Material Extract by patina - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • PATINA — Material by patina - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

pixal3d ​

  • Pixal3d — Pixal3d is pixal3d's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.

PixVerse ​

  • PixVerse V6 Extend — V6 Extend is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

sandbase-ai ​

  • Phota Text to Image — Phota is sandbase-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Hy Wu Edit — Hy Wu Edit by sandbase-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Firered Image Edit V1.1 — Firered Image Edit V1.1 is sandbase-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • try-on — Cat Vton by sandbase-ai - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Sana — Sana by sandbase-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Creative Upscaler — Creative Upscaler is sandbase-ai's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

Stability AI ​

  • Stable Diffusion 3.5 Large — Sd 3.5 Large by stability-ai - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Stable Diffusion 3.5 Medium — Sd 3.5 Medium is stability-ai's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Stable Diffusion V3 — Stable Diffusion V3 Medium by Stability AI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • SDXL ControlNet Union — Sdxl Controlnet Union by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • SDXL ControlNet Union — Sdxl Controlnet Union is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • SDXL ControlNet Union — Sdxl Controlnet Union Inpainting by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Stable Cascade — Stable Cascade by Stability AI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Stable Diffusion XL — Fast Sdxl is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Stable Diffusion V3 — Stable Diffusion V3 Medium is Stability AI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • SoteDiffusion — Stable Cascade Sote Diffusion is Stability AI's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Stable Diffusion XL — Fast Sdxl by Stability AI - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Stable Diffusion v1.5 — Stable Diffusion V15 by Stability AI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • …and 3 more models in the sidebar.

Sync Labs ​

  • sync-3 Lipsync — Lipsync V3 by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Sync React-1 — Lipsync React 1 is Sync Labs's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Sync Lipsync — Lipsync V2 Pro by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • Sync Lipsync 2.0 — Lipsync V2 by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.
  • sync.so -- lipsync 1.9.0-beta — Sync Lipsync by Sync Labs - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

Tencent ​

  • Hunyuan 3D 3.1 Rapid Text to 3D — Hunyuan 3d 3.1 Rapid is Tencent's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Hunyuan Image 3.0 Edit — Hunyuan Image 3.0 Edit by Tencent - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Hunyuan Image 3.0 Instruct — Hunyuan Image 3.0 Instruct by Tencent - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Hunyuan 3D Smart Topology — Hunyuan 3d 3.1 Smart Topology by Tencent - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Hunyuan 3D 3.1 Rapid Image to 3D — Hunyuan 3d 3.1 Rapid by Tencent - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Hunyuan 3D 3.1 Pro Text to 3D — Hunyuan 3d 3.1 Pro is Tencent's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Hunyuan 3D 3.1 Pro Image to 3D — Hunyuan 3d 3.1 Pro by Tencent - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Hunyuan 3D Part Splitter — Hunyuan 3d 3.1 Part by Tencent - advanced AI model for 3d-to-3d. Delivers high-quality results with fast inference, suitable for both creative and production workflows.
  • Hunyuan Motion Fast — Hunyuan Motion Fast is Tencent's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Hunyuan Motion — Hunyuan Motion by Tencent - generate 3D models from text descriptions with AI. Create detailed 3D assets from natural language for games, visualization, and digital production.
  • Hunyuan3d V3 — Hunyuan 3d V3 by Tencent - generate 3D models from text descriptions with AI. Create detailed 3D assets from natural language for games, visualization, and digital production.
  • Hunyuan3d V3 — Hunyuan 3d V3 Sketch To 3d by Tencent - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • …and 11 more models in the sidebar.

Topaz Labs ​

  • Topaz Upscale — Upscale Image is Topaz Labs's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

trellis ​

  • Trellis 2 — Trellis 2 Retexture is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Trellis 2 — Trellis 2 is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Trellis — Trellis Multi is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.
  • Trellis — Trellis is trellis's image-to-3D AI model. Transform photographs into production-ready 3D meshes with accurate geometry and texture mapping.

Tripo3D ​

  • Tripo H3.1 Multiview to 3D — Tripo H3.1 Multiview To 3d by Tripo3D - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Tripo H3.1 Text to 3D — Tripo H3.1 is Tripo3D's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Tripo H3.1 Image to 3D — Tripo H3.1 by Tripo3D - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.
  • Tripo P1 Text to 3D — Tripo P1 is Tripo3D's text-to-3D AI model. Turn written descriptions into textured 3D objects with realistic geometry and materials.
  • Tripo P1 Image to 3D — Tripo P1 by Tripo3D - convert 2D images into 3D models with AI. Generate textured 3D assets from single photos for games, AR/VR, e-commerce, and digital content creation.

VEED ​

  • Subtitles — Subtitles is VEED's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

Vidu ​

  • Vidu Q2 Reference to Image — Vidu Q2 Reference To Image by Shengshu - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Vidu Q2 Text to Image — Vidu Q2 is Shengshu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.
  • Vidu 2.0 Reference to Image — Vidu 2.0 Reference To Image is Shengshu's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.

xAI ​

  • Grok Imagine Image Quality — Grok Imagine Image Quality by xAI - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.
  • Grok Imagine Image Quality Edit — Grok Imagine Image Quality Edit is xAI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Grok Imagine Video Extend — Grok Imagine Video Extend is xAI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.
  • Grok Imagine Image Edit — Grok Imagine Image Edit is xAI's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to artistic style conversion.
  • Grok Imagine Video Edit — Grok Imagine Video Edit is xAI's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

zhipu ​

  • Glm Image — Glm Image Edit by zhipu - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images effortlessly.
  • Glm Image — Glm Image is zhipu's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt adherence.

Capability coverage ​

3d-to-3d, audio-to-audio, audio-to-text, audio-to-video, commercial, image-editing, image-to-3d, image-to-image, image-to-text, llm, speech-to-text, text-to-3d, text-to-image, training, video-to-text, video-to-video, vision