SandBase is live — $1 in free credits on signupStart free ›

MODEL PROVIDER

Alibaba

Explore 190 models and APIs from Alibaba, available through one SandBase integration.

Browse all models
190models available
01

Large Language Models (LLMs)

16 on this page

LLM

alibaba/qwen3-32b

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode ...

$0.08 / $0.28 per 1M tokens131.1K context
LLM

alibaba/qwen3-vl-235b-a22b-instruct

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language...

$0.20 / $0.88 per 1M tokens262.1K context
LLM

alibaba/qwen2.5-vl-72b-instruct

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

$0.25 / $0.75 per 1M tokens131.1K context
LLM

alibaba/qwen3-vl-30b-a3b-instruct

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general mu...

$0.13 / $0.52 per 1M tokens262.1K context
LLM

alibaba/qwen3-vl-8b-instruct

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal...

$0.08 / $0.50 per 1M tokens256K context
LLM

alibaba/qwen3.7-max

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivi...

$1.25 / $3.75 per 1M tokens1M context
LLM

alibaba/qwen3.7-plus

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

$0.40 / $1.60 per 1M tokens1M context
LLM

alibaba/qwen3.6-plus

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to t...

$0.33 / $1.95 per 1M tokens1M context
LLM

alibaba/qwen3-8b

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode f...

$0.05 / $0.40 per 1M tokens131.1K context
LLM

alibaba/qwen3-30b-a3b-instruct-2507

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quali...

$0.05 / $0.19 per 1M tokens131.1K context
LLM

alibaba/qwen3-vl-32b-instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines ...

$0.10 / $0.42 per 1M tokens262.1K context
LLM

alibaba/qwen-plus

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

$0.26 / $0.78 per 1M tokens1M context
LLM

alibaba/qwen3.8-2.4t-a95b

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen, with 95 billion active parameters out of 2.4 trillion total. It is the open-weight variant of Qwen3.8 Max.

$2.00 / $6.00 per 1M tokens262.1K context
LLM

alibaba/qwen-2.5-72b-instruct

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

$0.36 / $0.40 per 1M tokens131.1K context
LLM

alibaba/qwen3-235b-a22b-thinking-2507

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass...

$0.10 / $0.10 per 1M tokens262.1K context
LLM

alibaba/qwen3.6-35b-a3b

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architectur...

$0.14 / $1.00 per 1M tokens262.1K context
02

Image Models

16 on this page

IMAGE

alibaba/wan/2.2/5b/text-to-image

Wan 2.2 5b by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

$0.0160/call
IMAGE

alibaba/z-image/turbo/tiling/lora

Z Image Turbo Tiling is Alibaba's advanced text-to-image AI model. Create photorealistic images, illustrations, and concept art from natural language descriptions with exceptional detail and prompt ad...

$0.0250/call
IMAGE

alibaba/qwen-image-edit-plus/lora-gallery/remove-lighting

Qwen Image Edit Plus Lora Gallery Remove Lighting is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to ar...

$0.0350/call
IMAGE

alibaba/qwen-image-edit/lora

Qwen Image Edit Lora by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images ef...

$0.0350/call
IMAGE

alibaba/qwen-image-edit-2509-lora-gallery/remove-element

Qwen Image Edit 2509 Lora Gallery Remove Element by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change s...

$0.0350/call
IMAGE

alibaba/qwen-image-edit-plus/lora-gallery/lighting-restoration

Qwen Image Edit Plus Lora Gallery Lighting Restoration by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, ch...

$0.0350/call
IMAGE

alibaba/wan-vace

Wan Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.2000/call
IMAGE

alibaba/qwen-image-edit-plus/lora-gallery/next-scene

Qwen Image Edit Plus Lora Gallery Next Scene by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change style...

$0.0350/call
IMAGE

alibaba/qwen-image-edit/2509-lora-gallery/face-to-full-portrait

Qwen Image Edit 2509 Lora Gallery Face To Full Portrait is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement...

$0.0350/call
IMAGE

alibaba/z-image/turbo/inpaint/lora

Z Image Turbo Inpaint by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change styles, and enhance images e...

$0.0115/call
IMAGE

alibaba/qwen-image-edit/2509-lora-gallery/add-background

Qwen Image Edit 2509 Lora Gallery Add Background by Alibaba - AI-powered image editing, style transfer, and transformation. Edit photos with natural language instructions, remove backgrounds, change s...

$0.0350/call
IMAGE

alibaba/qwen-image-edit/2509-lora-gallery/integrate-product

Qwen Image Edit 2509 Lora Gallery Integrate Product is Alibaba's intelligent image editing model. Transform, retouch, and reimagine existing images using text prompts - from background replacement to ...

$0.0350/call
IMAGE

alibaba/qwen-image-3

Qwen Image 3 by Alibaba - generate stunning images from text prompts with state-of-the-art AI. Supports multiple aspect ratios, styles, and high-resolution output for creative and commercial use.

$0.0420/call
IMAGE

alibaba/face-swap

Face Swap replaces the face in a target image with the face from a source image, producing a realistic face-swapped result.

$0.0130/call
IMAGE

alibaba/head-swap

Head Swap replaces the head in a target image with the head from a source image, producing a realistic head-swapped result.

$0.0130/call
IMAGE

alibaba/qwen-image-3/edit

Alibaba Qwen Image 3 image editing model with support for up to three reference images.

$0.0452/call
03

Video Models

17 on this page

VIDEO

alibaba/wan/2.2/vace-fun/reframe

Wan 2.2 Vace Fun by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.1000/call
VIDEO

alibaba/wan/2.2/vace-fun/outpainting

Wan 2.2 Vace Fun by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.1000/call
VIDEO

alibaba/wan/2.1/image-to-video/lora

Wan 2.1 Lora is Alibaba's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

$0.7500/call
VIDEO

alibaba/wan/vision-enhancer

Wan Vision Enhancer is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.3000/call
VIDEO

alibaba/wan/2.1/vace/depth

Wan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.6400/call
VIDEO

alibaba/wan/2.2/fun-control

Wan 2.2 Fun Control is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.1000/call
VIDEO

alibaba/wan/alpha

Wan Alpha is Alibaba's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

$0.0400/call
VIDEO

alibaba/wan/2.2/vace-fun/depth

Wan 2.2 Vace Fun by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.1000/call
VIDEO

alibaba/wan/move

Wan Move by Alibaba - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

$0.2000/call
VIDEO

alibaba/wan/2.1/text-to-video/lora

Wan 2.1 Lora by Alibaba - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.

$0.7500/call
VIDEO

alibaba/wan/2.1/vace/pose

Wan 2.1 Vace is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.6400/call
VIDEO

alibaba/wan/2.1/vace

Wan 2.1 Vace by Alibaba - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.

$0.6400/call
VIDEO

alibaba/wan/2.2/vace-fun/inpainting

Wan 2.2 Vace Fun is Alibaba's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

$0.1000/call
VIDEO

alibaba/wan/ati

Wan Ati is Alibaba's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.

$0.1500/call
VIDEO

alibaba/wan/v2.2-5b/text-to-video/distill

Wan V2.2 5b Distill is Alibaba's text-to-video AI model. Turn written scripts and prompts into professional-quality video clips with realistic motion, lighting, and scene composition.

$0.0800/call
VIDEO

alibaba/wan-animate

Wan-Animate generates an animated video from an input image and a driving video, transferring motion onto the image subject.

$0.0530/call
VIDEO

alibaba/flashvsr

FlashVSR is a video super-resolution model that upscales videos to higher resolutions (720p / 1080p / 2K / 4K) with fast inference.

$0.0160/call
04

Audio Models

1 on this page

EXPLORE

Other model providers

View all providers ↗