API Free WeekHundreds of APIs are free to call this week.Browse free APIs

Qwen: Qwen3.8 Omni Flash

alibaba/qwen3.8-omni-flash

Qwen3.8 Omni Flash is Alibaba's omni-modal reasoning model in the Qwen3.8 family, accepting text, image, audio and video input and returning text, with a 1M-token context window. Alibaba positions it for audio and video understanding and content analysis, including transcription-aware summarization, meeting and lecture review, and multimodal agent steps, with tool calling and adjustable reasoning effort. Pick Qwen3.8 Flash for text and image workloads that do not need audio or video input.

Input price
$0.15USD / 1M tokens
Output price
$0.47USD / 1M tokens
Context window
1M
Max output
131.1K
Production route

Try the model

Playground

Open playground
Input⌘ / Ctrl + Enter

Try a prompt

1M context131.1K max outputVision
Input$0.15/MOutput$0.47/MCache read$0.02/MCache write$0.15/M
OutputReady

Your response will appear here

Choose an example or write a prompt, then click Run.

Model card

Specifications

Pricing

Input
$0.15 / 1M tokens
Output
$0.47 / 1M tokens
Cache read
$0.02 / 1M tokens
Cache write
$0.15 / 1M tokens

Context & modalities

Context window
1,000,000 tokens
Max output
131,072 tokens
Input
Text, image
Output
Text

Capabilities

Chat
Supported
Vision
Supported
Reasoning
Supported
Structured output
Supported
Function calling
Supported
Audio input
Supported

Access

Provider
Alibaba
Model ID
alibaba/qwen3.8-omni-flash
Execution
sync
Base URL
https://api.sandbase.ai
API
Chat Completions API
Endpoint
/v1/chat/completions

Start building

Send your first request

OpenAI-compatible endpoint with unified authentication and usage tracking.

Production API
Chat Completions API endpoint
https://api.sandbase.ai/v1/chat/completions
Model ID
alibaba/qwen3.8-omni-flash
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
  -X POST "https://api.sandbase.ai/v1/chat/completions" \
  -H "Authorization: Bearer $SANDBASE_API_KEY" \
  -H "Content-Type: application/json" \
  --data-binary @- <<'SANDBASE_JSON'
{
  "model": "alibaba/qwen3.8-omni-flash",
  "messages": [
    {
      "role": "user",
      "content": "Hello"
    }
  ],
  "max_tokens": 2048
}
SANDBASE_JSON
)
printf '%s\n' "$result"

Choose your model

Compare models

All language models
ModelContextInput / 1MOutput / 1MReleased
Alibaba1M$0.15$0.47Sep 18, 2026
Alibaba1M$4.00$12.00Sep 23, 2026
Alibaba1M$2.00$6.00Sep 3, 2026
Alibaba1M$0.50$3.00Aug 14, 2026
Alibaba1M$2.00$6.00Aug 12, 2026
Alibaba1M$2.00$6.00Aug 3, 2026

Questions

FAQ

How do I call Alibaba Qwen3.8 Omni Flash through SandBase?

Create a SandBase API key, then send requests with the model ID "alibaba/qwen3.8-omni-flash" to Chat Completions API (/v1/chat/completions) at https://api.sandbase.ai. The request examples on this page show the exact payload.

How much does Alibaba Qwen3.8 Omni Flash cost?

$0.15 per 1M input tokens and $0.47 per 1M output tokens, billed per request from your SandBase balance. Discounts, when available, are shown in the pricing on this page.

What is the context window of Alibaba Qwen3.8 Omni Flash?

1M tokens of context, with up to 131.1K output tokens per response.

Which features does Alibaba Qwen3.8 Omni Flash support?

Alibaba Qwen3.8 Omni Flash supports vision (image input), reasoning, function calling, structured output. See Specifications above for the full list.

Do I need a separate Alibaba account?

No. One SandBase API key and balance gives you access to Alibaba Qwen3.8 Omni Flash and the other models in the catalog; you do not need to sign up with Alibaba separately.