Qwen: Qwen3.8 Omni Flash
alibaba/qwen3.8-omni-flashQwen3.8 Omni Flash is Alibaba's omni-modal reasoning model in the Qwen3.8 family, accepting text, image, audio and video input and returning text, with a 1M-token context window. Alibaba positions it for audio and video understanding and content analysis, including transcription-aware summarization, meeting and lecture review, and multimodal agent steps, with tool calling and adjustable reasoning effort. Pick Qwen3.8 Flash for text and image workloads that do not need audio or video input.
- Input price
- $0.15USD / 1M tokens
- Output price
- $0.47USD / 1M tokens
- Context window
- 1M
- Max output
- 131.1K
Try the model
Playground
Try a prompt
Your response will appear here
Choose an example or write a prompt, then click Run.
Model card
Specifications
Pricing
- Input
- $0.15 / 1M tokens
- Output
- $0.47 / 1M tokens
- Cache read
- $0.02 / 1M tokens
- Cache write
- $0.15 / 1M tokens
Context & modalities
- Context window
- 1,000,000 tokens
- Max output
- 131,072 tokens
- Input
- Text, image
- Output
- Text
Capabilities
- Chat
- Supported
- Vision
- Supported
- Reasoning
- Supported
- Structured output
- Supported
- Function calling
- Supported
- Audio input
- Supported
Access
- Provider
- Alibaba
- Model ID
- alibaba/qwen3.8-omni-flash
- Execution
- sync
- Base URL
- https://api.sandbase.ai
- API
- Chat Completions API
- Endpoint
- /v1/chat/completions
Start building
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
-X POST "https://api.sandbase.ai/v1/chat/completions" \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
-H "Content-Type: application/json" \
--data-binary @- <<'SANDBASE_JSON'
{
"model": "alibaba/qwen3.8-omni-flash",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"max_tokens": 2048
}
SANDBASE_JSON
)
printf '%s\n' "$result"Choose your model
Compare models
| Model | Context | Input / 1M | Output / 1M | Released |
|---|---|---|---|---|
Qwen3.8 Omni FlashThis model Alibaba | 1M | $0.15 | $0.47 | Sep 18, 2026 |
Alibaba | 1M | $4.00 | $12.00 | Sep 23, 2026 |
Alibaba | 1M | $2.00 | $6.00 | Sep 3, 2026 |
Alibaba | 1M | $0.50 | $3.00 | Aug 14, 2026 |
Alibaba | 1M | $2.00 | $6.00 | Aug 12, 2026 |
Alibaba | 1M | $2.00 | $6.00 | Aug 3, 2026 |
Questions
FAQ
How do I call Alibaba Qwen3.8 Omni Flash through SandBase?
Create a SandBase API key, then send requests with the model ID "alibaba/qwen3.8-omni-flash" to Chat Completions API (/v1/chat/completions) at https://api.sandbase.ai. The request examples on this page show the exact payload.
How much does Alibaba Qwen3.8 Omni Flash cost?
$0.15 per 1M input tokens and $0.47 per 1M output tokens, billed per request from your SandBase balance. Discounts, when available, are shown in the pricing on this page.
What is the context window of Alibaba Qwen3.8 Omni Flash?
1M tokens of context, with up to 131.1K output tokens per response.
Which features does Alibaba Qwen3.8 Omni Flash support?
Alibaba Qwen3.8 Omni Flash supports vision (image input), reasoning, function calling, structured output. See Specifications above for the full list.
Do I need a separate Alibaba account?
No. One SandBase API key and balance gives you access to Alibaba Qwen3.8 Omni Flash and the other models in the catalog; you do not need to sign up with Alibaba separately.
