Gemini 3.5 Flash Lite
google/gemini-3.5-flash-liteGemini 3.5 Flash Lite is a high-efficiency multimodal model from Google with upgraded agentic capabilities. It is suited for focused subagent tasks in complex multi-agent workflows.
- Input price
- $0.30USD / 1M tokens
- Output price
- $2.50USD / 1M tokens
- Context window
- N/A
- Max output
- 65.5K
Try the model
Playground
Input⌘ / Ctrl + Enter
Try a prompt
OutputReady
Your response will appear here
Choose an example or write a prompt, then click Run.
Specifications
Pricing
- Input
- $0.30 / 1M tokens
- Output
- $2.50 / 1M tokens
- Billing formula
- usage.prompt_tokens * 0.3 / 1000000 + usage.completion_tokens * 2.5 / 1000000
Context & modalities
- Max output
- 65,536 tokens
- Input
- Text, image
- Output
- Text
Capabilities
- Chat
- Supported
- Vision
- Supported
- Reasoning
- Supported
- Structured output
- Supported
- Function calling
- Supported
- Audio input
- Supported
Access
- Provider
- Model ID
- google/gemini-3.5-flash-lite
- Execution
- sync
- API
- Chat Completions API
- Endpoint
- /v1/chat/completions
Start building
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
-X POST "https://api.sandbase.ai/v1/chat/completions" \
-H "Authorization: Bearer $SANDBASE_API_KEY" \
-H "Content-Type: application/json" \
--data-binary @- <<'SANDBASE_JSON'
{
"model": "google/gemini-3.5-flash-lite",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"max_tokens": 2048
}
SANDBASE_JSON
)
printf '%s\n' "$result"Choose your model
Compare models
| Model | Context | Input / 1M | Output / 1M | Released |
|---|---|---|---|---|
Google: Gemini 3.5 Flash LiteThis model Google | — | $0.30 | $2.50 | Jul 21, 2026 |
Google | 1M | — | — | Sep 2, 2026 |
Google | — | — | — | Aug 27, 2026 |
Google | 1M | — | — | Aug 13, 2026 |
Google | 1M | — | — | Jul 21, 2026 |
Google | — | — | — | Jun 30, 2026 |
