Inclusionai modelsChat Completions model

inclusionai/ling-2.6-flash

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficienc...

Input⌘ / Ctrl + Enter

Try a prompt

32.8K max output
Input $0.08/M · Output $0.24/M · Cache read $0.02/M
OutputReady

Your response will appear here

Choose an example or write a prompt, then click Run.

API details and access

Model Details

ProviderInclusionai
TypeLlm
Model IDinclusionai/ling-2.6-flash

Capabilities

InputText
OutputText
Context-
Max Output32,768
VisionNot supported
Function CallingSupported

Access details

Chat Completions
Base URLhttps://api.sandbase.ai
API Endpoint/v1/chat/completions
{
  "model": "inclusionai/ling-2.6-flash",
  "messages": [
    {
      "role": "user",
      "content": "Hello"
    }
  ],
  "max_tokens": 512
}

Pricing

Input$0.08 / 1M tokens
Output$0.24 / 1M tokens
Cache read$0.02 / 1M tokens

Related Models