SandBase is live — $1 in free credits on signupStart free ›

Inception modelsChat Completions model

inception/mercury-2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving >1,000 tokens/s...

Input⌘ / Ctrl + Enter

Try a prompt

128K context50K max output
Input $0.25/M · Output $0.75/M · Cache read $0.03/M
OutputReady

Your response will appear here

Choose an example or write a prompt, then click Run.

API details and access

Model Details

ProviderInception
TypeLlm
Model IDinception/mercury-2

Capabilities

InputText
OutputText
Context128,000
Max Output50,000
VisionNot supported
Function CallingSupported

Access details

Chat Completions
Base URLhttps://api.sandbase.ai
API Endpoint/v1/chat/completions
{
  "model": "inception/mercury-2",
  "messages": [
    {
      "role": "user",
      "content": "Hello"
    }
  ],
  "max_tokens": 512
}

Pricing

Input$0.25 / 1M tokens
Output$0.75 / 1M tokens
Cache read$0.03 / 1M tokens