API reference · Nousresearch
nousresearch/hermes-3-llama-3.1-70b
Integrate this model through SandBase's unified API, with production-ready schemas and examples.
Production endpoint
Send your first request
OpenAI-compatible endpoint with unified authentication and usage tracking.
POST
https://api.sandbase.ai/v1/chat/completionsModel ID
nousresearch/hermes-3-llama-3.1-70b01
Input Schema
11 parameters · 1 required · 10 optional
| Parameter | Type | Required | Description |
|---|---|---|---|
messages | object[] | Required | — |
seed | integer | Optional | — |
stop | string[] | Optional | — |
top_k | integer | Optional | Min: 0 |
top_p | number | Optional | Min: 0 · Max: 1 |
stream | boolean | Optional | Default: false |
max_tokens | integer | Optional | Max: 16384 |
temperature | number | Optional | Min: 0 · Max: 2 · Default: 1 |
response_format | object | Optional | — |
presence_penalty | number | Optional | Min: -2 · Max: 2 |
frequency_penalty | number | Optional | Min: -2 · Max: 2 |
02
Output Schema
| Field | Type | Description |
|---|---|---|
id | string | Unique identifier for the completion |
object | string | Object type, e.g. "chat.completion" |
model | string | Model used for the completion |
choices | array | List of completion choices |
choices[].message.content | string | Generated text content |
choices[].finish_reason | string | Reason the generation stopped |
usage.prompt_tokens | integer | Number of tokens in the prompt |
usage.completion_tokens | integer | Number of tokens in the completion |
usage.total_tokens | integer | Total tokens used |
03
Code Examples
Ready-to-run snippets
curl https://api.sandbase.ai/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "nousresearch/hermes-3-llama-3.1-70b",
"messages": [{"role": "user", "content": "Hello"}],
"temperature": 0.7,
"max_tokens": 2048,
"stream": true
}'
