API reference · MiniMax

minimax/speech/2.8/hd

Integrate this model through SandBase's unified API, with production-ready schemas and examples.

AUDIOAsyncOpen model
Production endpoint

Send your first request

OpenAI-compatible endpoint with unified authentication and usage tracking.

POSThttps://api.sandbase.ai/v1/run
Model IDminimax/speech/2.8/hd
01

Input Schema

12 parameters · 2 required · 10 optional

ParameterTypeRequiredDescription
textstringRequiredText to convert to speech. Use <#x#> between words to control pause duration (0.01-99.99s). · Min length: 1 · Max length: 5000 · Default: "Hello, welcome to Sandbase"
voice_idstringRequiredVoice ID for speech synthesis. Use a predefined system voice or a custom cloned voice ID. · Default: "Wise_Woman"
pitchintegerOptionalSpeech pitch. Range: -12 to 12, where 0 is normal pitch. · Min: -12 · Max: 12 · Default: 0
speednumberOptionalSpeech speed. Range: 0.5-2.0, where 1.0 is normal speed. · Min: 0.5 · Max: 2 · Default: 1
formatstringOptionalFormat of generated sound. · Options: mp3, pcm, flac
mp3pcmflac
volumenumberOptionalSpeech volume. Range: 0.1-10.0, where 1.0 is normal volume. · Min: 0.1 · Max: 10 · Default: 1
bitrateintegerOptionalBitrate of generated sound. · Options: 32000, 64000, 128000, 256000
3200064000128000256000
channelintegerOptionalThe number of channels. 1: mono, 2: stereo. · Options: 1, 2
12
emotionstringOptionalThe emotion of the generated speech. · Options: happy, sad, angry, fearful, disgusted, surprised, neutral · Default: "happy"
happysadangryfearfuldisgustedsurprisedneutral
sample_rateintegerOptionalSample rate of generated sound. · Options: 8000, 16000, 22050, 24000, 32000, 44100
80001600022050240003200044100
language_booststringOptionalEnhance the ability to recognize specified languages and dialects. · Options: Chinese, Chinese,Yue, English, Arabic, Russian, Spanish, French, Portuguese, German, Turkish, Dutch, Ukrainian, Vietnamese, Indonesian, Japanese, Italian, Korean, Thai, Polish, Romanian, Greek, Czech, Finnish, Hindi, Bulgarian, Danish, Hebrew, Malay, Slovak, Swedish, Croatian, Hungarian, Norwegian, Slovenian, Catalan, Nynorsk, Afrikaans, auto
ChineseChinese,YueEnglishArabicRussianSpanishFrenchPortugueseGermanTurkishDutchUkrainianVietnameseIndonesianJapaneseItalianKoreanThaiPolishRomanianGreekCzechFinnishHindiBulgarianDanishHebrewMalaySlovakSwedishCroatianHungarianNorwegianSlovenianCatalanNynorskAfrikaansauto
english_normalizationbooleanOptionalImproves performance in number-reading scenarios. · Default: false
02

Output Schema

FieldTypeDescription
idstringUnique identifier for the generation task
statusstringTask status: pending, running, completed, failed, timeout
modelstringModel used for the generation
outputsarrayArray of output items
outputs[].urlstringURL of the generated artifact
outputs[].content_typestringMIME type (e.g. image/png, video/mp4)
errorobject | nullError details if failed, null on success
error.typestringMachine-readable error type code
error.messagestringHuman-readable error description

Async Workflow

This model uses asynchronous execution. Submit a request and poll for the result.

  1. Submit — POST to /v1/run, receive an id
  2. Poll — GET /v1/run/{id} until status is completed, failed, or timeout
  3. Retrieve — Read outputs from the completed response
03

Code Examples

Ready-to-run snippets

# Step 1: Submit
curl -X POST https://api.sandbase.ai/v1/run \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
  "model": "minimax/speech/2.8/hd",
  "text": "Hello, welcome to Sandbase",
  "pitch": 0,
  "speed": 1,
  "volume": 1,
  "emotion": "happy",
  "voice_id": "Wise_Woman",
  "english_normalization": false,
  "prompt": "a beautiful sunset over mountains"
}'

# Step 2: Poll result (replace <id>)
curl https://api.sandbase.ai/v1/run/<id> \
  -H "Authorization: Bearer YOUR_API_KEY"