Skip to content

Qwen 3 TTS - Clone Voice [1.7B]

POST/v1/run

Qwen 3 Tts Clone Voice 1.7b by Alibaba - advanced AI model for audio-to-audio. Delivers high-quality results with fast inference, suitable for both creative and production workflows.

Request body

Submit an async generation request. The model field selects the model; other fields are model-specific input parameters.

stringmodelrequired

Model identifier. Set to alibaba/qwen-3-tts/clone-voice/1.7b.

Default: alibaba/qwen-3-tts/clone-voice/1.7b

Optional<string>audio

URL to the reference audio file used for voice cloning.

Optional<string>reference_text

Optional reference text that was used when creating the speaker embedding. Providing this can improve synthesis quality when using a cloned voice.

Response Schema

The submit endpoint returns an accepted generation task. Poll the result endpoint with the returned id for terminal outputs or errors.

Optional<string>error

Error message if the task failed. Empty on success.

stringidrequired

Unique identifier for the generation task.

Optional<string>model

Model ID used for the prediction.

Optional<array>outputs

Array of generated content. Empty when status is not completed.

stringstatusrequired

Status of the task: pending, running, completed, failed, or timeout.

Allowed values: pending, running, completed, failed, timeout

Model capabilities

array<string>capability_tagsrequired

Capabilities declared by the model registry.

Default: audio-to-audio

stringexecution_moderequired

Execution mode declared by the model registry.

Default: async