Skip to content

GPT-5 Chat

POST/v1/chat/completions

GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.

Request body

Parameters supported by this model. Values, defaults, and limits are read from the model registry.

stringmodelrequired

Model identifier. Set to openai/gpt-5-chat.

Default: openai/gpt-5-chat

array<object>messagesrequired

Conversation messages in system, user, or assistant order.

Optional<integer>max_tokens

Maximum number of tokens the model may generate in the response.

Range: −∞ to 16384

Optional<number>temperature

Sampling temperature. Lower values are more deterministic; higher values are more creative.

Range: 0 to 2

Default: 1

Optional<boolean>stream

When true, returns incremental Server-Sent Events instead of one completed response.

Default: false

Optional<object>response_format

Controls the response format, including JSON mode or structured JSON output when supported.

Optional<integer>seed

Seed used to make sampling more reproducible when the provider supports it.

Response Schema

Fields returned by this model API response.

array<object>choicesrequired

Generated completion choices.

stringidrequired

Unique chat completion identifier.

stringmodelrequired

Model that generated the response.

Optional<object>usage

Token usage when available.

Model capabilities

array<string>capability_tagsrequired

Capabilities declared by the model registry.

Default: chat, vision, structured_output

integercontext_lengthrequired

Maximum context window accepted by this model.

Default: 128000 tokens

stringexecution_moderequired

Execution mode declared by the model registry.

Default: sync