Skip to content

GPT-5 Pro

POST/v1/responses

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy in high-stakes use cases. It supports test-time routing features and advanced prompt understanding, including user-specified intent like "think hard about this." Improvements include reductions in hallucination, sycophancy, and better performance in coding, writing, and health-related tasks.

Request body

Parameters supported by the OpenAI Responses API. Values, defaults, and limits are read from the model registry.

string | array<object>inputrequired

Input text or structured content for the response.

Optional<integer>max_tokens

Maximum number of tokens the model may generate in the response.

Maximum: 128000

Optional<number>temperature

Sampling temperature. Lower values are more deterministic; higher values are more creative.

Range: 0 to 2

Default: 1

Optional<boolean>stream

When true, returns incremental Server-Sent Events instead of one completed response.

Default: false

Optional<array>tools

Tool definitions that the model may call during the response.

Optional<string>tool_choice

Controls whether the model may call a tool and, when supported, which tool it must call.

Optional<object>response_format

Controls the response format, including JSON mode or structured JSON output when supported.

Optional<integer>seed

Seed used to make sampling more reproducible when the provider supports it.

Response Schema

Fields returned by the OpenAI Responses API.

stringidrequired

Unique response identifier.

stringobjectrequired

Response object type.

stringstatusrequired

Response lifecycle status.

stringmodelrequired

Model that generated the response.

array<object>outputrequired

Generated response output items.

Optional<object>usage

Token usage when available.

Model capabilities

array<string>capability_tagsrequired

Capabilities declared by the model registry.

Default: chat, vision, reasoning, structured_output, function_calling

integercontext_lengthrequired

Maximum context window accepted by this model.

Default: 400000 tokens

stringexecution_moderequired

Execution mode declared by the model registry.

Default: sync