Skip to main content
POST

Authorizations

Authorization
string
header
required

Bearer authentication header of the form Bearer <token>, where <token> is your auth token.

Body

application/json
model
string
required
Example:

"gpt-5.5"

input
required

Responses input. Can be a string or an array of message / tool-result items.

instructions
string

System or developer instructions.

max_output_tokens
integer

Maximum reasoning plus final-answer tokens. Model-specific limits apply; exhausting the budget can leave the answer incomplete.

Required range: x >= 1
temperature
number
Required range: 0 <= x <= 2
top_p
number
Required range: 0 <= x <= 1
reasoning
object
text
object
tools
object[]

Responses tools, including function tools and built-in tools such as web_search_preview.

tool_choice
stream
boolean

If true, responses are streamed as OpenAI Responses SSE events.

metadata
object
reasoning_effort
enum<string>

Compatibility alias for reasoning.effort. When both effort forms are present, their values must match; conflicting values return 400. Reasoning effort. These are syntactic values, not a support guarantee for every model. Supported values vary by model; unsupported values return 400. Kimi K3 supports low, high and max (default max); GPT-5.4/5.5 support none, low, medium, high, xhigh. See /models/chat-models for Claude and Gemini limits. Omission keeps the model default. Per-model levels, defaults and whether reasoning can be disabled are published in GET /v1/models under capability_metadata.reasoning.

Available options:
none,
minimal,
low,
medium,
high,
xhigh,
max

Response

OpenAI-compatible Responses response.

id
string
required
object
string
required
Example:

"response"

status
string
required
model
string
required
output
object[]
required
created_at
integer
output_text
string
usage
object