Skip to main content
POST
Create a response (Alpha)
Beta endpoint, available to all API users. Venice translates Responses requests through Chat Completions, so some Responses features are ignored today: instructions, text.format, previous_response_id, reasoning replay, and provider-hosted tools. The Venice system prompt is also on by default here. See the Responses API guide for the full list and workarounds.

Authorizations

Authorization
string
header
required

Bearer authentication header of the form Bearer <token>, where <token> is your auth token.

Body

application/json

Request body for the Responses API endpoint. E2EE-capable models are not supported on /api/v1/responses; use /api/v1/chat/completions with the required E2EE headers instead.

model
string
required

The ID of the model to use. E2EE-capable models are not supported on /api/v1/responses; use /api/v1/chat/completions with the required E2EE headers instead.

Example:

"zai-org-glm-5-1"

input
required

The input to generate a response for. Can be a simple string or an array of messages.

include
string[]

Additional response fields to include (OpenAI-compatible).

max_output_tokens
integer

Maximum number of tokens to generate.

Required range: x > 0
temperature
number

Sampling temperature between 0 and 2.

Required range: 0 <= x <= 2
top_p
number

Nucleus sampling parameter.

Required range: 0 <= x <= 1
fallbacks
object[]

Anthropic beta parameter for Claude Fable 5 server-side refusal fallback. Forwarded only for direct Anthropic routes; ignored for other providers.

Maximum array length: 10
Example:
reasoning
Reasoning Configuration · object | null
tools
(Function Tool · object | Web Search Tool · object | X Search Tool · object | Code Interpreter Tool · object | File Search Tool · object | Computer Use Tool · object | Generic Tool · object)[]

A list of tools the model may call.

tool_choice

Controls which tool is called by the model.

Available options:
auto

Enable web search for this request.

stream
boolean

Whether to stream back partial progress.

anon_user_id
string

Optional identifier for the API customer's end user. Combined with the Venice user id when attributing the request to upstream providers. Distinct from the discarded OpenAI user field. Must be printable ASCII and must not contain ||.

Required string length: 1 - 128
Pattern: ^[\x20-\x7E]+$
Example:

"end-user-123"

venice_parameters
Venice Parameters · object

Response

Successful response

Response from the Responses API endpoint.

id
string
required

Unique identifier for the response.

Example:

"resp_abc123"

object
enum<string>
required

The object type.

Available options:
response
created_at
integer
required

Unix timestamp of when the response was created.

model
string
required

The model used for the response.

status
enum<string>
required

The status of the response.

Available options:
completed,
failed,
in_progress,
cancelled,
incomplete
output
(Reasoning Output · object | Message Output · object | Function Call Output · object | Web Search Call Output · object)[]
required

The output items generated by the model.

incomplete_details
object

Why generation ended before completion; partial output and usage are retained.

usage
Usage · object

Token usage statistics.

error
Error · object

Error information if the response failed.