Skip to main content
POST
Create chat completion

請求主體

string
必填
要使用的模型 ID(例如 gpt-4o、claude-3.5-sonnet、gemini-2.0-flash)。
array
必填
組成對話的訊息列表。
boolean
預設值:false
如果為 true,回傳 Server-Sent Events (SSE) 串流。
number
預設值:1
取樣溫度,介於 0 到 2 之間。較低的值更具確定性。
integer
要生成的最大 token 數。
number
預設值:1
核取樣參數。
string | string[]
將停止生成的停止序列。
number
預設值:0
token 頻率懲罰(-2.0 到 2.0)。
number
預設值:0
token 存在懲罰(-2.0 到 2.0)。
array
模型可呼叫的工具列表(函數呼叫)。
string | object
控制呼叫哪個工具。auto、none 或指定的工具。
object
強制指定輸出格式(例如 {"type": "json_object"})。
integer
用於可重現性的確定性取樣種子。

OpenModex 擴充功能

object
為此請求設定智慧路由。
object
設定提示快取。

回應

範例

授權

Authorization
string
header
必填

API key authentication. Pass your OpenModex API key as a Bearer token in the Authorization header: Authorization: Bearer omx_sk_...

主體

application/json

Request body for creating a chat completion.

model
string
必填

ID of the model to use.

messages
object[]
必填

A list of messages comprising the conversation so far.

stream
boolean
預設值:false

If true, partial message deltas will be sent as server-sent events.

temperature
number

Sampling temperature between 0 and 2. Higher values make output more random.

必填範圍: 0 <= x <= 2
top_p
number

Nucleus sampling parameter. Only consider tokens with top_p probability mass.

n
integer

How many completions to generate for each prompt.

max_tokens
integer

The maximum number of tokens to generate in the completion.

stop

Up to 4 sequences where the API will stop generating further tokens.

frequency_penalty
number

Penalizes new tokens based on their existing frequency in the text so far.

必填範圍: -2 <= x <= 2
presence_penalty
number

Penalizes new tokens based on whether they appear in the text so far.

必填範圍: -2 <= x <= 2
logit_bias
object

Modify the likelihood of specified tokens appearing in the completion. Maps token IDs to bias values from -100 to 100.

tools
object[]

A list of tools the model may call.

tool_choice

Controls which tool is called by the model. Can be 'none', 'auto', 'required', or a specific tool object.

response_format
object

An object specifying the format the model must output (e.g., {"type": "json_object"}).

seed
integer

If specified, the system will attempt to sample deterministically.

user
string

A unique identifier representing your end-user, for abuse monitoring.

routing
object

Configuration for intelligent request routing.

cache
object

Configuration for response caching.

回應

Chat completion response. When stream=true, returns SSE stream of ChatCompletionChunk objects.

Response from a chat completion request.

id
string

A unique identifier for the completion.

object
enum<string>

The object type, always 'chat.completion'.

可用選項:
chat.completion
created
integer

Unix timestamp (in seconds) of when the completion was created.

model
string

The model used for the completion.

choices
object[]

A list of completion choices.

usage
object

Token usage statistics for a completion request.

system_fingerprint
string

A fingerprint representing the backend configuration.

openmodex
object

OpenModex-specific metadata included in completion responses.