Skip to content

Latest commit

 

History

History
19 lines (16 loc) · 22.3 KB

File metadata and controls

19 lines (16 loc) · 22.3 KB

AgentsCompletionRequest

Fields

Field Type Required Description Example
messages List[models.AgentsCompletionRequestMessages] ✔️ The prompt(s) to generate completions for, encoded as a list of dict with role and content. [
{
"role": "user",
"content": "Who is the best French painter? Answer in one short sentence."
}
]
agent_id str ✔️ The ID of the agent to use for this completion.
max_tokens OptionalNullable[int] The maximum number of tokens to generate in the completion. The token count of your prompt plus max_tokens cannot exceed the model's context length.
stream Optional[bool] Whether to stream back partial progress. If set, tokens will be sent as data-only server-side events as they become available, with the stream terminated by a data: [DONE] message. Otherwise, the server will hold the request open until the timeout or until completion, with the response containing the full result as JSON.
stop Optional[models.AgentsCompletionRequestStop] Stop generation if this token is detected. Or if one of these tokens is detected when providing an array
random_seed OptionalNullable[int] The seed to use for random sampling. If set, different calls will generate deterministic results.
response_format Optional[models.ResponseFormat] N/A
tools List[models.Tool] N/A
tool_choice Optional[models.AgentsCompletionRequestToolChoice] N/A
presence_penalty Optional[float] presence_penalty determines how much the model penalizes the repetition of words or phrases. A higher presence penalty encourages the model to use a wider variety of words and phrases, making the output more diverse and creative.
frequency_penalty Optional[float] frequency_penalty penalizes the repetition of words based on their frequency in the generated text. A higher frequency penalty discourages the model from repeating words that have already appeared frequently in the output, promoting diversity and reducing repetition.
n OptionalNullable[int] Number of completions to return for each request, input tokens are only billed once.