gpt-4o
Auto-routeropenai/gpt-4o
Context
—
Max output
16K
Input / 1M
$2.50
Output / 1M
$10.00
Modality
Chat
Cutoff
Jun 2024
About this model
GPT-4o provides a balance between intelligence, speed, and cost that makes it an attractive model for many use cases.
Best suited for
- Multimodal assistants, real-time voice/vision use cases, interactive AI features, complex content creation, and rich user-facing applications.
Built-in tools
Hosted by the gateway — enable them per request without wiring your own endpoint.
Capabilities
Vision
Accepts images alongside text in the same message.
Tools
Native function calling, so agents can invoke your endpoints.
System prompt
Honours a dedicated system role, separate from the user turn.
Supported parameters
creativity_levelControls the creativity of responses. Higher values (e.g., 0.7) increase creativity; lower values (e.g., 0.2) make responses more predictable.
max_tokensMax Tokens LimitSpecifies the maximum number of text units (tokens) allowed in a response, limiting its length.
toolsLists tool definitions or capabilities available to the model.
tool_choiceDecides whether to use tools or just the model for generating responses.
response_typeDefines the format or type of the generated response.
parallel_tool_callsEnables parallel execution of tools, allowing multiple tools to run simultaneously.
streamSends the response in real-time as it's being generated.
service_tiertemperature