Openai logoOpenai

    gpt-4o

    Auto-router

    openai/gpt-4o

    ChatVisionToolsSystem prompt

    Context

    Max output

    16K

    Input / 1M

    $2.50

    Output / 1M

    $10.00

    Modality

    Chat

    Cutoff

    Jun 2024

    About this model

    GPT-4o provides a balance between intelligence, speed, and cost that makes it an attractive model for many use cases.

    Best suited for

    • Multimodal assistants, real-time voice/vision use cases, interactive AI features, complex content creation, and rich user-facing applications.

    Built-in tools

    Image generationWeb searchGtwy web search

    Hosted by the gateway — enable them per request without wiring your own endpoint.

    Capabilities

    Vision

    Accepts images alongside text in the same message.

    Tools

    Native function calling, so agents can invoke your endpoints.

    System prompt

    Honours a dedicated system role, separate from the user turn.

    Supported parameters

    creativity_level

    Controls the creativity of responses. Higher values (e.g., 0.7) increase creativity; lower values (e.g., 0.2) make responses more predictable.

    max_tokensMax Tokens Limit

    Specifies the maximum number of text units (tokens) allowed in a response, limiting its length.

    tools

    Lists tool definitions or capabilities available to the model.

    tool_choice

    Decides whether to use tools or just the model for generating responses.

    response_type

    Defines the format or type of the generated response.

    parallel_tool_calls

    Enables parallel execution of tools, allowing multiple tools to run simultaneously.

    stream

    Sends the response in real-time as it's being generated.

    service_tier
    temperature