Openai logoOpenai

    gpt-4.1-nano

    Auto-router

    openai/gpt-4.1-nano

    ChatVisionFilesToolsSystem prompt

    Context

    1.0M

    Max output

    33K

    Input / 1M

    $0.10

    Output / 1M

    $0.40

    Cached / 1M

    $0.025

    Cutoff

    Jun 2024

    About this model

    GPT-4.1 nano is the fastest, most cost-effective GPT-4.1 model.

    Best suited for

    • Basic classification, formatting, extraction, and large-scale processing where speed and cost matter more than reasoning depth.

    Built-in tools

    Gtwy web searchImage generation

    Hosted by the gateway — enable them per request without wiring your own endpoint.

    Capabilities

    Vision

    Accepts images alongside text in the same message.

    Files

    Accepts file attachments — PDFs, transcripts, spreadsheets.

    Tools

    Native function calling, so agents can invoke your endpoints.

    System prompt

    Honours a dedicated system role, separate from the user turn.

    Supported parameters

    creativity_level

    Controls the creativity of responses. Higher values (e.g., 0.7) increase creativity; lower values (e.g., 0.2) make responses more predictable.

    max_tokensMax Tokens Limit

    Specifies the maximum number of text units (tokens) allowed in a response, limiting its length.

    tools

    Lists tool definitions or capabilities available to the model.

    tool_choice

    Decides whether to use tools or just the model for generating responses.

    response_type

    Defines the format or type of the generated response.

    parallel_tool_calls

    Enables parallel execution of tools, allowing multiple tools to run simultaneously.

    stream

    Sends the response in real-time as it's being generated.

    service_tier