Models

    Every Model. One Gateway.

    Compare pricing, context windows and capabilities across every provider GTWY routes to — then switch between them with a single field in your request body.

    104

    Models

    12

    Providers

    Catalogue

    Showing all 104 models across 12 providers

    Anthropic logoAuto-router

    Anthropic

    claude-fable-5

    AI

    Context

    200K

    In / 1M

    $5.00

    Out / 1M

    $15.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-haiku-4-5-20251001

    Fastest model with near-frontier intelligence

    Context

    200K

    In / 1M

    $1.00

    Out / 1M

    $5.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-opus-4-1

    test

    Context

    200K

    In / 1M

    $15.00

    Out / 1M

    $75.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-opus-4-1-20250805

    AI

    Context

    200K

    In / 1M

    $5.00

    Out / 1M

    $15.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-opus-4-5-20251101

    Premium model combining maximum intelligence with practical performance

    Context

    200K

    In / 1M

    $5.00

    Out / 1M

    $25.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-opus-4-6

    Our most intelligent model for building agents and coding

    Context

    200K

    In / 1M

    $5.00

    Out / 1M

    $25.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-opus-4-7

    Most powerful model for complex reasoning and coding

    Context

    1M

    In / 1M

    $5.00

    Out / 1M

    $25.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-opus-5

    1

    Context

    200K

    In / 1M

    $1.00

    Out / 1M

    $1.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-sonnet-4-5-20250929

    Highest intelligence across most tasks with exceptional agent and coding capabilities

    Context

    200K

    In / 1M

    $3.00

    Out / 1M

    $15.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-sonnet-4-6

    The best combination of speed and intelligence

    Context

    200K

    In / 1M

    $3.00

    Out / 1M

    $15.00

    Chat

    Anthropic logoAuto-router

    Anthropic

    claude-sonnet-5

    Anthropic's most agentic Sonnet model, closing much of the gap to Opus 4.8 on coding, agentic, and knowledge-work tasks at Sonnet-tier pricing.

    Context

    1M

    In / 1M

    $3.00

    Out / 1M

    $15.00

    Chat

    Deepgram logo

    Deepgram

    nova-2

    Context

    In / 1M

    Out / 1M

    Chat

    Deepgram logo

    Deepgram

    nova-3

    Context

    In / 1M

    Out / 1M

    Chat

    Deepseek logoAuto-router

    Deepseek

    deepseek-v4-flash

    DeepSeek-V4-Flash is a fast, capable model supporting thinking and non-thinking modes with 1M context. Accepts text inputs and produces text outputs.

    Context

    1M

    In / 1M

    $0.14

    Out / 1M

    $0.28

    Chat

    Deepseek logoAuto-router

    Deepseek

    deepseek-v4-pro

    DeepSeek-V4-Pro is a powerful flagship model supporting thinking and non-thinking modes with 1M context. Accepts text inputs and produces text outputs.

    Context

    1M

    In / 1M

    $0.435

    Out / 1M

    $0.87

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-2.5-flash

    best model in terms of price-performance, offering well-rounded capabilities. 2.5 Flash is best for large scale processing, low-latency, high volume tasks that require thinking, and agentic use cases.

    Context

    1.0M

    In / 1M

    $0.30

    Out / 1M

    $2.50

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-2.5-flash-image

    Gemini 2.5 Flash Image is optimized for image understanding and generation and offers a balance of price and performance. Gemini 2.5 Flash Image uses the speed and cost-effectiveness of Gemini 2.5 Flash to provide fast and efficient image generation and editing capabilities.

    Context

    In / 1M

    Out / 1M

    Image

    Gemini logoAuto-router

    Gemini

    gemini-2.5-flash-lite

    The most cost-efficient and fastest model in the 2.5 family, optimized for extreme low-latency and high-throughput scenarios.

    Context

    1.0M

    In / 1M

    $0.10

    Out / 1M

    $0.40

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-2.5-pro

    state-of-the-art thinking model, capable of reasoning over complex problems in code, math, and STEM, as well as analyzing large datasets, codebases, and documents using long context.

    Context

    1.0M

    In / 1M

    $1.25

    Out / 1M

    $10.00

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-3-flash-preview

    Gemini 3 Flash combines Gemini 3 Pro's reasoning capabilities with the Flash line's levels on latency, efficiency, and cost

    Context

    1.0M

    In / 1M

    $0.50

    Out / 1M

    $3.00

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-3-pro-image-preview

    The gemini-3-pro-image-preview model, also known as Nano Banana Pro, is Google's image generation and editing model. It launched in late 2025. Built on the Gemini 3 Pro architecture, it has a Thinking Mode. This mode allows the model to reason through complex instructions. It results in higher accuracy and factuality compared to earlier versions. The model supports resolutions up to 4K. It is the first to integrate Grounding with Google Search.

    Context

    In / 1M

    Out / 1M

    Image

    Gemini logoAuto-router

    Gemini

    gemini-3-pro-preview

    Gemini 3 Pro Preview is our most powerful agentic and coding model.

    Context

    1.0M

    In / 1M

    $2.00

    Out / 1M

    $12.00

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-3.1-pro-preview

    Gemini 3.1 Pro is our most advanced reasoning Gemini model, capable of solving complex problems.

    Context

    1.0M

    In / 1M

    $2.00

    Out / 1M

    $12.00

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-3.5-flash

    Newest and most capable Speed oriented Frontier model from Google.

    Context

    1.0M

    In / 1M

    $15.00

    Out / 1M

    $35.00

    Chat

    Gemini logoAuto-router

    Gemini

    gemini-3.7-flash

    Gemini 3.7 Flash is Google's fast, cost-efficient multimodal model released August 2026, featuring a 1M-token context window, tunable thinking (reasoning) levels, and improved coding and agentic performance over Gemini 3.6 Flash.

    Context

    1M

    In / 1M

    $0.75

    Out / 1M

    $3.75

    Chat

    Gemini logoAuto-router

    Gemini

    imagen-4.0-fast-generate-001

    This is the Low Latency variant. It is designed for high-speed generation.

    Context

    In / 1M

    Out / 1M

    Image

    Gemini logoAuto-router

    Gemini

    imagen-4.0-generate-001

    This Standard model balances quality and speed. It supports resolutions up to 2K (2048x2048) and multilingual prompts in 9 languages.

    Context

    In / 1M

    Out / 1M

    Image

    Gemini logoAuto-router

    Gemini

    imagen-4.0-ultra-generate-001

    This is the Highest Quality model. It offers image fidelity and text rendering and is optimized for premium production needs. This model generates one image at a time.

    Context

    In / 1M

    Out / 1M

    Image

    Grok logo

    Grok

    grok-4-0709

    Grok Models

    Context

    2M

    In / 1M

    $0.20

    Out / 1M

    $0.50

    Chat

    Grok logo

    Grok

    grok-4-fast

    Grok Models

    Context

    2M

    In / 1M

    $0.20

    Out / 1M

    $0.50

    Chat

    Grok logo

    Grok

    grok-4-fast-reasoning

    Grok-4 Fast Reasoning – optimized for quick structured reasoning and planning tasks

    Context

    2M

    In / 1M

    $0.25

    Out / 1M

    $0.60

    Chat

    Groq logo

    Groq

    llama-3.1-8b-instant

    Llama 3.1 8B on Groq provides low-latency, high-quality responses suitable for real-time conversational interfaces, content filtering systems, and data analysis applications. This model offers a balance of speed and performance with significant cost savings compared to larger models. Technical capabilities include native function calling support, JSON mode for structured output generation, and a 128K token context window for handling large documents.

    Context

    131K

    In / 1M

    $0.05

    Out / 1M

    $0.08

    Chat

    Groq logo

    Groq

    llama-3.3-70b-versatile

    Llama-3.3-70B-Versatile is Meta's advanced multilingual large language model, optimized for a wide range of natural language processing tasks. With 70 billion parameters, it offers high performance across various benchmarks while maintaining efficiency suitable for diverse applications.

    Context

    33K

    In / 1M

    $0.59

    Out / 1M

    $0.79

    Chat

    Groq logo

    Groq

    meta-llama/llama-4-scout-17b-16e-instruct

    Llama 4 Scout is Meta's natively multimodal model that enables text and image understanding. With a 17 billion parameter mixture-of-experts architecture (16 experts), this model offers industry-leading performance for multimodal tasks like natural assistant-like chat, image recognition, and coding tasks. With a 128K token context window and support for 12 languages (Arabic, English, French, German, Hindi, Indonesian, Italian, Portuguese, Spanish, Tagalog, Thai, and Vietnamese), the model delivers exceptional capabilities, especially when paired with Groq for fast inference.

    Context

    In / 1M

    $0.11

    Out / 1M

    $0.34

    Chat

    Groq logo

    Groq

    openai/gpt-oss-120b

    GPT-oss-120b is the most powerful open-weight model. A 120B-parameter open-weight Mixture-of-Experts model delivering state-of-the-art reasoning, coding, and tool-use performance

    Context

    131K

    In / 1M

    Out / 1M

    Chat

    Groq logo

    Groq

    openai/gpt-oss-20b

    GPT-oss-120b is the most powerful open-weight model. A 120B-parameter open-weight Mixture-of-Experts model delivering state-of-the-art reasoning, coding, and tool-use performance

    Context

    131K

    In / 1M

    Out / 1M

    Chat

    Minimax logo

    Minimax

    minimax-m3

    MiniMax-M3 is an open-weight, natively multimodal Mixture-of-Experts model with ~428B total parameters (~23B activated), built on MiniMax Sparse Attention (MSA) for efficient long-context processing. Accepts text, image, video, and PDF inputs and produces text outputs, with a deep thinking mode for complex reasoning.

    Context

    1.0M

    In / 1M

    $0.30

    Out / 1M

    $1.20

    Chat

    Mistral logoAuto-router

    Mistral

    codestral-latest

    Lightweight, fast, and proficient in over 80 programming languages.

    Context

    256K

    In / 1M

    $0.30

    Out / 1M

    $0.90

    Chat

    Mistral logoAuto-router

    Mistral

    magistral-medium-latest

    Context

    41K

    In / 1M

    $0.40

    Out / 1M

    $2.00

    Chat

    Mistral logoAuto-router

    Mistral

    magistral-small-latest

    Lightweight, fast, and proficient in over 80 programming languages.

    Context

    128K

    In / 1M

    $0.50

    Out / 1M

    $1.50

    Chat

    Mistral logoAuto-router

    Mistral

    mistral-medium-latest

    Context

    128K

    In / 1M

    $0.40

    Out / 1M

    $2.00

    Chat

    Mistral logoAuto-router

    Mistral

    mistral-small-latest

    Lightweight, fast, and proficient in over 80 programming languages.

    Context

    128K

    In / 1M

    $0.10

    Out / 1M

    $0.30

    Chat

    Moonshot logo

    Moonshot

    kimi-k2.5

    Kimi K2.5 is Moonshot AI's multimodal model supporting text, image, and video input, with thinking and non-thinking modes, and both dialogue and agent tasks. Context length 256K with support for long thinking and deep reasoning. Supports automatic context caching, ToolCalls, JSON Mode, Partial Mode, and internet search.

    Context

    262K

    In / 1M

    $0.60

    Out / 1M

    $3.00

    Chat

    Moonshot logo

    Moonshot

    kimi-k2.6

    Kimi K2.6 is Moonshot AI's latest and most intelligent model, with stronger and more stable long-horizon code generation, significantly improved instruction following and self-correction. It features a native multimodal architecture supporting text, image, and video input, thinking and non-thinking modes, and both dialogue and agent tasks. Supports automatic context caching, ToolCalls, JSON Mode, Partial Mode, and internet search.

    Context

    262K

    In / 1M

    $0.95

    Out / 1M

    $4.00

    Chat

    Moonshot logo

    Moonshot

    kimi-k2.7-code

    Kimi K2.6 is Moonshot AI's latest and most intelligent model, with stronger and more stable long-horizon code generation, significantly improved instruction following and self-correction. It features a native multimodal architecture supporting text, image, and video input, thinking and non-thinking modes, and both dialogue and agent tasks. Supports automatic context caching, ToolCalls, JSON Mode, Partial Mode, and internet search.

    Context

    262K

    In / 1M

    $0.95

    Out / 1M

    $4.00

    Chat

    Moonshot logo

    Moonshot

    kimi-k2.7-code-highspeed

    Kimi K2.6 is Moonshot AI's latest and most intelligent model, with stronger and more stable long-horizon code generation, significantly improved instruction following and self-correction. It features a native multimodal architecture supporting text, image, and video input, thinking and non-thinking modes, and both dialogue and agent tasks. Supports automatic context caching, ToolCalls, JSON Mode, Partial Mode, and internet search.

    Context

    262K

    In / 1M

    $1.90

    Out / 1M

    $8.00

    Chat

    Moonshot logo

    Moonshot

    moonshot-v1-128k

    Moonshot V1 128K is Moonshot AI's generation model with a 131,072-token context window, suitable for generating very long texts and handling long-context chat, generation, and tool-calling tasks. Supports ToolCalls, JSON Mode, and Partial Mode.

    Context

    131K

    In / 1M

    $2.00

    Out / 1M

    $5.00

    Chat

    Moonshot logo

    Moonshot

    moonshot-v1-128k-vision-preview

    Moonshot V1 128K Vision is Moonshot AI's vision-capable generation model with a 131,072-token context window. It understands image content and outputs text, suited for long-context multimodal chat, image understanding, and tool-calling tasks. Supports ToolCalls, JSON Mode, and Partial Mode.

    Context

    131K

    In / 1M

    $2.00

    Out / 1M

    $5.00

    Chat

    Moonshot logo

    Moonshot

    moonshot-v1-32k

    Moonshot V1 32K is Moonshot AI's generation model with a 32,768-token context window, suited for medium-context chat, generation, and tool-calling tasks. Supports ToolCalls, JSON Mode, and Partial Mode.

    Context

    33K

    In / 1M

    $1.00

    Out / 1M

    $3.00

    Chat

    Moonshot logo

    Moonshot

    moonshot-v1-32k-vision-preview

    Moonshot V1 32K Vision is Moonshot AI's vision-capable generation model with a 32,768-token context window. It understands image content and outputs text, suited for medium-context multimodal chat, image understanding, and tool-calling tasks. Supports ToolCalls, JSON Mode, and Partial Mode.

    Context

    33K

    In / 1M

    $1.00

    Out / 1M

    $3.00

    Chat

    Moonshot logo

    Moonshot

    moonshot-v1-8k

    Moonshot V1 8K is Moonshot AI's generation model with an 8,192-token context window, suited for short-context chat, generation, and tool-calling tasks. Supports ToolCalls, JSON Mode, and Partial Mode.

    Context

    8K

    In / 1M

    $0.20

    Out / 1M

    $2.00

    Chat

    Moonshot logo

    Moonshot

    moonshot-v1-8k-vision-preview

    Moonshot V1 8K Vision is Moonshot AI's vision-capable generation model with an 8,192-token context window. It understands image content and outputs text, suited for short-context multimodal chat, image understanding, and tool-calling tasks. Supports ToolCalls, JSON Mode, and Partial Mode.

    Context

    8K

    In / 1M

    $0.20

    Out / 1M

    $2.00

    Chat

    Neev Cloud logo

    Neev Cloud

    deepseek-v3-2

    Balances reasoning capability with output length, designed for Q&A and general agent workflows, optimized for long-context scenarios.

    Context

    In / 1M

    $26.66

    Out / 1M

    $39.99

    Chat

    Neev Cloud logo

    Neev Cloud

    glm-4-7

    Flagship GLM model with stronger coding, reliable multi-step reasoning, improved agentic workflows, and enhanced front-end generation quality.

    Context

    In / 1M

    $57.12

    Out / 1M

    $209.45

    Chat

    Neev Cloud logo

    Neev Cloud

    gpt-oss-120b

    OpenAI’s open-weight model designed for powerful reasoning, agentic tasks, and versatile developer use cases.

    Context

    In / 1M

    $9.52

    Out / 1M

    $47.60

    Chat

    Neev Cloud logo

    Neev Cloud

    gpt-oss-20b

    Compact open-weight Mixture-of-Experts model optimized for cost-efficient deployment and agentic workflows.

    Context

    In / 1M

    $7.14

    Out / 1M

    $28.56

    Chat

    Neev Cloud logo

    Neev Cloud

    llama-3.1-8b-instant

    Low-latency model suitable for real-time conversational interfaces, content filtering, and general analysis.

    Context

    In / 1M

    $4.76

    Out / 1M

    $7.62

    Chat

    Neev Cloud logo

    Neev Cloud

    llama-3.3-70b-versatile

    Advanced multilingual large language model optimized for a wide range of natural language tasks with strong performance and efficiency.

    Context

    In / 1M

    $56.17

    Out / 1M

    $75.21

    Chat

    Neev Cloud logo

    Neev Cloud

    minimax-m2.7

    MiniMax-M2.7 is a high-performance long-context open-weight model optimized for engineering, reasoning, and professional conversational tasks.

    Context

    In / 1M

    $0.2852

    Out / 1M

    $1.1409

    Chat

    Neev Cloud logo

    Neev Cloud

    minimax-m2.7-highspeed

    MiniMax-M2.7-highspeed delivers the same performance as MiniMax-M2.7 with significantly faster inference and lower latency.

    Context

    In / 1M

    $57.12

    Out / 1M

    $228.49

    Chat

    Open Router logo

    Open Router

    cognitivecomputations/dolphin-mistral-24b-venice-edition:free

    ljkjhgfd

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    cognitivecomputations/dolphin3.0-mistral-24b:free

    dolphin

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Completion

    Open Router logo

    Open Router

    deepseek/deepseek-chat-v3-0324:free

    DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team.

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    deepseek/deepseek-r1-0528-qwen3-8b:free

    DeepSeek-R1-0528 is a lightly upgraded release of DeepSeek R1 that taps more compute and smarter post-training tricks, pushing its reasoning and inference to the brink of flagship models like O3 and Gemini 2.5 Pro.

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    deepseek/deepseek-r1:free

    DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass.

    Context

    164K

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    google/gemini-3-flash-preview

    bjsdahfkjasjdfk;jso

    Context

    In / 1M

    $0.01

    Out / 1M

    $0.12

    Chat

    Open Router logo

    Open Router

    google/gemma-3n-e2b-it:free

    mono

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    mistralai/devstral-small-2505:free

    wd

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    mistralai/mistral-small-3.2-24b-instruct:free

    aed

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    moonshotai/kimi-k2:free

    asdc

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    openai/gpt-4o

    GPT-4o is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of GPT-4 Turbo while being twice as fast and 50% more cost-effective. GPT-4o also offers improved performance in processing non-English languages and enhanced visual capabilities.

    Context

    128K

    In / 1M

    $2.50

    Out / 1M

    $10.00

    Chat

    Open Router logo

    Open Router

    openai/gpt-5.4-mini

    GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments.

    Context

    In / 1M

    $0.75

    Out / 1M

    $4.50

    Chat

    Open Router logo

    Open Router

    qwen/qwen3-coder:free

    qwen/qwen3-coder:free

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    tencent/hunyuan-a13b-instruct:free

    sadv

    Context

    In / 1M

    Free

    Out / 1M

    Free

    Chat

    Open Router logo

    Open Router

    z-ai/glm-4.5-air:free

    qwen/qwen3-coder:free

    Context

    In / 1M

    $32.00

    Out / 1M

    $22.998

    Chat

    Openai logoAuto-router

    Openai

    gpt-4.1

    GPT-4.1 is our flagship model for complex tasks. It is well suited for problem solving across domains.

    Context

    1.0M

    In / 1M

    $2.00

    Out / 1M

    $8.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-4.1-mini

    GPT-4.1 mini provides a balance between intelligence, speed, and cost that makes it an attractive model for many use cases.

    Context

    1.0M

    In / 1M

    $0.40

    Out / 1M

    $1.60

    Chat

    Openai logoAuto-router

    Openai

    gpt-4.1-nano

    GPT-4.1 nano is the fastest, most cost-effective GPT-4.1 model.

    Context

    1.0M

    In / 1M

    $0.10

    Out / 1M

    $0.40

    Chat

    Openai logoAuto-router

    Openai

    gpt-4o

    GPT-4o provides a balance between intelligence, speed, and cost that makes it an attractive model for many use cases.

    Context

    In / 1M

    $2.50

    Out / 1M

    $10.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-4o-2024-08-06

    gpt-4o-2024-08-06 model which is fast, Intelligent model for all purpose tasks. It accepts both text and image inputs, and produces text outputs.

    Context

    128K

    In / 1M

    $2.50

    Out / 1M

    $10.00

    Completion

    Openai logoAuto-router

    Openai

    gpt-4o-mini

    GPT-4o mini (“o” for “omni”) is a fast, affordable small model for focused tasks. It accepts both text and image inputs, and produces text outputs (including Structured Outputs). It is ideal for fine-tuning, and model outputs from a larger model like GPT-4o can be distilled to GPT-4o-mini to produce similar results at lower cost and latency.

    Context

    128K

    In / 1M

    $0.15

    Out / 1M

    $0.60

    Chat

    Openai logoAuto-router

    Openai

    gpt-4o-mini-2024-07-18

    gpt-4o-mini-2024-07-18 is a fast, affordable small model for focused tasks. It accepts both text and image inputs, and produces text outputs (including Structured Outputs). It is ideal for fine-tuning, and model outputs from a larger model like GPT-4o can be distilled to GPT-4o-mini to produce similar results at lower cost and latency.

    Context

    128K

    In / 1M

    $0.15

    Out / 1M

    $0.60

    Completion

    Openai logoAuto-router

    Openai

    gpt-5

    GPT-5 model which is fast, Intelligent model for all purpose tasks. It accepts both text and image inputs, and produces text outputs.

    Context

    400K

    In / 1M

    $1.25

    Out / 1M

    $10.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5-mini

    GPT-5-mini model which is fast, Intelligent model for all purpose tasks. It accepts both text and image inputs, and produces text outputs.

    Context

    400K

    In / 1M

    $0.25

    Out / 1M

    $2.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5-nano

    GPT-5-nano model which is fast, Intelligent model for all purpose tasks. It accepts both text and image inputs, and produces text outputs.

    Context

    400K

    In / 1M

    $0.05

    Out / 1M

    $0.40

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.1

    GPT-5.1 is a frontier-grade OpenAI model with stronger general-purpose reasoning and better instruction following than GPT-5. It accepts text and image inputs and produces text outputs.

    Context

    400K

    In / 1M

    $1.25

    Out / 1M

    $10.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.2

    GPT-5.2 is our best general-purpose model, part of the GPT-5 flagship model family. Our most intelligent model yet for both general and agentic tasks

    Context

    400K

    In / 1M

    $1.75

    Out / 1M

    $14.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.3-codex

    GPT-5.3-Codex is OpenAI's most capable agentic coding model to date, optimized for agentic coding tasks in Codex and similar environments. Supports low/medium/high/xhigh reasoning effort.

    Context

    400K

    In / 1M

    $1.75

    Out / 1M

    $14.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.4

    GPT-5.4 is our most advanced general-purpose model in the GPT-5 family. Optimized for complex reasoning, long-context agentic workflows, multimodal processing, and superior code generation.

    Context

    600K

    In / 1M

    $2.25

    Out / 1M

    $18.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.4-mini

    GPT-5.4 mini is a fast, cost-efficient, tool-capable model designed for scalable agent workflows and real-time automation.

    Context

    400K

    In / 1M

    $0.75

    Out / 1M

    $4.50

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.4-nano

    GPT-5.4 nano is an ultra-fast, low-cost model optimized for simple tasks, high-scale automation, and lightweight tool usage.

    Context

    400K

    In / 1M

    $0.20

    Out / 1M

    $1.25

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.5

    GPT-5.5 is OpenAI's most advanced frontier model, optimized for complex real-world work including agentic reasoning, long-context workflows, tool use, and high-performance code generation. It can plan, execute, and complete multi-step tasks with minimal guidance while maintaining strong context over long durations.

    Context

    1.1M

    In / 1M

    $5.00

    Out / 1M

    $30.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.5-pro

    GPT-5.5 Pro uses more compute to think harder and provide consistently better answers.

    Context

    400K

    In / 1M

    $30.00

    Out / 1M

    $180.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.6-luna

    GPT-5.6 Luna is OpenAI's fastest and most affordable model in the GPT-5.6 family. Built for high-volume, latency-sensitive workloads like chat, classification, and lightweight agentic tasks, it delivers strong reasoning at the lowest cost in the lineup — with a 1M+ token context window and 128K max output.

    Context

    1.1M

    In / 1M

    $0.20

    Out / 1M

    $1.20

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.6-sol

    GPT-5.6 Sol is OpenAI's flagship frontier model for complex reasoning and professional work. 1.05M context, knowledge cutoff Feb 2026.

    Context

    1.1M

    In / 1M

    $5.00

    Out / 1M

    $30.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-5.6-terra

    GPT-5.6 Terra balances intelligence and cost in the GPT-5.6 family. 1.05M context, knowledge cutoff Feb 2026.

    Context

    1.1M

    In / 1M

    $2.00

    Out / 1M

    $12.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-6-astra

    GPT-6 Astra is OpenAI's most capable broadly deployed model, built for the hardest end-to-end work. It excels at complex reasoning, coding, computer use, research, and document creation, and is OpenAI's first model to reach the Critical level of cybersecurity capability under its Preparedness Framework.

    Context

    1.1M

    In / 1M

    $10.00

    Out / 1M

    $50.00

    Chat

    Openai logoAuto-router

    Openai

    gpt-image-1

    GPT Image 1 is a natively multimodal language model that accepts both text and image inputs, and produces image outputs.

    Context

    In / 1M

    Out / 1M

    Image

    Openai logoAuto-router

    Openai

    gpt-image-1-mini

    A cost-efficient version of GPT Image 1. It is a natively multimodal language model that accepts both text and image inputs, and produces image outputs.

    Context

    In / 1M

    Out / 1M

    Image

    Openai logoAuto-router

    Openai

    gpt-image-1.5

    GPT Image 1.5 is our latest image generation model, with better instruction following and adherence to prompts.

    Context

    In / 1M

    Out / 1M

    Image

    Openai logoAuto-router

    Openai

    gpt-image-2-2026-04-21

    You are the world's best creative graphic designer and illustration specialist with industry experience of two decades.

    Context

    400K

    In / 1M

    $8.00

    Out / 1M

    Image

    Openai logoAuto-router

    Openai

    o1

    The o1 series of models are trained with reinforcement learning to perform complex reasoning. o1 models think before they answer, producing a long internal chain of thought before responding to the user.

    Context

    200K

    In / 1M

    $15.00

    Out / 1M

    $60.00

    Reasoning

    Openai logoAuto-router

    Openai

    o3-mini

    o3-mini is our newest small reasoning model, providing high intelligence at the same cost and latency targets of o1-mini. o3-mini supports key developer features, like Structured Outputs, function calling, and Batch API.

    Context

    200K

    In / 1M

    $1.10

    Out / 1M

    $4.40

    Reasoning

    Openai logoAuto-router

    Openai

    o4-mini

    o4-mini is our latest small o-series model. It's optimized for fast, effective reasoning with exceptionally efficient performance in coding and visual tasks.

    Context

    200K

    In / 1M

    $1.10

    Out / 1M

    $4.40

    Reasoning