Openai logoOpenai

    gpt-5.6-luna

    Auto-router

    openai/gpt-5.6-luna

    ChatVisionFilesToolsSystem promptReasoning
    Use this model

    Context

    1.1M

    Max output

    128K

    Input / 1M

    $0.20

    Output / 1M

    $1.20

    Cached / 1M

    Free

    Cutoff

    Feb 2026

    About this model

    GPT-5.6 Luna is OpenAI's fastest and most affordable model in the GPT-5.6 family. Built for high-volume, latency-sensitive workloads like chat, classification, and lightweight agentic tasks, it delivers strong reasoning at the lowest cost in the lineup — with a 1M+ token context window and 128K max output.

    Best suited for

    • High-volume chat & customer support — fast, low-cost responses for user-facing chatbots and support assistants at scale
    • Classification & routing — intent detection, sentiment analysis, content moderation, ticket triage, and query routing pipelines
    • Data extraction & structuring — pulling structured JSON from documents, emails, forms, and unstructured text
    • Lightweight agentic workflows — quick tool-calling steps, sub-agent tasks, and orchestration steps that don't need deep reasoning
    • Summarization at scale — condensing conversations, documents, and logs in bulk, aided by the 1M+ token context window
    • Autocomplete & real-time assistance — latency-sensitive features like suggestions, rewriting, and inline completions
    • Cost-optimized fallbacks — a default cheap tier in multi-model routing, escalating to Terra/Sol only when tasks demand it

    Built-in tools

    Image generationWeb searchGtwy web search

    Hosted by the gateway — enable them per request without wiring your own endpoint.

    Capabilities

    Vision

    Accepts images alongside text in the same message.

    Files

    Accepts file attachments — PDFs, transcripts, spreadsheets.

    Tools

    Native function calling, so agents can invoke your endpoints.

    System prompt

    Honours a dedicated system role, separate from the user turn.

    Reasoning

    Emits a separate thinking pass before the answer.

    Supported parameters

    max_tokens

    Controls the creativity of responses. Higher values (e.g., 0.7) increase creativity; lower values (e.g., 0.2) make responses more predictable.

    tools

    Lists tool definitions or capabilities available to the model.

    tool_choice

    Decides whether to use tools or just the model for generating responses.

    response_type
    parallel_tool_calls
    reasoning
    verbosity
    stream

    Sends the response in real-time as it's being generated.

    Use this model

    OpenAI-SDK compatible, with gateway fallback and routing across providers.

    Other Openai models

    ModelContextMax outputInputOutputCapabilities
    gpt-4o

    Chat

    16K$2.50$10.00Vision, Tools, System prompt
    gpt-4o-mini

    Chat

    128K8K$0.15$0.60Vision, Files, Tools, System prompt
    gpt-4.1

    Chat

    1.0M33K$2.00$8.00Vision, Files, Tools, System prompt
    gpt-4.1-mini

    Chat

    1.0M33K$0.40$1.60Vision, Files, Tools, System prompt
    gpt-4.1-nano

    Chat

    1.0M33K$0.10$0.40Vision, Files, Tools, System prompt
    gpt-5

    Chat

    400K128K$1.25$10.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5-mini

    Chat

    400K128K$0.25$2.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5-nano

    Chat

    400K128K$0.05$0.40Vision, Files, Tools, System prompt, Reasoning
    gpt-5.1

    Chat

    400K128K$1.25$10.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5.2

    Chat

    400K128K$1.75$14.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5.5

    Chat

    1.1M200K$5.00$30.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5.4-mini

    Chat

    400K128K$0.75$4.50Vision, Files, Tools, System prompt, Reasoning
    gpt-5.4-nano

    Chat

    400K128K$0.20$1.25Vision, Files, Tools, System prompt, Reasoning
    gpt-5.4

    Chat

    600K200K$2.25$18.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5.5-pro

    Chat

    400K128K$30.00$180.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5.6-sol

    Chat

    1.1M128K$5.00$30.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5.6-terra

    Chat

    1.1M128K$2.00$12.00Vision, Files, Tools, System prompt, Reasoning
    gpt-5.3-codex

    Chat

    400K$1.75$14.00Vision, Tools, System prompt
    gpt-6-astra

    Chat

    1.1M$10.00$50.00Vision, Files, Tools, System prompt
    gpt-4o-2024-08-06

    Completion

    128K16K$2.50$10.00Vision, Tools, System prompt
    gpt-4o-mini-2024-07-18

    Completion

    128K8K$0.15$0.60Vision, Tools, System prompt
    o1

    Reasoning

    200K100K$15.00$60.00Vision, Files, Tools, System prompt, Reasoning
    o3-mini

    Reasoning

    200K100K$1.10$4.40Files, Tools, System prompt, Reasoning
    o4-mini

    Reasoning

    200K100K$1.10$4.40Vision, Files, Tools, System prompt, Reasoning
    gpt-image-1.5

    Image

    Vision
    gpt-image-1

    Image

    Vision
    gpt-image-1-mini

    Image

    Vision
    gpt-image-2-2026-04-21

    Image

    400K228K$8.00Vision, Files, Tools, System prompt, Reasoning

    Pricing & provider details

    Openai logo

    Openai

    openai/gpt-5.6-luna

    Input · per 1M
    $0.20
    Output · per 1M
    $1.20
    Cached input · per 1M
    Free
    Context window
    1,050,000 tokens
    Max output
    128,000 tokens
    Knowledge cutoff
    Feb 2026
    Auto-router
    Supported

    Start building today

    Route gpt-5.6-luna — and every other model in the catalogue — through one endpoint, with failover built in.