OpenAIProprietary

GPT-5.6 Luna

Compare this model

The lowest-cost model developed by OpenAI. It achieves high-speed processing.

Parameters

Undisclosed

Context Window

License

Proprietary

Release Date

2026-06-26

Japanese Language Capability

High-Quality JP

Multilingual model with strong Japanese language processing capabilities.

API Pricing

Input Price (per 1M tokens)

$1

Output Price (per 1M tokens)

$

Billing Mode: standard

Strengths

    Weaknesses

      Use Cases

        Deep Analysis

        Artificial Analysis Intelligence Index

        51

        vs Terra: 55, Sol: 59 (lowest in family)

        Cost per Task (AA Index)

        $0.21

        vs Terra: $0.55, Sol: $1.04

        Terminal-Bench 2.1

        84.3%

        Close to Terra (82.5%) and GPT-5.5 (88.0%)

        Input Price

        $1/1M tokens

        Lowest in GPT-5.6 family (Terra $2.50, Sol $5)

        Output Price

        $6/1M tokens

        Lowest in GPT-5.6 family (Terra $15, Sol $30)

        Context Window

        1,050,000 tokens

        Identical to flagship Sol

        Strengths

        • Lowest cost per task ($0.21) in the GPT-5.6 family for high-volume work.
        • Fastest inference speed in the family, ideal for latency-sensitive pipelines.
        • Includes the full agentic tool stack (Programmatic Tool Calling, MCP, etc.) at the lowest price tier.

        Weaknesses

        • Least intelligent model in the GPT-5.6 family (AA Intelligence Index: 51).
        • Poor long-context retrieval performance (MRCR score: 41.3%).
        • No fine-tuning support at launch and not available in consumer ChatGPT.

        Competitor Comparison

        ModelArenaSWEGPQAPrice
        GPT-5.6 TerraN/A63.4%92.9%$2.50/$15
        GPT-5.6 SolN/A64.6%94.6%$5/$30
        Claude Fable 5N/A80%N/A$10/$50

        GPT-5.6 Luna is OpenAI's most cost-efficient and fastest model, released as part of the GPT-5.6 family on July 9, 2026. Positioned as the 'efficiency tier' below the flagship Sol and balanced Terra, Luna is designed for high-volume, latency-sensitive, and well-scoped tasks like summarization, classification, and batch processing. It shares the family's 1-million-token context window and full agentic toolbox but is optimized for cost and speed over raw intelligence. Independent benchmarks confirm its leadership on cost-performance: it offers the lowest cost per task in the family while outperforming previous-generation models and more expensive competitors on specific, parallelizable workloads.

        The launch of Luna, alongside Sol and Terra, represents OpenAI's strategy to dominate the cost-performance frontier across all price points. While it trails its siblings on complex reasoning and long-horizon tasks, Luna makes frontier AI capabilities accessible for massive-scale applications where per-task cost is the primary constraint. Its introduction, along with features like Programmatic Tool Calling and explicit cache breakpoints, lowers the barrier for building sophisticated, high-volume AI pipelines.

        Analysis generated: 2026-07-17