OpenAIProprietary

GPT-5.5 Instant

Compare this model

A high-speed foundation model developed by OpenAI. It balances low latency with high throughput and is the predecessor to GPT-5.6 Luna.

Parameters

Undisclosed

Context Window

License

Proprietary

Release Date

2026-04-01

Japanese Language Capability

High-Quality JP

Multilingual model with strong Japanese language processing capabilities.

API Pricing

API pricing for this model is not yet available

Strengths

    Weaknesses

      Use Cases

        Deep Analysis

        Artificial Analysis Intelligence Index

        34

        #77 of 186 models; above average (median: 30)

        GPQA Diamond

        84.6%

        #84 across all tracked models

        AIME 2025

        81.2%

        Up from 65.4% on GPT-5.3 Instant

        MMMU-Pro (Multimodal Reasoning)

        76%

        Up from 69.2% on GPT-5.3 Instant

        Input Price

        $5.00/1M tokens

        Cache hit: $0.50/1M (90% discount)

        Output Price

        $30.00/1M tokens

        Blended 3:1 ratio: $11.25/1M

        Hallucination Reduction

        52.5% fewer

        vs GPT-5.3 Instant on high-stakes prompts

        Context Window

        400K tokens

        ~600 pages; 128K in ChatGPT UI

        Strengths

        • 52.5% fewer hallucinations on high-stakes medical, legal, and financial prompts vs predecessor
        • Significantly improved math and multimodal reasoning (AIME +24.2%, MMMU-Pro +9.8%)
        • Deep personalization via Memory Sources with full user control over cited context

        Weaknesses

        • Premium pricing ($5/$30 per 1M tokens) — significantly above market average ($1.71/$8.70)
        • Still outperformed by Claude Opus 4.7 on long-form prose writing and detailed reasoning tasks
        • Citation behavior shifted heavily toward Reddit (3× most-cited) with brand site citations halved to 6%

        Competitor Comparison

        ModelPrice
        GPT-5.3 Instant (predecessor)$5/$30 per 1M
        Claude Opus 4.7 (Anthropic)Higher (exact pricing not disclosed in sources)
        Gemini 3.1 Pro (Google)N/A (competitive on multimodal tasks)

        GPT-5.5 Instant, released May 5, 2026, is OpenAI's latest default model for ChatGPT, replacing GPT-5.3 Instant across all user tiers including the free plan. Positioned as the everyday workhorse of the GPT-5.5 family, it prioritizes low-latency, high-factuality conversational responses over the extended reasoning of OpenAI's dedicated thinking models. The model is available via the API as the chat-latest alias and is integrated into Microsoft 365 Copilot (since May 7, 2026) and GitHub Copilot.

        The headline improvements center on factual reliability: OpenAI reports 52.5% fewer hallucinated claims on high-stakes prompts and 37.3% fewer inaccurate claims on conversations users previously flagged for errors. These claims are supported by meaningful benchmark jumps — AIME 2025 went from 65.4% to 81.2%, MMMU-Pro from 69.2% to 76%, and HealthBench Professional from 32.9 to 38.4. The model also introduced real-time self-correction mid-response, where it can catch and fix its own reasoning errors before completing an answer.

        Beyond raw intelligence, GPT-5.5 Instant's most distinctive innovation is its Memory Sources personalization system. The model draws on past conversations, uploaded files, and connected Gmail (for Plus/Pro users) to provide contextually tailored responses, then transparently shows which sources were referenced so users can edit or remove them. Combined with a 30.2% reduction in word count and fewer unnecessary formatting elements, the model represents OpenAI's push toward making AI feel less like a tool and more like a persistent, trusted assistant. However, at $5/$30 per million tokens via API, it remains premium-priced, and independent tests suggest its prose quality still trails Anthropic's Claude Opus 4.7.

        Analysis generated: 2026-07-17