OpenAI독점

GPT-5.5 Instant

이 모델 비교

OpenAI에서 개발한 고속 기반 모델. 낮은 지연 시간과 높은 처리량을 동시에 구현하며, GPT-5.6 Luna의 전작에 해당합니다.

파라미터

Undisclosed

컨텍스트

라이선스

Proprietary

출시일

2026-04-01

일본어 처리 능력

High-Quality JP

Multilingual model with strong Japanese language processing capabilities.

API 가격

이 모델의 API 가격 정보는 현재 공개되지 않았습니다

강점

    약점

      활용 사례

        심층 분석

        Artificial Analysis Intelligence Index

        34

        #77 of 186 models; above average (median: 30)

        GPQA Diamond

        84.6%

        #84 across all tracked models

        AIME 2025

        81.2%

        Up from 65.4% on GPT-5.3 Instant

        MMMU-Pro (Multimodal Reasoning)

        76%

        Up from 69.2% on GPT-5.3 Instant

        Input Price

        $5.00/1M tokens

        Cache hit: $0.50/1M (90% discount)

        Output Price

        $30.00/1M tokens

        Blended 3:1 ratio: $11.25/1M

        Hallucination Reduction

        52.5% fewer

        vs GPT-5.3 Instant on high-stakes prompts

        Context Window

        400K tokens

        ~600 pages; 128K in ChatGPT UI

        강점

        • 52.5% fewer hallucinations on high-stakes medical, legal, and financial prompts vs predecessor
        • Significantly improved math and multimodal reasoning (AIME +24.2%, MMMU-Pro +9.8%)
        • Deep personalization via Memory Sources with full user control over cited context

        약점

        • Premium pricing ($5/$30 per 1M tokens) — significantly above market average ($1.71/$8.70)
        • Still outperformed by Claude Opus 4.7 on long-form prose writing and detailed reasoning tasks
        • Citation behavior shifted heavily toward Reddit (3× most-cited) with brand site citations halved to 6%

        경쟁사 비교

        ModelPrice
        GPT-5.3 Instant (predecessor)$5/$30 per 1M
        Claude Opus 4.7 (Anthropic)Higher (exact pricing not disclosed in sources)
        Gemini 3.1 Pro (Google)N/A (competitive on multimodal tasks)

        GPT-5.5 Instant, released May 5, 2026, is OpenAI's latest default model for ChatGPT, replacing GPT-5.3 Instant across all user tiers including the free plan. Positioned as the everyday workhorse of the GPT-5.5 family, it prioritizes low-latency, high-factuality conversational responses over the extended reasoning of OpenAI's dedicated thinking models. The model is available via the API as the chat-latest alias and is integrated into Microsoft 365 Copilot (since May 7, 2026) and GitHub Copilot.

        The headline improvements center on factual reliability: OpenAI reports 52.5% fewer hallucinated claims on high-stakes prompts and 37.3% fewer inaccurate claims on conversations users previously flagged for errors. These claims are supported by meaningful benchmark jumps — AIME 2025 went from 65.4% to 81.2%, MMMU-Pro from 69.2% to 76%, and HealthBench Professional from 32.9 to 38.4. The model also introduced real-time self-correction mid-response, where it can catch and fix its own reasoning errors before completing an answer.

        Beyond raw intelligence, GPT-5.5 Instant's most distinctive innovation is its Memory Sources personalization system. The model draws on past conversations, uploaded files, and connected Gmail (for Plus/Pro users) to provide contextually tailored responses, then transparently shows which sources were referenced so users can edit or remove them. Combined with a 30.2% reduction in word count and fewer unnecessary formatting elements, the model represents OpenAI's push toward making AI feel less like a tool and more like a persistent, trusted assistant. However, at $5/$30 per million tokens via API, it remains premium-priced, and independent tests suggest its prose quality still trails Anthropic's Claude Opus 4.7.

        분석 생성일: 2026-07-17