OpenAI에서 개발한 고속 기반 모델. 낮은 지연 시간과 높은 처리량을 동시에 구현하며, GPT-5.6 Luna의 전작에 해당합니다.
파라미터
Undisclosed
컨텍스트
라이선스
Proprietary
출시일
2026-04-01
일본어 처리 능력
Multilingual model with strong Japanese language processing capabilities.
API 가격
이 모델의 API 가격 정보는 현재 공개되지 않았습니다
강점
약점
활용 사례
심층 분석
Artificial Analysis Intelligence Index
34
#77 of 186 models; above average (median: 30)
GPQA Diamond
84.6%
#84 across all tracked models
AIME 2025
81.2%
Up from 65.4% on GPT-5.3 Instant
MMMU-Pro (Multimodal Reasoning)
76%
Up from 69.2% on GPT-5.3 Instant
Input Price
$5.00/1M tokens
Cache hit: $0.50/1M (90% discount)
Output Price
$30.00/1M tokens
Blended 3:1 ratio: $11.25/1M
Hallucination Reduction
52.5% fewer
vs GPT-5.3 Instant on high-stakes prompts
Context Window
400K tokens
~600 pages; 128K in ChatGPT UI
강점
- ・52.5% fewer hallucinations on high-stakes medical, legal, and financial prompts vs predecessor
- ・Significantly improved math and multimodal reasoning (AIME +24.2%, MMMU-Pro +9.8%)
- ・Deep personalization via Memory Sources with full user control over cited context
약점
- ・Premium pricing ($5/$30 per 1M tokens) — significantly above market average ($1.71/$8.70)
- ・Still outperformed by Claude Opus 4.7 on long-form prose writing and detailed reasoning tasks
- ・Citation behavior shifted heavily toward Reddit (3× most-cited) with brand site citations halved to 6%
경쟁사 비교
| Model | Price |
|---|---|
| GPT-5.3 Instant (predecessor) | $5/$30 per 1M |
| Claude Opus 4.7 (Anthropic) | Higher (exact pricing not disclosed in sources) |
| Gemini 3.1 Pro (Google) | N/A (competitive on multimodal tasks) |
GPT-5.5 Instant, released May 5, 2026, is OpenAI's latest default model for ChatGPT, replacing GPT-5.3 Instant across all user tiers including the free plan. Positioned as the everyday workhorse of the GPT-5.5 family, it prioritizes low-latency, high-factuality conversational responses over the extended reasoning of OpenAI's dedicated thinking models. The model is available via the API as the chat-latest alias and is integrated into Microsoft 365 Copilot (since May 7, 2026) and GitHub Copilot.
The headline improvements center on factual reliability: OpenAI reports 52.5% fewer hallucinated claims on high-stakes prompts and 37.3% fewer inaccurate claims on conversations users previously flagged for errors. These claims are supported by meaningful benchmark jumps — AIME 2025 went from 65.4% to 81.2%, MMMU-Pro from 69.2% to 76%, and HealthBench Professional from 32.9 to 38.4. The model also introduced real-time self-correction mid-response, where it can catch and fix its own reasoning errors before completing an answer.
Beyond raw intelligence, GPT-5.5 Instant's most distinctive innovation is its Memory Sources personalization system. The model draws on past conversations, uploaded files, and connected Gmail (for Plus/Pro users) to provide contextually tailored responses, then transparently shows which sources were referenced so users can edit or remove them. Combined with a 30.2% reduction in word count and fewer unnecessary formatting elements, the model represents OpenAI's push toward making AI feel less like a tool and more like a persistent, trusted assistant. However, at $5/$30 per million tokens via API, it remains premium-priced, and independent tests suggest its prose quality still trails Anthropic's Claude Opus 4.7.
출처
- GPT-5.5 Instant: smarter, clearer, and more personalized — OpenAI Official
- GPT-5.5 Instant (May 2026) — Intelligence, Performance & Price Analysis — Artificial Analysis
- GPT-5.5 Instant Benchmarks, Pricing & Context Window — LLM Stats
- GPT-5.5 Instant (May 2026) — Easy Benchmarks
- GPT-5.5 Instant (May 2026) by OpenAI — Sophon
- I tested OpenAI's three claims about GPT-5.5 Instant, and only one fully held up — The New Stack
- I tested GPT-5.5 Instant — and it finally stopped overexplaining everything — Tom's Guide
- GPT-5.5 Instant Review: I Tested It for 3 Days — PrimeAIcenter
- GPT-5.5 Instant: ChatGPT's Default Model (2026) — AI/TLDR
- GPT-5.5 Instant vs GPT-5.3: Three OpenAI Claims Tested — Pickuma
분석 생성일: 2026-07-17