A high-speed foundation model developed by OpenAI. It balances low latency with high throughput and is the predecessor to GPT-5.6 Luna.
Parameters
Undisclosed
Context Window
License
Proprietary
Release Date
2026-04-01
Japanese Language Capability
Multilingual model with strong Japanese language processing capabilities.
API Pricing
API pricing for this model is not yet available
Strengths
Weaknesses
Use Cases
Deep Analysis
Artificial Analysis Intelligence Index
34
#77 of 186 models; above average (median: 30)
GPQA Diamond
84.6%
#84 across all tracked models
AIME 2025
81.2%
Up from 65.4% on GPT-5.3 Instant
MMMU-Pro (Multimodal Reasoning)
76%
Up from 69.2% on GPT-5.3 Instant
Input Price
$5.00/1M tokens
Cache hit: $0.50/1M (90% discount)
Output Price
$30.00/1M tokens
Blended 3:1 ratio: $11.25/1M
Hallucination Reduction
52.5% fewer
vs GPT-5.3 Instant on high-stakes prompts
Context Window
400K tokens
~600 pages; 128K in ChatGPT UI
Strengths
- ・52.5% fewer hallucinations on high-stakes medical, legal, and financial prompts vs predecessor
- ・Significantly improved math and multimodal reasoning (AIME +24.2%, MMMU-Pro +9.8%)
- ・Deep personalization via Memory Sources with full user control over cited context
Weaknesses
- ・Premium pricing ($5/$30 per 1M tokens) — significantly above market average ($1.71/$8.70)
- ・Still outperformed by Claude Opus 4.7 on long-form prose writing and detailed reasoning tasks
- ・Citation behavior shifted heavily toward Reddit (3× most-cited) with brand site citations halved to 6%
Competitor Comparison
| Model | Price |
|---|---|
| GPT-5.3 Instant (predecessor) | $5/$30 per 1M |
| Claude Opus 4.7 (Anthropic) | Higher (exact pricing not disclosed in sources) |
| Gemini 3.1 Pro (Google) | N/A (competitive on multimodal tasks) |
GPT-5.5 Instant, released May 5, 2026, is OpenAI's latest default model for ChatGPT, replacing GPT-5.3 Instant across all user tiers including the free plan. Positioned as the everyday workhorse of the GPT-5.5 family, it prioritizes low-latency, high-factuality conversational responses over the extended reasoning of OpenAI's dedicated thinking models. The model is available via the API as the chat-latest alias and is integrated into Microsoft 365 Copilot (since May 7, 2026) and GitHub Copilot.
The headline improvements center on factual reliability: OpenAI reports 52.5% fewer hallucinated claims on high-stakes prompts and 37.3% fewer inaccurate claims on conversations users previously flagged for errors. These claims are supported by meaningful benchmark jumps — AIME 2025 went from 65.4% to 81.2%, MMMU-Pro from 69.2% to 76%, and HealthBench Professional from 32.9 to 38.4. The model also introduced real-time self-correction mid-response, where it can catch and fix its own reasoning errors before completing an answer.
Beyond raw intelligence, GPT-5.5 Instant's most distinctive innovation is its Memory Sources personalization system. The model draws on past conversations, uploaded files, and connected Gmail (for Plus/Pro users) to provide contextually tailored responses, then transparently shows which sources were referenced so users can edit or remove them. Combined with a 30.2% reduction in word count and fewer unnecessary formatting elements, the model represents OpenAI's push toward making AI feel less like a tool and more like a persistent, trusted assistant. However, at $5/$30 per million tokens via API, it remains premium-priced, and independent tests suggest its prose quality still trails Anthropic's Claude Opus 4.7.
Sources
- GPT-5.5 Instant: smarter, clearer, and more personalized — OpenAI Official
- GPT-5.5 Instant (May 2026) — Intelligence, Performance & Price Analysis — Artificial Analysis
- GPT-5.5 Instant Benchmarks, Pricing & Context Window — LLM Stats
- GPT-5.5 Instant (May 2026) — Easy Benchmarks
- GPT-5.5 Instant (May 2026) by OpenAI — Sophon
- I tested OpenAI's three claims about GPT-5.5 Instant, and only one fully held up — The New Stack
- I tested GPT-5.5 Instant — and it finally stopped overexplaining everything — Tom's Guide
- GPT-5.5 Instant Review: I Tested It for 3 Days — PrimeAIcenter
- GPT-5.5 Instant: ChatGPT's Default Model (2026) — AI/TLDR
- GPT-5.5 Instant vs GPT-5.3: Three OpenAI Claims Tested — Pickuma
Analysis generated: 2026-07-17