Kimi K3 is Moonshot AI's open-weight flagship, released July 16, 2026 — the largest open-weight model to date at 2.8 trillion parameters (104B active per token across 896 experts) with a 1,000,000-token context window. Licensed under the modified MIT-style Kimi K3 License, it runs native multimodal input (text, image, video) with always-on thinking and leads the open-weights field at 57 on the Artificial Analysis Intelligence Index (rank 4 of 189). It tops frontend coding (Frontend Code Arena 1,679 Elo, #1 globally) and scores 88.3% on Terminal-Bench 2.1, 93.5% on GPQA Diamond, and 91.2% on BrowseComp. API pricing is $3 per million input (cache hit $0.30) and $15 per million output tokens.
Parameters
Undisclosed
Context Window
1M
License
Proprietary
Release Date
2026-07-16
Japanese Language Capability
General multilingual model. Basic Japanese processing is possible, but inferior to specialized models.
API Pricing
Input Price (per 1M tokens)
$3
Output Price (per 1M tokens)
$15
Billing Mode: standard
Strengths
- •Legendary high parameter count (over 2.5T parameters)
- •Extended context length (over 1M tokens)
- •High expectations as a next-gen flagship model
Weaknesses
- •Lack of official information
- •Uncertain release date based on rumors
- •Missing licensing and evaluation benchmarks
Use Cases
- •General text generation
- •Reasoning tasks
- •Natural language understanding
Deep Analysis
Arena Elo
1679
#1 in Frontend Code Arena
SWE-Bench Verified
~80%
Est. from coding benchmarks, vs Fable 5 ~87.6%
Input Price (cache-miss)
$3.00/1M
vs GPT-5.6 Sol: $5.00/1M
Output Price
$15.00/1M
vs Claude Fable 5: $50.00/1M
Parameters
2.8T
Largest open-weight model to date
Context Window
1M tokens
Flat pricing, no length tiers
Strengths
- ・World's largest open-weight model (2.8T params) with near-frontier performance.
- ・Exceptional long-horizon coding and agentic task performance (SWE Marathon, BrowseComp leader).
- ・Competitive pricing (~40-50% cheaper than GPT-5.6 Sol on output).
Weaknesses
- ・Independent third-party benchmarks still pending; some metrics lag top proprietary models.
- ・Higher hallucination rate (51%) reported in independent testing vs predecessor.
- ・Inference at scale requires substantial hardware; not a single-server deployment.
Competitor Comparison
| Model | Arena | SWE | GPQA | Price |
|---|---|---|---|---|
| Claude Fable 5 | ~1800+ (est.) | ~87.6% (Opus 4.7 verified) | 92.6% | $10/$50 |
| GPT-5.6 Sol | ~1750 (est.) | ~80% | 94.1% | $5/$30 |
| Claude Opus 4.8 | 1600 | ~80% | 91.0% | ~$15/$75 (est.) |
Kimi K3, released by Moonshot AI on July 16, 2026, is a landmark open-weight model and the first to cross the 2.8-trillion parameter threshold. Built on novel Kimi Delta Attention (KDA) and Attention Residuals (AttnRes) architecture with a 1-million-token context window, it targets frontier-level long-horizon coding, agentic workflows, and knowledge work. While Moonshot's own benchmarks show it trailing the very top proprietary models (Claude Fable 5, GPT-5.6 Sol), K3 consistently outperforms all other tested systems, including Claude Opus 4.8 and GPT-5.5. Its launch signals the functional closing of the capability gap between the best open-weight and closed-source models, offering enterprises a high-performance, self-hostable alternative at aggressive pricing ($3/$15 per million tokens).
Kimi K3's strategic release, timed just before the 2026 World AI Conference, represents a major technical and business escalation. The architecture innovations deliver a reported 2.5x improvement in scaling efficiency over its predecessor, enabling strong performance in sustained multi-hour engineering tasks (leading SWE Marathon) and complex information retrieval (state-of-the-art on BrowseComp). The full open-weight release, expected by July 27, will allow the community to verify these claims and fine-tune the model for specific domains. However, the model's higher hallucination rate and the sheer infrastructure required for self-hosting present practical considerations for adopters.
Pricing marks a significant shift from Kimi's earlier ultra-cheap models. K3 is priced at parity with Western mid-tier offerings like Claude Sonnet 5, undercutting top-tier models like GPT-5.6 Sol by 40-50%. This positions Kimi K3 as a value-oriented frontier model, especially for cost-sensitive, high-volume coding and research workloads where its flat-rate, million-token context window provides a distinct economic advantage over tiered-pricing competitors.
Sources
- Kimi K3 Tech Blog: Open Frontier Intelligence
- Kimi's open model K3 nears GPT-5.6 Sol and Fable 5 while signaling the end of super cheap Chinese AI
- Kimi K3 Beats Fable 5, GPT 5.6 On Some Benchmarks In Frontier-Level Performance
- Kimi K3: Moonshot's 2.8T Open-Weight Model Explained
- Moonshot AI Releases Kimi K3: A 2.8 Trillion Parameter Open MoE Model With Kimi Delta Attention and 1M Context
- Kimi K3 vs Gemini 3.1 Pro: Open-Weight Coding Value vs Natively Multimodal Reasoning
- Kimi K3 vs GPT-5.6 Sol Pro: Full 2026 Comparison
- Kimi K3 vs Claude: 2.8T Open Model vs Opus 4.8
- China’s Moonshot AI releases Kimi K3, the largest open-source model ever, rivaling top U.S. systems
- Moonshot’s Kimi K3 pushes Chinese AI into Fable-level territory
Analysis generated: 2026-07-17