Moonshot AIProprietary

Kimi K3

Compare this model

Kimi K3 is Moonshot AI's open-weight flagship, released July 16, 2026 — the largest open-weight model to date at 2.8 trillion parameters (104B active per token across 896 experts) with a 1,000,000-token context window. Licensed under the modified MIT-style Kimi K3 License, it runs native multimodal input (text, image, video) with always-on thinking and leads the open-weights field at 57 on the Artificial Analysis Intelligence Index (rank 4 of 189). It tops frontend coding (Frontend Code Arena 1,679 Elo, #1 globally) and scores 88.3% on Terminal-Bench 2.1, 93.5% on GPQA Diamond, and 91.2% on BrowseComp. API pricing is $3 per million input (cache hit $0.30) and $15 per million output tokens.

Parameters

Undisclosed

Context Window

1M

License

Proprietary

Release Date

2026-07-16

Japanese Language Capability

🌐Multilingual

General multilingual model. Basic Japanese processing is possible, but inferior to specialized models.

API Pricing

Input Price (per 1M tokens)

$3

Output Price (per 1M tokens)

$15

Billing Mode: standard

Strengths

  • Legendary high parameter count (over 2.5T parameters)
  • Extended context length (over 1M tokens)
  • High expectations as a next-gen flagship model

Weaknesses

  • Lack of official information
  • Uncertain release date based on rumors
  • Missing licensing and evaluation benchmarks

Use Cases

  • General text generation
  • Reasoning tasks
  • Natural language understanding

Deep Analysis

Arena Elo

1679

#1 in Frontend Code Arena

SWE-Bench Verified

~80%

Est. from coding benchmarks, vs Fable 5 ~87.6%

Input Price (cache-miss)

$3.00/1M

vs GPT-5.6 Sol: $5.00/1M

Output Price

$15.00/1M

vs Claude Fable 5: $50.00/1M

Parameters

2.8T

Largest open-weight model to date

Context Window

1M tokens

Flat pricing, no length tiers

Strengths

  • World's largest open-weight model (2.8T params) with near-frontier performance.
  • Exceptional long-horizon coding and agentic task performance (SWE Marathon, BrowseComp leader).
  • Competitive pricing (~40-50% cheaper than GPT-5.6 Sol on output).

Weaknesses

  • Independent third-party benchmarks still pending; some metrics lag top proprietary models.
  • Higher hallucination rate (51%) reported in independent testing vs predecessor.
  • Inference at scale requires substantial hardware; not a single-server deployment.

Competitor Comparison

ModelArenaSWEGPQAPrice
Claude Fable 5~1800+ (est.)~87.6% (Opus 4.7 verified)92.6%$10/$50
GPT-5.6 Sol~1750 (est.)~80%94.1%$5/$30
Claude Opus 4.81600~80%91.0%~$15/$75 (est.)

Kimi K3, released by Moonshot AI on July 16, 2026, is a landmark open-weight model and the first to cross the 2.8-trillion parameter threshold. Built on novel Kimi Delta Attention (KDA) and Attention Residuals (AttnRes) architecture with a 1-million-token context window, it targets frontier-level long-horizon coding, agentic workflows, and knowledge work. While Moonshot's own benchmarks show it trailing the very top proprietary models (Claude Fable 5, GPT-5.6 Sol), K3 consistently outperforms all other tested systems, including Claude Opus 4.8 and GPT-5.5. Its launch signals the functional closing of the capability gap between the best open-weight and closed-source models, offering enterprises a high-performance, self-hostable alternative at aggressive pricing ($3/$15 per million tokens).

Kimi K3's strategic release, timed just before the 2026 World AI Conference, represents a major technical and business escalation. The architecture innovations deliver a reported 2.5x improvement in scaling efficiency over its predecessor, enabling strong performance in sustained multi-hour engineering tasks (leading SWE Marathon) and complex information retrieval (state-of-the-art on BrowseComp). The full open-weight release, expected by July 27, will allow the community to verify these claims and fine-tune the model for specific domains. However, the model's higher hallucination rate and the sheer infrastructure required for self-hosting present practical considerations for adopters.

Pricing marks a significant shift from Kimi's earlier ultra-cheap models. K3 is priced at parity with Western mid-tier offerings like Claude Sonnet 5, undercutting top-tier models like GPT-5.6 Sol by 40-50%. This positions Kimi K3 as a value-oriented frontier model, especially for cost-sensitive, high-volume coding and research workloads where its flat-rate, million-token context window provides a distinct economic advantage over tiered-pricing competitors.

Analysis generated: 2026-07-17