xAIProprietary

Grok 4.5

Compare this model

Grok 4.5 is xAI's Opus-class coding and agentic flagship, released July 8, 2026. It pairs a 500,000-token context window with configurable low/medium/high reasoning effort, tool calling, web and X search, and code execution. Priced at $2 per million input and $6 per million output tokens (rising to $4/$12 above 200K tokens), it undercuts Western frontier APIs while matching Opus-class quality — scoring ~54 on the Artificial Analysis Intelligence Index, 72.4 on the coding index, 64.7% on SWE-Bench Pro, and 83.3% on Terminal-Bench 2.1. It accepts text and image input, streams text, and is the default model in Grok Build and Cursor.

Parameters

Undisclosed

Context Window

500K

License

Proprietary

Release Date

2026-07-08

Japanese Language Capability

High-Quality JP

Multilingual model with strong Japanese language processing capabilities.

API Pricing

Input Price (per 1M tokens)

$2

Output Price (per 1M tokens)

$6

Billing Mode: standard

Strengths

  • Large-scale foundation model for general-purpose tasks.
  • Backed by xAI, a company with significant AI research resources.
  • Designed for broad applicability.

Weaknesses

  • Detailed performance metrics and benchmarks may not be publicly available.
  • Potential competition with other well-established foundation models.
  • Specific strengths and weaknesses are not well-documented publicly.

Use Cases

  • General text generation and understanding.
  • Serving as a base for fine-tuning on specific downstream tasks.
  • AI research and experimentation.

Deep Analysis

Artificial Analysis Intelligence Index

54

#4 of 168 models; behind Fable 5 (60), Opus 4.8 (56), GPT-5.5 (55)

Agentic Tool Use (τ³-Banking)

33%

#1 overall, ahead of GPT-5.5 (31%) and Claude Sonnet 4.6 (31%)

Coding Agent Index (Grok Build)

76

Tied with GPT-5.5 (Codex), 1 point behind Fable 5 (Claude Code)

Input / Output Price

$2.00 / $6.00 per 1M

60-75% cheaper than Opus 4.8 and GPT-5.5

Intelligence Index Task Cost

$0.31 per task

5x cheaper than Claude Sonnet 5 (max) with higher score

Token Efficiency

1.9M tokens/task

vs 6.2M (GPT-5.5 Codex) and 7.2M (Fable 5 Claude Code)

Strengths

  • Best-in-class agentic tool use (#1 on τ³-Banking) with exceptional token efficiency across agent workflows
  • Dramatically cheaper than frontier peers: $2/$6 per 1M tokens vs $5-$10 input and $25-$50 output for competitors
  • Fast output (~80-86 tokens/sec) and highly concise, completing tasks with 4x fewer tokens than comparable models

Weaknesses

  • Not the smartest model available — #4 on Intelligence Index; trails Fable 5 and Opus 4.8 on raw reasoning and harder coding benchmarks
  • Hallucination rate rose to 54% (up from 25% on Grok 4.3) as accuracy improved — more confident when wrong
  • Context window reduced to 500K tokens (down from Grok 4.3's 1M); no batch discount at launch; EU availability delayed

Competitor Comparison

ModelPrice
Claude Fable 5 (max)$10/$50
Claude Opus 4.8 (max)$5/$25
GPT-5.5 (xhigh)$5/$30

Grok 4.5, released July 8, 2026 by SpaceXAI (the merged SpaceX/xAI entity), represents a significant leap forward for xAI's model lineup, jumping 16 points on the Artificial Analysis Intelligence Index from Grok 4.3's score of 38 to 54, placing it fourth overall behind only Claude Fable 5, Claude Opus 4.8, and GPT-5.5. The model was jointly trained with Cursor on trillions of tokens of real developer-agent interaction data, following SpaceX's $60 billion acquisition of Anysphere (Cursor's maker) in June 2026. It is a 1.5 trillion parameter mixture-of-experts model (per Musk's disclosure, not officially confirmed by SpaceXAI) trained on tens of thousands of NVIDIA GB300 GPUs.

Grok 4.5's positioning is deliberate: it does not claim to be the absolute smartest model, but rather the best intelligence-per-dollar option in the near-frontier tier. Its standout achievement is the #1 ranking on agentic tool use (τ³-Banking at 33%), and its token efficiency is remarkable — averaging just 1.9 million tokens per coding agent task compared to 6.2-7.2 million for competitors. At $2/$6 per 1M input/output tokens, it costs 60-75% less than Claude Opus 4.8 or GPT-5.5, while completing tasks at roughly $2.49 per Coding Agent Index task versus $5.07 for GPT-5.5 and $11.80 for Fable 5.

The model ships as the default in Grok Build (xAI's coding agent CLI), is available across all Cursor plans, and is integrated into Microsoft Office add-ins for Word, PowerPoint, and Excel. Access is also available through OpenRouter, Vercel, Cloudflare, Snowflake, and Databricks Mosaic. The trade-offs are real: it trails the top models on raw coding accuracy (64.7% on SWE-Bench Pro vs. Fable 5's 80.4%), its hallucination rate has increased, and it reduced the context window from 1M to 500K tokens. But for high-volume agentic workloads where cost and speed matter as much as peak accuracy, Grok 4.5 has established itself as the most compelling value proposition on the market.

Analysis generated: 2026-07-17

Related Comparisons