Grok 4.6은 SpaceXAI가 2026년 8월 12일 발표한 최상위 모델이다. Grok 4.5에서 보강 학습과 에이전트 강화학습을 확장했으며, Artificial Analysis Intelligence Index에서 61점(GPT-5.6 Sol Max와 동점)을 기록했다. 장시간 실행 에이전트와 '아이디어에서 제품으로' 워크플로에 집중하며 API 가격은 100만 토큰당 $2/$6이다.
파라미터
Undisclosed
컨텍스트
TBA
라이선스
Proprietary
출시일
2026-08-12
벤치마크 성능
AA Intelligence Index
61.0
LMArena Elo
1753.0
HLE
—
ARC-AGI-2
—
SWE-bench Verified
—
GPQA Diamond
—
MMLU-Pro
—
LiveCodeBench
—
AIME 2025
—
MATH-500
—
API 가격
입력 가격 (1M 토큰당)
$2
출력 가격 (1M 토큰당)
$6
과금 모드: standard
강점
- •Artificial Analysis Intelligence Index 61점으로 최상위 모델과 동수준이면서 절반 이하 가격
- •강력한 장시간 실행 에이전트能力与 자체 검증 및 다단계 작업 수행
- •GDPVal-AA v2 Elo 1753점(1위)
- •API·OpenRouter·Vercel·Cloudflare·Cursor·Grok Build에서 폭넓게 이용 가능
약점
- •몇몇 고난도 코딩 벤치마크에서는 GPT-5.6 Sol이나 Fable 5에 뒤처짐
- •컨텍스트 윈도우 길이가 아직 공개되지 않음
- •기존과 마찬가지로 SpaceXAI 콘텐츠·안전 정책에 종속
활용 사례
- •에이전트형 코딩 지원 및 장기 소프트웨어 엔지니어링
- •자율 리서치 및 데이터 분석 에이전트
- •제품 기획안에서 실행 가능한 앱으로 시제품 제작
심층 분석
Artificial Analysis Intelligence Index
61
Tied with GPT-5.6 Sol Max; up from Grok 4.5 High's 56
GDPVal-AA v2 Elo
1753
#1, ahead of GPT-5.6 Sol Max (1728) and Claude Fable 5 (1741)
Input / Output Price
$2 / $6 per 1M
Half the cost of most frontier APIs; high-speed tier at double
Context Window
TBA
Not yet disclosed by SpaceXAI
Release
Aug 12, 2026
Available via API, OpenRouter, Vercel, Cloudflare, Cursor, Grok Build
강점
- ・Matches the top frontier models on the Artificial Analysis Intelligence Index (61) at roughly half the price of GPT-5.6 Sol or Claude Fable 5
- ・Strong long-horizon agent capability: self-testing, verification, and multi-step task execution for research, data analysis, and coding
- ・Competitive on agentic coding and knowledge-work benchmarks (GDPVal-AA v2 #1, plus strong showings on CursorBench v3.2 and DeepSWE v1.1)
- ・Wide availability across API, OpenRouter, Vercel, Cloudflare, Cursor, and Grok Build with first-week double quotas
약점
- ・Still trails GPT-5.6 Sol and Claude Fable 5 on several hard coding benchmarks despite the agentic focus
- ・Context window length not yet disclosed, leaving maximum task scope unclear
- ・Like its predecessors, subject to SpaceXAI's content and safety policy constraints
경쟁사 비교
| Model | Price |
|---|---|
| Claude Fable 5 (max) | $10/$50 |
| GPT-5.6 Sol (max) | $5/$30 |
| Grok 4.5 (high) | $2/$6 |
Grok 4.6 is SpaceXAI's upgraded frontier model, launched August 12, 2026. It extends Grok 4.5 with longer supplementary training and agentic reinforcement learning, pushing its Artificial Analysis Intelligence Index to 61 — tied with GPT-5.6 Sol Max — while keeping pricing at a disruptive $2/$6 per million tokens. The release leans hard into long-running agents and idea-to-product workflows.
출처
분석 생성일: 2026-08-13