SpaceXAI독점

Grok 4.6

이 모델 비교

Grok 4.6은 SpaceXAI가 2026년 8월 12일 발표한 최상위 모델이다. Grok 4.5에서 보강 학습과 에이전트 강화학습을 확장했으며, Artificial Analysis Intelligence Index에서 61점(GPT-5.6 Sol Max와 동점)을 기록했다. 장시간 실행 에이전트와 '아이디어에서 제품으로' 워크플로에 집중하며 API 가격은 100만 토큰당 $2/$6이다.

파라미터

Undisclosed

컨텍스트

TBA

라이선스

Proprietary

출시일

2026-08-12

벤치마크 성능

AA Intelligence Index

61.0

LMArena Elo

1753.0

HLE

ARC-AGI-2

SWE-bench Verified

GPQA Diamond

MMLU-Pro

LiveCodeBench

AIME 2025

MATH-500

API 가격

입력 가격 (1M 토큰당)

$2

출력 가격 (1M 토큰당)

$6

과금 모드: standard

강점

  • Artificial Analysis Intelligence Index 61점으로 최상위 모델과 동수준이면서 절반 이하 가격
  • 강력한 장시간 실행 에이전트能力与 자체 검증 및 다단계 작업 수행
  • GDPVal-AA v2 Elo 1753점(1위)
  • API·OpenRouter·Vercel·Cloudflare·Cursor·Grok Build에서 폭넓게 이용 가능

약점

  • 몇몇 고난도 코딩 벤치마크에서는 GPT-5.6 Sol이나 Fable 5에 뒤처짐
  • 컨텍스트 윈도우 길이가 아직 공개되지 않음
  • 기존과 마찬가지로 SpaceXAI 콘텐츠·안전 정책에 종속

활용 사례

  • 에이전트형 코딩 지원 및 장기 소프트웨어 엔지니어링
  • 자율 리서치 및 데이터 분석 에이전트
  • 제품 기획안에서 실행 가능한 앱으로 시제품 제작

심층 분석

Artificial Analysis Intelligence Index

61

Tied with GPT-5.6 Sol Max; up from Grok 4.5 High's 56

GDPVal-AA v2 Elo

1753

#1, ahead of GPT-5.6 Sol Max (1728) and Claude Fable 5 (1741)

Input / Output Price

$2 / $6 per 1M

Half the cost of most frontier APIs; high-speed tier at double

Context Window

TBA

Not yet disclosed by SpaceXAI

Release

Aug 12, 2026

Available via API, OpenRouter, Vercel, Cloudflare, Cursor, Grok Build

강점

  • Matches the top frontier models on the Artificial Analysis Intelligence Index (61) at roughly half the price of GPT-5.6 Sol or Claude Fable 5
  • Strong long-horizon agent capability: self-testing, verification, and multi-step task execution for research, data analysis, and coding
  • Competitive on agentic coding and knowledge-work benchmarks (GDPVal-AA v2 #1, plus strong showings on CursorBench v3.2 and DeepSWE v1.1)
  • Wide availability across API, OpenRouter, Vercel, Cloudflare, Cursor, and Grok Build with first-week double quotas

약점

  • Still trails GPT-5.6 Sol and Claude Fable 5 on several hard coding benchmarks despite the agentic focus
  • Context window length not yet disclosed, leaving maximum task scope unclear
  • Like its predecessors, subject to SpaceXAI's content and safety policy constraints

경쟁사 비교

ModelPrice
Claude Fable 5 (max)$10/$50
GPT-5.6 Sol (max)$5/$30
Grok 4.5 (high)$2/$6

Grok 4.6 is SpaceXAI's upgraded frontier model, launched August 12, 2026. It extends Grok 4.5 with longer supplementary training and agentic reinforcement learning, pushing its Artificial Analysis Intelligence Index to 61 — tied with GPT-5.6 Sol Max — while keeping pricing at a disruptive $2/$6 per million tokens. The release leans hard into long-running agents and idea-to-product workflows.

분석 생성일: 2026-08-13