Skip to content
All systems operational0 AI providers monitored, polled every 2 minutes
Live status

Grok 4.6

Flagship

by xAI

Grok 4.6 shipped on August 7, 2026 on the same 1.5 trillion parameter V9 foundation as Grok 4.5, with the entire gain coming from improved supervised fine-tuning and reinforcement learning rather than a larger model. The context window is 500,000 tokens. Pricing has a wrinkle worth budgeting for: prompts below 200,000 tokens bill at $2 per million input, $0.50 per million cached input, and $6 per million output, but once a request crosses 200,000 tokens those rates double to $4, $1, and $12, and the higher tier applies to every token in that request rather than only the overage. Do not extrapolate the $2/$6 headline across the full 500K window when you model cost of ownership. On capability, Artificial Analysis places Grok 4.6 above Kimi K3 and level with GPT-5.6 Sol, which puts it third overall on that index at a fraction of Sol's $5/$30 rate. xAI positioned it directly against Kimi K3, at roughly 2.8 trillion parameters, and Claude Opus 4.8, and the post-training-only approach is the argument: xAI is betting that a well-tuned 1.5T model beats a larger one at the same latency and cost envelope. A larger 2.1 trillion parameter Grok 4.7 has been signaled to follow within weeks, so treat 4.6 as a checkpoint rather than a plateau. Available through the xAI API, Grok in X, and SuperGrok.

Input Price

$2.00

per 1M tokens

Output Price

$6.00

per 1M tokens

Context Window

500K

tokens

Released

2026-08

API access

Capabilities

textvisiontool-usecodereasoning

Key Strengths

  • $2/$6 per 1M tokens below 200K, matching Grok 4.5 rates
  • Third on the Artificial Analysis index, level with GPT-5.6 Sol
  • 500K token context window
  • Post-training-only upgrade keeps 4.5 throughput and latency
  • Cached input at $0.50 per 1M below the 200K threshold

Best For

  • Reasoning and analysis at well under frontier flagship pricing
  • Real-time work that benefits from X platform integration
  • Coding and tool use where latency matters
  • Workloads that stay comfortably under the 200K token pricing threshold

Pricing Details

Input tokens

$2.00

per 1M tokens

Output tokens

$6.00

per 1M tokens

Estimated cost per 1K requests

$5.00

~1K input + ~500 output tokens avg

Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.

Related Models

View DocumentationCompare ModelsCost CalculatorFull Pricing Guide