Thursday Aug 13

Grok 4.6 Launches At Half Price

13AUG
RIVALSGROK $2

Frontier AI just got a price war. xAI shipped Grok 4.6 today at $2 per million tokens, half of rival models. Coding and agent work keeps getting cheaper, fast.

Grok 4.6 launched August 13 with a 1753 Elo score and a 500K-token context window. It undercuts GPT and Claude pricing by half.

xAI is racing on price, not just benchmarks. Grok 4.7, a bigger 2.1-trillion-parameter model, is coming within weeks. Grok 5 is targeted before year-end.

Cheaper frontier models make agentic workflows viable at higher volume. Whoever wins on cost per task, not raw score, wins enterprise budgets.

full brief & sources

⚡ Why this matters

  • Price, not just benchmark score, is becoming the main lever labs compete on.
  • Cheaper frontier-grade models make it viable to run agents at high volume in production.
  • A rapid release cadence resets how fast 'frontier' churns.

🔍 What happened

  • Aug 13: xAI released Grok 4.6, scoring 1753 Elo.
  • Priced at $2 per million input tokens, roughly half of comparable frontier models.
  • 500K-token context window, text and image input, text-only output.
  • Built for coding, agentic tasks, and knowledge work.
  • Grok 4.7 (2.1 trillion parameters) expected within weeks; Grok 5 targeted before year-end 2026.

💬 Smart takes

  • xAI: positions 4.6 as the cost leader for agentic and coding workloads.
  • Skeptic: Elo leaderboard scores move fast and rarely predict which model wins real enterprise contracts.

🧭 Where this goes

  1. LikelyOpenAI and Anthropic respond with their own price cuts within the quarter.
  2. Likelyagent-heavy startups default to whichever model is cheapest per task, not highest-scoring.
  3. PossibleGrok 4.7 ships before rivals finish responding to 4.6's pricing.
  4. Wild Cardthe price war compresses margins enough that a smaller lab exits the frontier race entirely.

🥄 The Spoon Take

The model race just became a price race too. xAI is betting cheap-and-fast beats smart-and-expensive for the agent workloads companies actually run at scale. If that bet is right, benchmark leaderboards stop mattering as much as the invoice.

🤔 Pushback

Elo scores are self-reported and gameable; the real test is whether enterprises actually switch, not whether xAI wins a leaderboard.