Home / Blog / Grok 4.6
ENGINEERING_BLOG · 2026.07.30

Grok 4.6 Release Date:
Is xAI's August 7 Target Realistic?

Elon Musk says Grok 4.6 is coming around August 7, with a larger Grok 4.7 to follow a few weeks later. He shared the timeline on X on July 28, 2026, in a reply to Vercel CEO Guillermo Rauch — not in an official xAI announcement. No benchmarks, pricing, or model card have been published yet. This article is for developers and technical leads tracking frontier model releases, Agent routing, and token economics. We map the timeline, specs, competitive context, and what remains unverified.

SECTION 01 Three common misreads before Grok 4.6 ships

  • Treating a tweet as an official launch: The only on-record source is Musk's July 28 X reply. xAI's blog and product pages have not confirmed the date or specs;
  • Equating parameter count with capability: Musk emphasized SFT and RL post-training upgrades, not raw scale — consistent with Grok 4.5's coding-first positioning;
  • Ignoring "Musk time": xAI, Tesla, and SpaceX timelines from Musk have historically slipped by days to weeks. Treat "around August 7" as a target, not a guarantee;
  • Mixing rumored rivals with shipped models: Claude Fable 5.1 is not officially confirmed by Anthropic. Compare against shipped products like Kimi K3 and Grok 4.5 with clear source labels;
  • Skipping vendor-risk context: In July 2026, xAI sued a user over alleged CSAM generation via Grok. That does not predict Grok 4.6's technical scores, but it matters for enterprise vendor review.

Bottom line: the Grok 4.6/4.7 roadmap is plausible but unverified by third parties. August 2026 may be the densest release month yet — choose models on token efficiency and real task cost, not parameter headlines alone.

SECTION 02 Timeline and specs: Grok 4.5 through 4.7

Key events from July through September 2026
Date Event
July 8, 2026 xAI ships Grok 4.5 for coding and agentic work, co-trained with Cursor on real developer sessions. 500K-token context, $2/$6 per million input/output tokens, published model card with 15 benchmark scores
July 16–26, 2026 Moonshot AI's Kimi K3 moves from hosted preview to full open-weight release (2.8T parameters, 1M-token context)
July 28, 2026 Musk posts the Grok 4.6/4.7 roadmap replying to Rauch. Same day, 1,200+ employees at OpenAI, Anthropic, Google DeepMind, and Meta publish the "Pacing the Frontier" letter
~August 7, 2026 Grok 4.6 target: 1.5T parameters, SFT/RL upgrade (announced via tweet, unshipped)
Late Aug–early Sep 2026 Grok 4.7 target: 2.1T parameters. Musk says it beats 4.6 across the board except serving speed, with better token efficiency
Grok 4.3–4.7 at a glance
Model Date Parameters Focus Status
Grok 4.3 Beta Apr 17, 2026 Undisclosed Prior baseline Shipped
Grok 4.5 Jul 8, 2026 Undisclosed (single SKU, not MoE) Coding/agentic, Cursor co-training Shipped, benchmarked
Grok 4.6 ~Aug 7, 2026 1.5T SFT/RL upgrade Announced via tweet, unshipped
Grok 4.7 ~late Aug–early Sep 2026 2.1T Broad upgrade over 4.6, better token efficiency Announced via tweet, unshipped

All Grok 4.6/4.7 figures are unverified vendor claims from a single social media post — treat them as directional, not confirmed specs.

SECTION 03 Why xAI is emphasizing post-training, not just scale

Supervised fine-tuning (SFT) trains a model on curated example outputs to shape behavior. Reinforcement learning (RL) uses reward signals to teach which action sequences work — critical for multi-step agentic tasks. Musk's "significantly improved SFT & RL" framing signals xAI is doubling down on the playbook that made Grok 4.5 competitive on agentic benchmarks: roughly 15,954 output tokens per SWE-Bench Pro task versus Opus 4.8's 67,020 — a 4.2x efficiency gap credited to post-training on real Cursor developer sessions.

Grok 4.6's jump to 1.5T parameters is a real scale increase, but Musk's Grok 4.7 framing — bigger at 2.1T, "better in every way except slightly slower to serve" — suggests xAI is deliberately building two SKUs with different trade-offs rather than one model for everything.

Grok 4.6's target date lands almost exactly 10 days after Kimi K3's full open-weight release rattled the industry. Kimi K3 topped the Frontend Code Arena leaderboard at 1,679 points — the first open-weight model to beat every closed model on that board — and ranked third on Artificial Analysis's Intelligence Index. Musk commented "impressive" on Kimi K3 benchmark posts. That competitive pressure is the most plausible explanation for xAI compressing its release cadence to three frontier models in roughly two months.

How Grok stacks up against the field
Model Vendor Parameters Context Pricing (input/output per 1M tokens)
Grok 4.5 xAI Undisclosed 500K $2 / $6
Grok 4.6 (announced) xAI 1.5T Undisclosed Undisclosed
Kimi K3 Moonshot AI 2.8T (MoE, ~16/896 experts active) 1M $0.30 (cache hit) / $3 (cache miss) input, $15 output
Claude Fable 5.1 (rumored) Anthropic Undisclosed Undisclosed Rumored unchanged from Fable 5 ($10 in / $50 out)
GPT-5.6 Sol OpenAI Undisclosed Undisclosed Undisclosed

Grok 4.6 and Claude Fable 5.1 rows are pre-release claims, not verified benchmarks — useful for release timing and positioning, not head-to-head performance comparisons.

SECTION 04 What to flag before you trust this timeline

  • The only source is a tweet. No xAI blog post, model card, or product page confirms Grok 4.6's specs or date;
  • Benchmarks and pricing are blank. Unlike Grok 4.5's detailed launch materials, Grok 4.6 has no independent evaluation or official model card;
  • xAI is fighting safety controversies on a separate front. In July 2026, xAI sued a user for allegedly using Grok to generate CSAM — the company's first lawsuit of its kind. A January 2026 Common Sense Media report rated Grok among the worst AI chatbots for child-safety risks;
  • The timing collides with an industry split on AI pacing. The same day Musk announced Grok 4.6/4.7, over 1,200 employees published "Pacing the Frontier," asking the US government to help slow automated AI development. OpenAI and Anthropic endorsed it as companies. xAI is absent from that list.

If Musk's timeline holds, Grok 4.6 and Grok 4.7 will land in the same month as a rumored Claude Fable 5.1 (per leaks from 36kr and WinCentral) and just weeks after Kimi K3's open-weight shock. For teams evaluating models, that compresses the useful shelf life of any single flagship to a matter of weeks — which makes token efficiency and real-world task cost the more durable basis for a model choice.

SECTION 05 Citable specs and a 6-step pre-launch checklist

  • Grok 4.5 benchmarks (shipped): Terminal-Bench 2.1 83.3%, SWE-Bench Pro 64.7% (xAI official model card);
  • Grok 4.5 token efficiency: ~15,954 output tokens per SWE-Bench Pro task vs. Opus 4.8's ~67,020 (third-party summaries);
  • Grok 4.6 announced: 1.5T parameters, SFT/RL focus (Musk X post, unverified);
  • Grok 4.7 announced: 2.1T parameters, broadly better than 4.6 but slower to serve (Musk X post);
  • Kimi K3 reference: Frontend Code Arena 1,679 points, 2.8T MoE, 1M-token context (Moonshot official);
  • Grok 4.5 pricing reference: $2 input / $6 output per million tokens (xAI official).
  1. Lock your source of truth: Follow xAI's blog and @xai on X. Do not reroute production on second-hand reports alone;
  2. Build a baseline matrix: Before Grok 4.6's model card ships, benchmark Grok 4.5 and Kimi K3 on your own repos. Log token spend and pass rates;
  3. Reserve an A/B window: Avoid locking annual API contracts in the first week of August. Budget headroom for Grok 4.6 and rumored Fable 5.1;
  4. Split "stronger" vs "faster" SKUs: Per Musk, 4.6 is balanced and 4.7 is quality-heavy. Plan separate routes for latency-sensitive and quality-sensitive workloads;
  5. Re-check safety and compliance: Factor xAI's recent CSAM lawsuit and child-safety ratings into vendor policy review;
  6. Reconcile pricing on launch day: Use the xAI console and API docs as ground truth. Do not assume Grok 4.5 prices for financial models.

Official and third-party sources — verify against live pages before you rely on any date, spec, or price:

xAI: Introducing Grok 4.5

Elon Musk on X (July 28, 2026 reply to @rauchg)

Moonshot AI: Kimi K3 on GitHub

If your team is building Agents and iOS/macOS CI pipelines on Grok, Kimi, or other frontier models, virtualized Mac environments add overhead on Xcode builds and Metal workloads. For production setups that need zero-loss native compute, stable iOS CI/CD, and 24/7 Agent automation, MACNOX dedicated cloud Mac hardware is usually the better fit: genuine Apple hardware, full Root access, no hypervisor tax, flexible daily/weekly/monthly billing. See also our Grok 4.5 review and Kimi K3 open-weight coverage.

SECTION 06 FAQ

When exactly is Grok 4.6 coming out?

Musk said "around August 7, 2026" in an X post, but xAI has not officially confirmed a date. Treat it as a target that could shift.

What's the difference between Grok 4.6 and Grok 4.7?

Grok 4.6 is a 1.5T-parameter model focused on SFT/RL post-training improvements. Grok 4.7, expected a few weeks later, is a larger 2.1T model that Musk says outperforms 4.6 across the board except for serving speed, where it trades some latency for better token efficiency.

Will Grok 4.6 beat Kimi K3 or Claude Fable 5.1?

Too early to tell. Grok 4.6 has no published benchmarks yet, Kimi K3 already has verified third-party scores, and Claude Fable 5.1 hasn't been officially confirmed by Anthropic. A real comparison isn't possible until Grok 4.6 ships with a model card.

How much will Grok 4.6 cost?

Unknown. Grok 4.5 launched at $2 per million input tokens and $6 per million output tokens, which is a reasonable reference point, but xAI hasn't disclosed Grok 4.6 pricing.

Where will I be able to use Grok 4.6?

Based on Grok 4.5's rollout, expect Grok Build, the xAI API, and the xAI console to get access first, with third-party platform integrations following shortly after — but this isn't confirmed for 4.6 yet.