What's actually confirmed vs. what's just Musk's word
Only one thing is on the record here: a tweet. Everything else — the "1.5 trillion parameters," the "significantly improved SFT & RL," the follow-up Grok 4.7 at 2.1 trillion parameters — comes from that single X post, with no accompanying model card, benchmark suite, or pricing page the way xAI published for Grok 4.5.
July 8, 2026: xAI ships Grok 4.5, its first model built specifically for coding and agentic work, co-trained with Cursor on real developer session data. It launched with a 500K-token context window, pricing at $2/$6 per million input/output tokens, and a published model card with 15 tracked benchmark scores.
July 16–26, 2026: Moonshot AI's Kimi K3 goes from hosted preview to fully open-weight release (2.8T parameters, 1M-token context), immediately topping Hugging Face's trending chart and drawing praise from Musk himself.
July 28, 2026: Musk posts the Grok 4.6/4.7 roadmap in reply to Rauch. Same day, more than 1,200 employees across OpenAI, Anthropic, Google DeepMind, and Meta publish the "Pacing the Frontier" letter.
~August 7, 2026 (target): Grok 4.6, 1.5T parameters, positioned as an SFT/RL upgrade rather than a raw scale-up.
Late August–early September 2026 (estimated): Grok 4.7, 2.1T parameters, which Musk says will be "better than 4.6 in every way, except slightly slower to serve, albeit with even better token efficiency."
"Musk time" has a track record — xAI, Tesla, and SpaceX timelines from Musk have historically slipped by days to weeks. Treat "around August 7" as a target, not a guarantee.
The numbers so far: Grok 4.6 vs. Kimi K3 and Claude Fable 5.1
| Model | Date | Parameters | Focus | Status |
|---|---|---|---|---|
| Grok 4.3 Beta | Apr 17, 2026 | Undisclosed | Baseline | Shipped |
| Grok 4.5 | Jul 8, 2026 | Undisclosed (single SKU, not MoE) | Coding/agentic, co-trained with Cursor | Shipped, benchmarked |
| Grok 4.6 | ~Aug 7, 2026 | 1.5T | SFT/RL upgrade | Announced via tweet, unshipped |
| Grok 4.7 | ~late Aug–early Sep 2026 | 2.1T | Broad upgrade over 4.6, better token efficiency | Announced via tweet, unshipped |
All Grok 4.6/4.7 figures are unverified vendor claims from a single social media post — treat them as directional, not confirmed specs.
| Model | Vendor | Parameters | Context | Pricing (input/output per 1M tokens) |
|---|---|---|---|---|
| Grok 4.5 | xAI | Undisclosed | 500K | $2 / $6 |
| Grok 4.6 (announced) | xAI | 1.5T | Undisclosed | Undisclosed |
| Kimi K3 | Moonshot AI | 2.8T (MoE, ~16/896 experts active) | 1M | $0.30 (cache hit) / $3 (cache miss) input, $15 output |
| Claude Fable 5.1 (rumored) | Anthropic | Undisclosed | Undisclosed | Rumored unchanged from Fable 5 ($10 in / $50 out) |
| GPT-5.6 Sol | OpenAI | Undisclosed | Undisclosed | Undisclosed |
Grok 4.6's target date lands almost exactly 10 days after Kimi K3's full open-weight release rattled the industry — Moonshot's model topped the Frontend Code Arena leaderboard at 1,679 points, becoming the first open-weight model to beat every closed model on that board.
Why xAI is emphasizing post-training, not just scale
Supervised fine-tuning (SFT) trains a model on curated example outputs to shape its behavior; reinforcement learning (RL) uses reward signals to teach a model which action sequences actually work, which matters most for multi-step agentic tasks.
Musk's choice of words — "significantly improved SFT & RL" — signals xAI is doubling down on the same playbook that made Grok 4.5 competitive on agentic benchmarks despite trailing rivals on raw intelligence scores: Grok 4.5 used roughly 15,954 output tokens per SWE-Bench Pro task versus Opus 4.8's 67,020, a 4.2x efficiency gap, largely credited to post-training on real Cursor developer sessions rather than sheer model size.
Grok 4.6's jump to 1.5T parameters is a real scale increase over Grok 4.5, but Musk's own framing of Grok 4.7 — bigger at 2.1T, "better in every way except slightly slower to serve" — suggests xAI is deliberately building two SKUs with different trade-offs rather than one model for everything.
Note: Chinese financial outlets reported that Kimi K3's release wiped an estimated $314 billion off combined OpenAI/Anthropic valuation expectations and $111 billion off Nvidia's market cap in a single day — striking figures, but they're analyst estimates relayed through Chinese media, not independently confirmed.
Six-step pre-launch evaluation Runbook for Grok 4.6
Lock your information sources: Tag Musk's July 28 X post as "informal preview." Subscribe to xAI's official blog and Grok Build announcements. Do not make architecture decisions on rumored specs before a formal release.
Build a baseline comparison set: Record SWE-Bench, Terminal-Bench, and your own agent task token consumption and latency against your current Grok 4.5, Kimi K3 API, or Claude Fable 5 deployment.
Pre-stage A/B routing: Add a Grok 4.6 model ID placeholder in your API gateway or LiteLLM router so you can gray-release within 24 hours of launch without changing application code.
Measure token efficiency, not just leaderboard rank: Grok 4.5's core advantage was output token efficiency (~19K tokens per SWE-Bench Pro task vs. Opus 4.8's 67K). Re-test actual bill cost on the same tasks after 4.6 ships.
Plan for the dual-SKU strategy: If 4.6 targets "faster" and 4.7 targets "stronger but slower," separate latency-sensitive and quality-sensitive workloads in your routing policy now.
Update vendor risk assessment: In July 2026, xAI sued a user for allegedly using Grok to generate CSAM — the company's first lawsuit of its kind. A January 2026 Common Sense Media report had already rated Grok among the worst AI chatbots for child-safety risks.
What to flag before you trust this timeline
The only source is a tweet. There's no xAI blog post, model card, or product page confirming Grok 4.6's specs or date — this is one executive's public statement, and xAI has no obligation to hit it.
Benchmarks and pricing are blank. Unlike Grok 4.5's detailed model card at launch, Grok 4.6 has no independent evaluation or official model card to substantiate its claimed capabilities.
The timing collides with an industry split on AI pacing. The same day Musk announced Grok 4.6/4.7, over 1,200 employees at OpenAI, Anthropic, Google DeepMind, and Meta published "Pacing the Frontier." xAI is notably absent from that list.
Warning: If Musk's timeline holds, Grok 4.6 and Grok 4.7 will land in the same month as a rumored Claude Fable 5.1 and just weeks after Kimi K3's open-weight shock. August 2026 is shaping up to be one of the densest release months in frontier model history.
For teams evaluating models, that compresses the useful shelf life of any single flagship to a matter of weeks — which makes token efficiency and real-world task cost, not leaderboard rank alone, the more durable basis for a model choice. Running long-lived agent orchestration or iOS CI/CD pipelines on a local Mac means fighting sleep timers, update prompts, and memory ceilings. For production environments that need stable Apple Silicon tooling alongside API-driven inference, MESHLAUNCH Mac Mini cloud rental is usually the better fit: dedicated hardware, 24/7 uptime, and flexible daily/weekly/monthly billing so you keep builds on a reliable cloud Mac while routing inference through whichever API wins this month's benchmark war.
Musk said "around August 7, 2026" in an X post, but xAI has not officially confirmed a date. Treat it as a target that could shift. See our pricing page for cloud Mac development options.
Grok 4.6 is a 1.5T-parameter model focused on SFT/RL post-training improvements. Grok 4.7, expected a few weeks later, is a larger 2.1T model that Musk says outperforms 4.6 across the board except for serving speed, where it trades some latency for better token efficiency.
Too early to tell. Grok 4.6 has no published benchmarks yet, Kimi K3 already has verified third-party scores, and Claude Fable 5.1 hasn't even been officially confirmed by Anthropic. Check our help center for cloud Mac setup guidance.
Unknown. Grok 4.5 launched at $2 per million input tokens and $6 per million output tokens, which is a reasonable reference point, but xAI hasn't disclosed Grok 4.6 pricing.
Based on Grok 4.5's rollout, expect Grok Build, the xAI API, and the xAI console to get access first, with third-party platform integrations (like Grok 4.5's day-one availability in Cursor) following shortly after — but this isn't confirmed for 4.6 yet.