Таймлайн релиза Grok 4.6 и roadmap параметров xAI

2026 Grok 4.6: дата релиза, Маск — 7 августа, 1,5T параметров и гайд по выбору конкурентов

28 июля 2026 Илон Маск в ответе на X к CEO Vercel Гильермо Рауху сообщил: xAI планирует Grok 4.6 с 1,5 триллионами параметров около 7 августа, затем Grok 4.7 с 2,1 триллионами через несколько недель. Месяц после Grok 4.5 — три frontier-модели в одном цикле. Официальные benchmark scores и pricing не опубликованы. Ниже — разделение verified facts и social-media preview, сравнение с конкурентами и техническая подготовка Agent runtime до релиза.

1. Три decision pain points: schedule, params, infra

  1. «Около 7 августа» как SLA deadline: единственный on-record source — один reply на X; xAI blog и product pages не синхронизированы. «Musk time» historically: slip от нескольких дней до двух недель. Для production planning — buffer, не hard commit.
  2. Model selection по parameter count без benchmark data: «1,5T» и «significantly improved SFT & RL» — vendor claims без independent eval и official model card. Kimi K3 имеет verified scores; Claude Fable 5.1 — unconfirmed rumor. Head-to-head «кто сильнее» — premature.
  3. Фокус на model, игнор agent execution environment: Grok 4.5 продемонстрировал token efficiency и value of Cursor co-training. Для 4.6 high-frequency agent loops требуют 7×24 uptime, stable network, auditable repo sync — hard production gate, orthogonal к выбору модели.

2. Timeline: Grok 4.5 → 4.6 → 4.7

  • 8 июля 2026 — xAI релизит Grok 4.5 для coding и agentic workloads, co-trained с Cursor на real developer sessions. Context window 500K tokens, pricing $2/$6 per 1M input/output tokens, published model card с 15 benchmark entries.
  • 16–26 июля 2026 — Moonshot AI Kimi K3: hosted preview → full open-weight release (2,8T params, 1M context), день раньше schedule.
  • 28 июля 2026 — Маск публикует roadmap 4.6/4.7 в reply к Rauch — primary source данного материала. В тот же день 1200+ employees OpenAI, Anthropic, Google DeepMind, Meta подписывают «Pacing the Frontier»; OpenAI и Anthropic endorse corporately. xAI не в списке signatories.
  • ~7 августа 2026 (target) — Grok 4.6, 1,5T parameters, positioning: SFT/RL post-training upgrade, не raw scale-up.
  • Конец августа – начало сентября 2026 (estimate) — Grok 4.7, 2,1T parameters; Маск: «better than 4.6 in every way, except slightly slower to serve, albeit with even better token efficiency.»

До релиза: ориентируйтесь на official xAI channels — social preview не substitute для production schedule.

3. Core data table (цитируемо)

Model Release / target Parameter scale Focus Status
Grok 4.3 Beta17 апр. 2026UndisclosedBaselineShipped
Grok 4.58 июл. 2026Undisclosed (single SKU)Coding/agentic, Cursor co-trainingShipped, benchmarked
Grok 4.6~7 авг. 20261,5TSFT/RL upgradeAnnounced via tweet
Grok 4.7~конец авг. – нач. сент.2,1TBroad upgrade, token efficiencyAnnounced via tweet

Parameter scale, pricing, benchmark scores — vendor/Musk statements; для 4.6/4.7 нет third-party independent eval — verify at official release.

4. Deep dive: upgrade в SFT/RL, не только scale

4.1 Что решают SFT и RL

Supervised fine-tuning (SFT) корректирует output behavior на curated high-quality samples; reinforcement learning (RL) оптимизирует action sequences через reward signals — critical для multi-step agentic pipelines. Musk явно позиционирует Grok 4.6 upgrade в «significantly improved SFT & RL», не в parameter count — continuation Grok 4.5 playbook: Terminal-Bench 2.1 83,3%, SWE-Bench Pro 64,7%, output tokens per comparable SWE task ~19K vs Claude Opus 4.8 ~67K (4,2× efficiency gap), attributed к post-training на real Cursor sessions.

4.2 Dual SKU strategy: scale vs latency trade-off

Grok 4.6 jump to 1,5T — material scale increase; Musk описывает 4.7 (2,1T) как «slightly slower to serve» при overall superiority — deliberate SKU segmentation: quality tier vs latency tier, аналог Sonnet/Opus или mini/full tiers у других vendors.

4.3 Competitive pressure от Kimi K3

Target date 4.6 — ~10 дней после Kimi K3 open-weight shock. K3 tops Frontend Code Arena 1679 points — first open-weight above all closed models; Artificial Analysis Intelligence Index rank #3 globally. Маск: «impressive» в benchmark threads. Compression xAI cadence (три frontier models ~два месяца) — plausible response к intensified competition с OpenAI, Anthropic и Chinese labs.

5. Horizontal compare: Grok vs Kimi K3, Claude Fable 5.1, GPT-5.6 Sol

Model Vendor Parameters Context Pricing (per 1M tokens) Source
Grok 4.5xAIUndisclosed500K$2 in / $6 outxAI official
Grok 4.6 (announced)xAI1,5TUndisclosedUndisclosedMusk X post (unverified)
Kimi K3Moonshot AI2,8T MoE (~16/896 active experts)1M$0,30/$3 in, $15 outMoonshot + Hugging Face
Claude Fable 5.1 (rumor)AnthropicUndisclosedUndisclosedRumor $10/$5036kr, WinCentral — unconfirmed
GPT-5.6 SolOpenAIUndisclosedUndisclosedUndisclosedOpenAI official

Rows 4.6 и Fable 5.1 — pre-release claims; не formal benchmark comparison; useful для release timing и positioning.

6. Controversies и technical caveats

  • Schedule unverified third-party: single X reply; no xAI blog mirror, no model card, no product page confirmation.
  • Benchmark и pricing void: unlike Grok 4.5 launch package, 4.6 has zero independent validation.
  • xAI content-safety controversies: незадолго до 4.6 preview xAI в июле 2026 sued user за alleged bypass of safety limits для CSAM generation — implicit admission что Grok under removed limits может produce such content. Common Sense Media (январь 2026) rated Grok among worst chatbots for child-safety. Separate technical capability eval от compliance/safety investment under rapid iteration cadence.
  • Industry pacing split: 28 июля — 4.6/4.7 roadmap и «Pacing the Frontier» same day; xAI absent от signatory list — acceleration vs deceleration как ongoing industry fracture.

7. August 2026: release crush

Если Musk timeline holds, Grok 4.6 и 4.7 land в same month с rumored Claude Fable 5.1 (leaks: август, before anticipated OpenAI GPT-6) плюс Kimi K3 open-weight shock из июля. Август 2026 — potentially highest frontier model release density. Useful shelf life flagship compresses to weeks — token efficiency и real task cost более durable basis для model choice чем parameter count или single leaderboard rank. Для ops: no architecture cutover в release week без fallback и rollback plan.

8. Five-step HowTo: prep checklist перед Grok 4.6

  1. Baseline Grok 4.5 cost и token metrics: log output tokens и USD/EUR per task type (terminal fix, cross-repo refactor, test gen) в Cursor Agent / OpenClaw routing.
  2. Pre-wire multi-model fallback: Grok 4.5 default в openclaw.json или Cursor Rules; Claude Fable 5/Opus для high-precision code review; Kimi K3 open-weight endpoint reserved.
  3. Pin 7×24 agent host: migrate long agent loops на always-on Apple Silicon node — laptop sleep breaks multi-step inference chain, conflicts с RL agent scenarios 4.6.
  4. SFTP/rsync sync codebase и config: align remote node repo, .cursorrules, CI artifacts перед parallel A/B load test без pollution main branch.
  5. Launch-day routing decision table: preset triggers для 4.6 vs 4.7 vs Kimi K3 по speed-sensitive / quality-sensitive / cost-sensitive columns — no blind full cutover в release week; adjust после official model card и third-party benchmarks.

9. Agent host decision matrix

Host option Best fit Main limitation Release-week recommendation
Personal laptop + Cursor Light completions, short chat Sleep interrupt, network jitter, no 7×24 soak ⚠️ Not production agent substrate
Generic cloud Linux VM API-only scripts, no GUI No native Cursor/macOS toolchain, no Apple Silicon stack ⚠️ API path OK; weak для Cursor co-training pipeline
SFTPMAC remote Apple Silicon Mac Cursor CLI/SDK, OpenClaw gateway, team SFTP/rsync sync Plan tier и bandwidth ✅ Migrate agent loops pre-release; parallel 4.6/4.7 soak tests

10. FAQ

Когда точно Grok 4.6?
Маск: «около 7 августа» на X; xAI не confirmed officially — movable target.

Разница 4.6 vs 4.7?
4.6: 1,5T, SFT/RL focus. 4.7: 2,1T, weeks later, overall better except slightly slower serve, higher token efficiency — per Musk.

Обгонит Kimi K3 или Fable 5.1?
Too early — no 4.6 benchmarks; K3 verified; Fable unconfirmed.

Pricing?
Unknown; 4.5 $2/$6 reference only.

Где access?
Expect Grok Build, xAI API, console first — Cursor integrations likely follow 4.5 pattern, unconfirmed for 4.6.

Sources: xAI «Introducing Grok 4.5» и model card (media.x.ai), Musk 28 июл. 2026 X (@rauchg), Chain Tech Daily / Tron Weekly / American Bazaar / Qazinform, Moonshot Kimi K3 и moonshotai/Kimi-K3, Fable 5.1 leaks (Emergent.sh / WinCentral / 36kr), «Pacing the Frontier» (The Verge / TechTimes), xAI safety coverage (Ars Technica / TechCrunch / The Guardian). As of 30 июл. 2026. Verify Grok 4.6/4.7 via official xAI channels before relying on dates, specs, prices.

11. Summary: agent substrate first, then Grok 4.6 landing

Grok 4.6 around 7 августа — xAI fast follow после Kimi K3 open-weight shock: clear narrative 1,5T + SFT/RL refinement, но benchmarks, pricing, official release date still empty. Teams on Grok 4.5: run token-efficiency edge на stable infra, reserve routing для 4.6/4.7 dual SKU — не rebuild entire stack на tweet day.

Whether standardize на 4.6, 4.7 или Kimi K3 — production agent workloads require 7×24 uptime, auditable repo sync, native Cursor compatibility, independent of parameter counts.

Для parallel soak tests across multiple frontier models в August release crush: host Cursor CLI, OpenClaw gateways, team repos на always-on Apple Silicon remote Mac с SFTP/rsync sync. SFTPMAC remote Mac rental — macOS environments tuned для AI agents: native Cursor, low-latency API callbacks, uninterrupted 24/7 operation — better launch-week flexibility чем personal device как agent host.