grok-code-fast-1 轉為 grok-build-0.1 由 beta 繼承落嚟嘅核心功能(V1.0 一樣有齊):
grok-build-0.1(源自 Grok 4.x 架構) grok-code-fast-1;現有模型嘅 benchmark 仲未獨立公佈) Grok Build V1.0 進入一個競爭激烈嘅市場,雖然 raw benchmark 分數落後,但喺架構上有獨特優勢:
| 功能 | Grok Build V1.0 | Claude Code | Codex CLI |
|---|---|---|---|
| 開發商 | xAI | Anthropic | OpenAI |
| 預設模型 | grok-build-0.1 (Grok 4.x) | Claude Opus 4.7 / Sonnet 4.6 | GPT-5.5 |
| Context window | 256K | 200K (Opus 4.7 有 1M) | 200K–1M |
| 多代理架構 | 最多 8 個 parallel sub-agents,每個有獨立 worktree | Task-tool sub-agents | Sub-task tool |
| Plan mode | ✅ | ✅ (Shift+Tab) | ✅ |
| SWE-Bench Verified | 70.8% | 87.6% | 88.7% |
| API input 價錢 | 大約 1 蚊/100 萬 tokens | 較高 (前線模型) | 視乎情況 |
| 入門價錢 | 包含喺 SuperGrok / X Premium+ | Claude Pro 月費 20 蚊 | ChatGPT 月費計劃 |
| 獨特賣點 | Parallel sub-agents + Arena Mode;coding 用嘅 API 最平 | 最審慎嘅 plan-then-execute;深度多檔案重構 | 內置 review agent;sandbox 先行;omnimodal 骨幹 |
總結: Grok Build 喺 raw benchmark 分數上落後(70.8% 對比競爭對手大約 88%),但喺 價錢(大約 1/2 蚊 per 100 萬 tokens,比起其他前線模型平好多)同 parallel sub-agent 架構(最多 8 個 concurrent agents)上有競爭力 。評測指出,Grok Build 最適合已經身處 xAI 生態圈嘅開發者(SuperGrok/X Premium+),而 Claude Code 喺複雜嘅生產環境重構表現更好,Codex CLI 就啱啲鍾意 sandbox 先行嘅工作流程
。
呢個雙重推出策略係咁樣行嘅: