Google 于 2026 年 8 月 13 日推出 Gemini 3.7 Flash,主打编程与 AI Agent,性能大幅提升;DeepSWE v1.1 从 49% 跃升至 65.3%,FrontierCode 1.1 达 43.6%,AutomationBench 成绩近翻倍至 30.4%。 价格极具颠覆性:年底前输入/输出价格分别为每百万 token 0.75/3.75 美元,约为同级别竞品 Claude Sonnet 5 和 GPT 5.6 Terra 的三分之一;模型已通过 Gemini API、AI Studio 及消费级“Gemini Spark”平台上线。
研究答案

Create a landscape editorial hero image for this Studio Global article: What did Google announce with Gemini 3.7 Flash on August 14, 2026, and what is the status of the delayed Gemini 3.5 Pro flagship model, incl. Article summary: Here is a full rundown based on the available evidence from August 2026.. Topic tags: general, general web, news, user generated. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts with fake numbers, clickbait thumbnails, icons, and tiny thumbnail layouts. Make it useful as an illustrative visual, not as factual evidence.
2026 年 8 月 13 日,Google 正式推出 Gemini 3.7 Flash,这是其面向编程和 AI Agent 工作流的最新“工作马”模型,距前代 3.6 Flash 发布仅三周。然而,此次发布并未伴随外界期待已久的旗舰模型 Gemini 3.5 Pro 的任何更新,加之 DeepMind 高层近日发生剧烈动荡,Google 的 AI 路线图正笼罩在迷雾之中 。
Google 将 Gemini 3.7 Flash 定位为“迄今为止最智能的、面向编程和 Agent 的工作马模型” 。根据官方公告,该模型在软件工程、Web 开发和 Agent 工作流方面均有显著提升
。Google 声称,3.7 Flash 在九项基准测试中均优于 Anthropic 和 OpenAI 的同级别模型
。
| 基准测试 | 3.6 Flash 成绩 | 3.7 Flash 成绩 | 提升幅度 |
|---|---|---|---|
| DeepSWE v1.1(软件工程基准) | 49.0% | 65.3% | +16.3 个百分点 |
| FrontierCode 1.1 Main(前沿编程) | 34.4% | 43.6% | +9.2 个百分点 |
| AutomationBench(自动化基准) | 17.0% | 30.4% | +13.4 个百分点(接近翻倍) |
| GDP.pdf(文档处理) | 22.0% | 34.0% | +12.0 个百分点 |
| WebDev Arena Elo(Web 开发排名) | ~1538 | 1588 | +50 分 |
Gemini 3.7 Flash 的定价策略被视为其最大杀招 :
作为对比,其主要竞品为 Anthropic 的 Claude Sonnet 5(限时价:输入 $2/输出 $10 每百万 token)和 OpenAI 的 GPT-5.6 Terra(定价为 $2.50/$15)。在限时推广期间,Gemini 3.7 Flash 在输入上的成本仅为 Sonnet 5 的约三分之一和 Terra 的三分之一,同时提供了极具竞争力的编程成绩(DeepSWE 65.3% vs Sonnet 5 的 63.2%)。不过,在特定的自动化编程基准 Terminal-Bench 2.1 上,Terra(87.1%)表现优于 Sonnet 5(78.4%),而 Google 未公布 3.7 Flash 在此项的数据
。
当前状态: Gemini 3.5 Pro 仍然 延期,没有公布任何发布日期。自 2026 年 5 月在 Google I/O 上承诺给开发者后,该旗舰模型一再被推迟 。在 3.7 Flash 的发布会上,Google 完全未提及 3.5 Pro 的进展
。这意味着面向开发者的“工作马”Flash 系列已经领先于旗舰 Pro 系列至少两个版本
。
定价(未经官方确认,仅为估算): Google 从未公布 3.5 Pro 的官方定价 。根据行业预测:
被取消的猜想: 伴随着漫长的延期、高层人事动荡以及内部已开始早期 Gemini 4 训练 的消息,多家分析指出 Google 可能跳过 3.5 Pro 直接发布 4.0 旗舰 。彭博社和路透社均评论称,对于一款延期如此之久的旗舰模型,Google 的沉默是前所未有的
。
在 3.7 Flash 发布的同一周,Google 经历了 AI 部门自 2023 年 DeepMind 与 Google Brain 合并以来最大的内部变革。
1. 联合创始人施压:“All in Gemini”
Google 联合创始人 Sergey Brin 近期频繁接触核心 AI 员工,要求他们“全力投入 Gemini”,以缩小与竞争对手的差距。据知情人士透露,Brin 甚至直接督促研究人员加快进度 。
2. Demis Hassabis 退居二线
2026 年 8 月 5 日,Google 宣布 Demis Hassabis 将卸任 DeepMind 的 CEO 职务,转而担任 主席兼首席科学家,专注于通用人工智能(AGI)战略和药物研发 。
3. Koray Kavukcuoglu 接掌大权
DeepMind 首席技术官 Koray Kavukcuoglu 将成为实际上的运营领袖,直接向集团 CEO Sundar Pichai 汇报 。此举也被视为将 AI 决策权从伦敦进一步向山景城总部的集中
。
4. 核心人才流失
《财富》杂志报道称,此次人才出走潮的诱因是“模型开发停滞、未能达到既定目标以及员工过劳” 。
总结: Google 发布了一款价格极具杀伤力的强力新模型,在编程和 Agent 领域向 Anthropic 和 OpenAI 发起了有力挑战。但这一切发生在公司自家人工智能领导层出现严重动荡的背景下。旗舰模型的“爽约”以及核心人才的不断流失,让 Google 的 AI 蓝图在近期充满了不确定性。
Studio Global AI
此页面包含一个有来源支持的答案,您可以在 Studio Global 内继续。
Google 于 2026 年 8 月 13 日推出 Gemini 3.7 Flash,主打编程与 AI Agent,性能大幅提升;DeepSWE v1.1 从 49% 跃升至 65.3%,FrontierCode 1.1 达 43.6%,AutomationBench 成绩近翻倍至 30.4%。
Google 于 2026 年 8 月 13 日推出 Gemini 3.7 Flash,主打编程与 AI Agent,性能大幅提升;DeepSWE v1.1 从 49% 跃升至 65.3%,FrontierCode 1.1 达 43.6%,AutomationBench 成绩近翻倍至 30.4%。 价格极具颠覆性:年底前输入/输出价格分别为每百万 token 0.75/3.75 美元,约为同级别竞品 Claude Sonnet 5 和 GPT 5.6 Terra 的三分之一;模型已通过 Gemini API、AI Studio 及消费级“Gemini Spark”平台上线。
旗舰模型 Gemini 3.5 Pro 继续延期且无发布时间表,内部已有猜测称 Google 可能跳过 3.5 Pro 直接开发 Gemini 4。