Google announces Gemini 3.6 Flash and cybersecurity AI, teases 3.5 Pro and Gemini 4
Google announces Gemini 3.6 Flash and cybersecurity AI, teases 3.5 Pro and Gemini 4
Google 发布 Gemini 3.6 Flash 和网络安全 AI,并预告 3.5 Pro 与 Gemini 4
Google announced a significant evolution of its AI models at I/O in May with the release of Gemini 3.5 Flash, and it’s not slowing down. The company revealed three new AI models today, including its first version of Gemini geared toward cybersecurity. However, none of the new models is the delayed Gemini 3.5 Pro, which was supposed to launch in June.
Google 在 5 月的 I/O 大会上发布了 Gemini 3.5 Flash,标志着其 AI 模型的重要演进,而这一步伐并未放缓。该公司今日揭晓了三款全新的 AI 模型,其中包括首个专为网络安全设计的 Gemini 版本。然而,这些新模型中并不包含原定于 6 月发布但已延期的 Gemini 3.5 Pro。
Gemini 3.5 Flash, which was the star of the show at I/O, has already been deprecated. In its place, developers and users will find Gemini 3.6 Flash. Google makes the usual claims about this model—it’s marginally more capable and better at coding, and it has great multimodal features. Google says the changes to 3.6 Flash were made in response to user feedback on the 3.5 release. In general, Gemini 3.5 Flash didn’t appear to live up to Google’s promises around code generation. Perhaps that is simply a consequence of Google’s intense focus on efficiency as businesses have started to fret over the cost of AI tokens.
作为 I/O 大会明星产品的 Gemini 3.5 Flash 现已被弃用,取而代之的是 Gemini 3.6 Flash。Google 对该模型给出了常规评价:其能力略有提升,编程表现更好,并具备出色的多模态功能。Google 表示,3.6 Flash 的改进是针对用户对 3.5 版本反馈的回应。总体而言,Gemini 3.5 Flash 在代码生成方面似乎并未达到 Google 的预期。这或许是因为随着企业开始担忧 AI Token 的成本,Google 将重心过度转向了效率。
In the DeepSWE test for coding, 3.6 Flash jumps to 49 percent versus 37 percent for 3.5 Flash. The new model now supports computer use as a standard feature in the Gemini API, too. The OSWorld test for computer use shows a modest boost to 83 percent from 3.5’s 78.4 percent score. Efficiency was a big focus for Gemini 3.5 Flash, and Google says that effort has been amped up with 3.6. Even with small benchmark gains, Gemini 3.6 Flash uses about 17 percent fewer tokens. In agentic workflows (like the one below), Gemini 3.6 Flash should complete tasks more accurately, in fewer steps, and with fewer tokens. That could save developers (and Google) a lot of money.
在 DeepSWE 编程测试中,3.6 Flash 的得分从 3.5 Flash 的 37% 跃升至 49%。新模型现在还在 Gemini API 中将“计算机使用”(computer use)作为标准功能提供支持。在 OSWorld 的计算机使用测试中,其得分从 3.5 的 78.4% 小幅提升至 83%。效率一直是 Gemini 3.5 Flash 的重点,而 Google 表示 3.6 版本进一步加强了这一努力。即便基准测试提升幅度不大,Gemini 3.6 Flash 的 Token 使用量也减少了约 17%。在智能体工作流中,Gemini 3.6 Flash 能够以更少的步骤和更少的 Token 更准确地完成任务,这可能为开发者(以及 Google)节省大量成本。
The new model has a lower API cost, at $1.50/1M input tokens and $7.50/1M output tokens. It was $1.50 and $9, respectively, for 3.5 Flash.
新模型的 API 成本有所降低,输入 Token 为每百万 1.50 美元,输出 Token 为每百万 7.50 美元。相比之下,3.5 Flash 的价格分别为 1.50 美元和 9 美元。
Google is not done with the 3.5 branch yet, though. It has also released Gemini 3.5 Flash Lite and 3.5 Flash Cyber. The new Flash Lite is Google’s most efficient modern AI, hitting an impressive 350 tokens per second. The company claims this model is ideal for scaling agentic systems without breaking the bank. Based on benchmark numbers, the new Flash Lite is almost on par with frontier models from about a year ago, but it’s cheap. Pricing is set at $0.30/1M input tokens and $2.50/1M output tokens, though that is slightly higher than the previous 3.1 Flash Lite ($0.25 and $1.50).
不过,Google 在 3.5 系列上的动作并未结束。它还发布了 Gemini 3.5 Flash Lite 和 3.5 Flash Cyber。全新的 Flash Lite 是 Google 目前最高效的现代 AI,每秒处理速度高达惊人的 350 个 Token。该公司称,该模型非常适合在不增加高昂成本的情况下扩展智能体系统。根据基准测试数据,新的 Flash Lite 几乎与一年前的前沿模型相当,但价格低廉。其定价为输入 Token 每百万 0.30 美元,输出 Token 每百万 2.50 美元,尽管这比之前的 3.1 Flash Lite(0.25 美元和 1.50 美元)略高。
Gemini 3.6 Flash will begin rolling out in the API today, and it will take over from 3.5 Flash in the Gemini app. Likewise, Gemini 3.5 Flash Lite is available to developers and in the Gemini app. Google also notes that you’ll see a lot of 3.5 Flash Lite in Google search, where its higher speed probably makes it ideal for AI Overviews. So those may get a smidge better.
Gemini 3.6 Flash 将于今日开始在 API 中推送,并将取代 Gemini 应用中的 3.5 Flash。同样,Gemini 3.5 Flash Lite 也已面向开发者开放,并可在 Gemini 应用中使用。Google 还指出,你会在 Google 搜索中频繁看到 3.5 Flash Lite 的身影,其高速度使其成为 AI 概览(AI Overviews)的理想选择,因此搜索体验可能会有小幅提升。
Google’s AI roadmap
Google 的 AI 路线图
Google also has some news on upcoming AI models. First up will be a limited release of Gemini 3.5 Flash Cyber. This is Google’s first LLM tuned specifically for cybersecurity. The company says this model is almost as good at finding and fixing cybersecurity issues as the much larger and more expensive Claude Mythos. At the same time, it has the efficiency of a Flash model. Of course, Google acknowledges the “dual-use” nature of such models, which can just as easily be used to identify vulnerabilities for malicious purposes. The company borrows a page from Anthropic here, painting Gemini 3.5 Flash Cyber as too dangerous to release publicly. Instead, the model will launch soon as a limited pilot in Google DeepMind’s CodeMender agent, which is available exclusively to trusted partners and governments.
Google 还发布了一些关于未来 AI 模型的消息。首先是限量发布的 Gemini 3.5 Flash Cyber。这是 Google 首个专门针对网络安全进行调优的大语言模型。该公司表示,该模型在发现和修复网络安全问题方面的能力,几乎可以媲美规模更大、价格更昂贵的 Claude Mythos,同时又具备 Flash 模型的高效性。当然,Google 也承认此类模型具有“双重用途”,即同样可以被用于识别漏洞以进行恶意攻击。在此,Google 借鉴了 Anthropic 的做法,将 Gemini 3.5 Flash Cyber 描述为过于危险而不宜公开发布。相反,该模型将作为有限试点,很快在 Google DeepMind 的 CodeMender 智能体中推出,仅供受信任的合作伙伴和政府使用。
Then there’s the curious case of Gemini 3.5 Pro, which Google claimed was slated for a June release back at I/O. That never happened, and Google hasn’t had anything to say about it until now. There’s not much of an update, though. The company claims its new flagship model, which is supposed to rival GPT 5.6 and Claude Fable/Sonnet 5, is currently in testing with unnamed partners. The model will be released “as soon as it’s ready.” Earlier reports claimed that Google delayed 3.5 Pro because it couldn’t match competing models in coding. Gemini 3.5 Flash may not be top-of-the-line for long when it arrives.
接下来是令人好奇的 Gemini 3.5 Pro,Google 在 I/O 大会上曾称其定于 6 月发布。但这并未实现,且 Google 此前一直对此保持沉默。不过,目前并没有太多实质性更新。该公司声称,其新的旗舰模型旨在与 GPT 5.6 和 Claude Fable/Sonnet 5 竞争,目前正与未具名的合作伙伴进行测试。该模型将在“准备就绪后立即发布”。早先有报道称,Google 推迟 3.5 Pro 是因为其在编程能力上无法与竞争对手的模型相媲美。当它最终发布时,Gemini 3.5 Flash 可能很快就会失去其顶级地位。
Google also notes that it has started pre-training for Gemini 4, a process that is apparently more ambitious than its previous AI efforts. There’s no timeline for when we’ll see Gemini 4, and we don’t know if there will be more 3.x releases before that.
Google 还指出,它已经开始了 Gemini 4 的预训练工作,这一过程显然比其之前的 AI 项目更具雄心。目前还没有关于何时能见到 Gemini 4 的时间表,我们也不知道在此之前是否还会有更多的 3.x 版本发布。