Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
隆重推出 Gemini 3.6 Flash、3.5 Flash-Lite 和 3.5 Flash Cyber
Our newest Gemini models deliver the efficiency, latency, and reliability to build AI agents at scale. 我们最新的 Gemini 模型提供了构建大规模 AI 智能体所需的效率、延迟表现和可靠性。
Developers and customers building production AI agents need higher token efficiency, lower latency, and more reliable performance. Our Flash series of models is built to meet the sweet spot of efficiency and quality to enable scaling agentic workflows. 正在构建生产级 AI 智能体的开发者和客户需要更高的 Token 效率、更低的延迟以及更可靠的性能。我们的 Flash 系列模型旨在实现效率与质量的最佳平衡,从而助力智能体工作流的扩展。
Building on Gemini 3.5 Flash, we’re introducing new Gemini models: 在 Gemini 3.5 Flash 的基础上,我们推出了全新的 Gemini 模型:
-
3.6 Flash: Our workhorse model that delivers better coding, knowledge work, and multimodal performance. According to the Artificial Analysis Index, it reduces output token usage by 17% compared to 3.5 Flash, and in some benchmarks like DeepSWE by Datacurve, we observe up to 65%, all at a lower cost per output token. 3.6 Flash: 我们的主力模型,在编程、知识工作和多模态性能方面表现更佳。根据 Artificial Analysis Index 的数据,与 3.5 Flash 相比,它减少了 17% 的输出 Token 使用量;在 Datacurve 的 DeepSWE 等部分基准测试中,这一降幅甚至高达 65%,且每个输出 Token 的成本更低。
-
3.5 Flash-Lite: Our fastest, most cost-effective 3.5-class model, delivering 350 output tokens per second according to the Artificial Analysis Index, also significantly outperforming prior Flash-Lite generations in agentic workflows. 3.5 Flash-Lite: 我们速度最快、最具成本效益的 3.5 系列模型。根据 Artificial Analysis Index,其输出速度可达每秒 350 个 Token,在智能体工作流中的表现也显著优于前几代 Flash-Lite。
-
3.5 Flash Cyber in CodeMender: Successful cybersecurity applications require careful orchestration of a model alongside an agent infrastructure. We’re introducing a combination of a new, highly efficient, specialized cyber-focused model paired with our CodeMender code security agent that delivers competitive performance at the frontier. CodeMender 中的 3.5 Flash Cyber: 成功的网络安全应用需要将模型与智能体基础设施进行精细编排。我们推出了一款全新的、高效的、专注于网络安全领域的模型,并将其与我们的 CodeMender 代码安全智能体相结合,从而在行业前沿提供极具竞争力的性能。
Beyond today’s releases, Gemini 3.5 Pro is currently testing with partners and we plan to make it broadly available as soon as it’s ready. In parallel, our team is already focusing on building the next generation of models. We have started our most ambitious pre-training run yet, for Gemini 4, and are excited by the progress. 除了今天发布的内容外,Gemini 3.5 Pro 目前正在与合作伙伴进行测试,我们计划在准备就绪后尽快向公众开放。与此同时,我们的团队已经开始专注于构建下一代模型。我们已经启动了迄今为止最雄心勃勃的预训练项目——Gemini 4,并对目前的进展感到振奋。
3.6 Flash: More efficient and better quality than 3.5 Flash
3.6 Flash:比 3.5 Flash 更高效、质量更高
Gemini 3.6 Flash builds directly on developer and customer feedback from 3.5 Flash. 3.6 Flash not only delivers a step up in coding and knowledge work, but it does this while meaningfully improving token efficiency. For example, on the Artificial Analysis Index, we see 3.6 Flash consuming 17% fewer output tokens than 3.5 Flash. It also takes fewer reasoning steps and tool calls to accomplish multi-step workflows. Gemini 3.6 Flash 直接基于开发者和客户对 3.5 Flash 的反馈进行了改进。3.6 Flash 不仅在编程和知识工作方面实现了质的飞跃,同时还显著提升了 Token 效率。例如,在 Artificial Analysis Index 中,3.6 Flash 的输出 Token 消耗量比 3.5 Flash 少了 17%。此外,它在完成多步骤工作流时所需的推理步骤和工具调用也更少。
This enhanced efficiency is also combined with a lower price than 3.5 Flash. At $1.50/1M input tokens and $7.50/1M output tokens, 3.6 Flash reduces the overall cost per agentic task, making agents more cost-effective to build and run. 这种效率的提升还伴随着比 3.5 Flash 更低的价格。3.6 Flash 的输入 Token 价格为每百万 1.50 美元,输出 Token 价格为每百万 7.50 美元,这降低了每个智能体任务的总体成本,使得构建和运行智能体更具性价比。
Even while being more efficient, 3.6 Flash sees performance gains compared to 3.5 Flash across use cases: 尽管效率更高,但 3.6 Flash 在各类用例中的性能表现相比 3.5 Flash 仍有提升:
-
3.6 Flash delivers higher precision with fewer unwanted code edits and reduced execution loops, as seen in DeepSWE (49% vs. 37%), and shows significant improvement in ML Research, as seen in MLE Bench (63.9% vs. 49.7%). 3.6 Flash 提供了更高的精度,减少了不必要的代码编辑和执行循环(如 DeepSWE 测试中表现为 49% 对 37%),并在机器学习研究方面表现出显著进步(如 MLE Bench 测试中为 63.9% 对 49.7%)。
-
It has improved computer use capabilities as seen in OSWorld-Verified (83.0% vs. 78.4%). Computer use is now a built-in client side tool via the Gemini API and Gemini Enterprise. 其计算机使用能力得到了提升(如 OSWorld-Verified 测试中为 83.0% 对 78.4%)。现在,计算机使用功能已成为通过 Gemini API 和 Gemini Enterprise 提供的内置客户端工具。
-
It outperforms 3.5 Flash in knowledge work, as shown by benchmarks like GDPval-AA v2 (1421 vs. 1349). Customers like Hebbia and Harvey have found it particularly capable at multimodal tasks like document parsing, chart and data analysis, and report drafting. 在知识工作方面,它优于 3.5 Flash,如 GDPval-AA v2 基准测试所示(1421 对 1349)。Hebbia 和 Harvey 等客户发现它在文档解析、图表与数据分析以及报告撰写等多模态任务中表现尤为出色。
Built with safety
安全构建
3.6 Flash is shipping with enhanced Frontier Safety safeguards in the domains of Chemical, Biological, Radiological, and Nuclear (CBRN) and cyber offense misuses. These safeguards make the model substantially more resistant to jailbreaks. At the same time, the model has been trained to minimize refusals for beneficial uses. 3.6 Flash 在化学、生物、放射性和核(CBRN)领域以及网络攻击滥用方面配备了增强的前沿安全防护措施。这些防护措施使模型在抵御“越狱”攻击方面更加稳健。同时,该模型经过训练,旨在最大限度地减少对有益用途的拒绝。
3.5 Flash-Lite: Built to scale agentic workflows
3.5 Flash-Lite:为扩展智能体工作流而生
Beyond Flash, we’re also releasing Gemini 3.5 Flash-Lite, designed for both low-latency tasks and tasks where high throughput is critical for developers workflows, like agentic search and document processing. 除了 Flash 之外,我们还发布了 Gemini 3.5 Flash-Lite,它专为低延迟任务以及对开发者工作流至关重要的高吞吐量任务(如智能体搜索和文档处理)而设计。
3.5 Flash-Lite is the fastest model in the 3.5 series. As measured by Artificial Analysis, it runs at 350 output tokens/s. Priced at $0.3/1M input tokens and $2.5/1M output tokens and with significantly better quality than 3.1 Flash-Lite, 3.5 Flash-Lite offers a strong price-to-performance ratio for developers and customers running high throughput production traffic. 3.5 Flash-Lite 是 3.5 系列中速度最快的模型。据 Artificial Analysis 测算,其运行速度为每秒 350 个输出 Token。其输入 Token 价格为每百万 0.3 美元,输出 Token 价格为每百万 2.5 美元,且质量显著优于 3.1 Flash-Lite,为运行高吞吐量生产流量的开发者和客户提供了极佳的性价比。