Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

隆重推出 Gemini 3.8 Live 与 3.8 Live Extended Thinking

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice. Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking 是我们迄今为止最先进的实时对话模型。在智能水平和并行推理方面的重大升级,使得与它们协作变得更加直观,并能通过语音更高效地执行复杂任务。

We are launching Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to make voice interactions more natural, fluid, and intelligent. These models handle complex reasoning, real-time visual context, and background task execution without interrupting your conversation. You can start using these features today through the Gemini API, Google Workspace, and the Gemini app. 我们正式推出 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,旨在让语音交互变得更加自然、流畅且智能。这些模型能够处理复杂的推理、实时视觉语境以及后台任务执行,且不会中断您的对话。您即日起即可通过 Gemini API、Google Workspace 和 Gemini 应用开始使用这些功能。

Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent. 今天,我们推出了两款新模型,它们在近实时推理方面带来了突破,能更有效地赋能语音智能体,并使与人工智能的对话感觉更加直观和智能。

  • Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
  • Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with increased intelligence and multi-step reasoning.
  • Gemini 3.8 Live: 专为规模化和成本效益而设计,将对话智能与流畅的交流及视觉基础能力相结合。
  • Gemini 3.8 Live Extended Thinking: 专为高复杂度任务而打造,具备更强的智能水平和多步推理能力。

For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice. 对于开发者和企业而言,这些模型为构建可靠、可投入生产的语音智能体提供了基础模块。它们还使得在 Gemini 应用、Google Workspace 和搜索中与 Gemini 的交流更加流畅和协作化,帮助您仅通过语音即可处理复杂任务。

Experience more fluid, intelligent conversations

体验更流畅、更智能的对话

Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis’ Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models. Gemini 3.8 Live Extended Thinking 提供企业级的任务完成能力和智能水平,在 Artificial Analysis 的“语音对语音质量指数”(Speech to Speech Quality Index)中以 82.6 分位列总榜第一,并在智能体任务完成度方面表现领先,在 τ-Voice 测试中达到 68.6%,在 Sierra 的 τ-Voice-banking 基准测试中达到 35.1%。它还具备强大的推理能力,在 Big Bench Audio 测试中得分 97.7%,同时与其他前沿模型相比,保持了极具竞争力的价格优势。

Gemini 3.8 Live has shown a high preference among users, securing a second place in the Speech Agent Arena. In addition to this performance, it remains highly cost-effective — providing developers and enterprises with a capable and efficient model built for scale. On ServiceNow’s EVA-Bench, a benchmark for evaluating voice agents, our models push the Pareto Frontier for complex workflows by successfully balancing accuracy with conversational quality. Gemini 3.8 Live 在用户中表现出极高的偏好度,在“语音智能体竞技场”(Speech Agent Arena)中位居第二。除了出色的性能外,它还保持了极高的成本效益,为开发者和企业提供了一个能够规模化应用的高效模型。在用于评估语音智能体的 ServiceNow EVA-Bench 基准测试中,我们的模型通过成功平衡准确性与对话质量,推动了复杂工作流的帕累托最优边界。

Gemini 3.8 Live processes visual inputs in near real-time, enriching conversations with context for more helpful responses. It automatically detects and transitions between 97 supported languages mid-conversation. It executes tools and API calls in the background while continuing the conversation, so the model can acknowledge requests and keep chatting while tasks finish in the background. Gemini 3.8 Live 能够近乎实时地处理视觉输入,通过语境丰富对话内容,从而提供更有帮助的回复。它可以在对话过程中自动检测并切换 97 种支持的语言。它能在后台执行工具调用和 API 请求的同时继续对话,因此模型可以在任务于后台完成时,一边确认请求一边保持交流。

For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. It delivers increased intelligence for complex workflows while maintaining an uninterrupted conversational flow — using early verbal cues like “Let me check that…” to acknowledge prompts naturally, and live progress narration to walk users through multi-step background tasks as they progress. 对于需要更深层推理的任务,3.8 Live Extended Thinking 可以边推理边说话。它在为复杂工作流提供更高智能的同时,保持了不间断的对话流——通过使用“让我查一下……”等早期口头提示来自然地确认指令,并利用实时进度叙述引导用户了解多步后台任务的进展。

Across Google Workspace and Search, our Live models deliver more intuitive, collaborative experiences — especially when tackling your most complex tasks. Try Gemini 3.8 Live Extended Thinking in Google Workspace with Docs Live, Gmail Live, and Keep Live. Get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live — right inside Search Live. 在 Google Workspace 和搜索中,我们的 Live 模型提供了更直观、更具协作性的体验,尤其是在处理最复杂的任务时。您可以在 Google Workspace 中通过 Docs Live、Gmail Live 和 Keep Live 体验 Gemini 3.8 Live Extended Thinking。您还可以直接在 Search Live 中获得由 Gemini 3.8 Live 驱动的实时分步故障排除帮助。

Empowering the developer and enterprise voice ecosystem

赋能开发者与企业语音生态系统

By using the Gemini Live API, developer platforms such as Agora, Fishjam, LangChain, LiveKit, Pipecat, Vercel, and Vision Agents enable developers to build… 通过使用 Gemini Live API,Agora、Fishjam、LangChain、LiveKit、Pipecat、Vercel 和 Vision Agents 等开发者平台,使开发者能够构建……