Gemini 3.8 Live and 3.8 Live Extended Thinking
Gemini 3.8 Live and 3.8 Live Extended Thinking
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking 隆重推出 Gemini 3.8 Live 与 3.8 Live Extended Thinking
Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice. Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking 是我们迄今为止最先进的实时对话模型。在智能和并行推理方面的重大升级,使它们在协作时更加直观,并能通过语音更高效地执行复杂任务。
We are launching Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to make voice interactions more natural, fluid, and intelligent. These models handle complex reasoning, real-time visual context, and background task execution without interrupting your conversation. You can start using these features today through the Gemini API, Google Workspace, and the Gemini app. 我们正式发布 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking,旨在让语音交互变得更加自然、流畅且智能。这些模型能够处理复杂的推理、实时视觉上下文以及后台任务执行,且不会中断您的对话。您即日起即可通过 Gemini API、Google Workspace 和 Gemini 应用开始使用这些功能。
Google just launched two new AI models that make talking to your devices feel way more natural. They can handle interruptions, switch between languages, and even explain their thought process while they work. Whether you’re solving a complex problem or just chatting, the AI now feels like it’s actually listening and thinking along with you. It’s a big step toward making AI feel like a real conversation partner. 谷歌刚刚发布了两款全新的 AI 模型,让与设备的对话感觉更加自然。它们能够处理打断、切换语言,甚至在工作时解释其思考过程。无论您是在解决复杂问题还是仅仅在闲聊,AI 现在给人的感觉就像是在真正倾听并与您同步思考。这是让 AI 成为真正对话伙伴的一大进步。
Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent. 今天,我们推出了两款新模型,它们在近实时推理方面带来了突破,能更有效地赋能语音代理,并使与 AI 的对话感觉更加直观和智能。
- Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live: 专为规模化和成本效益而构建,将对话智能与流畅的对话体验及视觉基础能力相结合。
- Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with increased intelligence and multi-step reasoning. Gemini 3.8 Live Extended Thinking: 专为高复杂度任务而构建,具备更强的智能和多步推理能力。
For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice. 对于开发者和企业而言,这些模型为构建可靠、可投入生产的语音代理提供了基础模块。它们还使您在 Gemini 应用、Google Workspace 和搜索中与 Gemini 的交流更加流畅和协作,帮助您仅通过语音即可处理复杂任务。
Experience more fluid, intelligent conversations 体验更流畅、更智能的对话
Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis’ Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models. Gemini 3.8 Live Extended Thinking 提供企业级的任务完成能力和智能,在 Artificial Analysis 的“语音对语音质量指数”中位列第一(82.6 分),并在代理任务完成度方面处于领先地位(在 τ-Voice 上达到 68.6%,在 Sierra 的 τ-Voice-banking 基准测试中达到 35.1%)。它还具备强大的推理能力,在 Big Bench Audio 上得分 97.7%,同时与其他前沿模型相比,保持了极具竞争力的价格优势。
Gemini 3.8 Live has shown a high preference among users, securing a second place in the Speech Agent Arena. In addition to this performance, it remains highly cost-effective — providing developers and enterprises with a capable and efficient model built for scale. On ServiceNow’s EVA-Bench, a benchmark for evaluating voice agents, our models push the Pareto Frontier for complex workflows by successfully balancing accuracy with conversational quality. Gemini 3.8 Live 在用户中表现出极高的偏好度,在“语音代理竞技场”(Speech Agent Arena)中排名第二。除了出色的性能外,它还保持了极高的成本效益,为开发者和企业提供了一个专为规模化构建的强大且高效的模型。在用于评估语音代理的基准测试 ServiceNow EVA-Bench 上,我们的模型通过成功平衡准确性与对话质量,推动了复杂工作流的帕累托前沿(Pareto Frontier)。
Gemini 3.8 Live processes visual inputs in near real-time, enriching conversations with context for more helpful responses. It automatically detects and transitions between 97 supported languages mid-conversation. It executes tools and API calls in the background while continuing the conversation, so the model can acknowledge requests and keep chatting while tasks finish in the background. Gemini 3.8 Live 能近乎实时地处理视觉输入,通过上下文丰富对话内容,从而提供更有帮助的回复。它能在对话过程中自动检测并切换 97 种支持的语言。它可以在后台执行工具和 API 调用,同时继续对话,因此模型可以在任务于后台完成的同时确认请求并保持交流。
For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. It delivers increased intelligence for complex workflows while maintaining an uninterrupted conversational flow — using early verbal cues like “Let me check that…” to acknowledge prompts naturally, and live progress narration to walk users through multi-step background tasks as they progress. 对于需要更深层推理的任务,3.8 Live Extended Thinking 可以边推理边说话。它在为复杂工作流提供更高智能的同时,保持了不间断的对话流——通过使用“让我查一下……”等早期口头提示来自然地确认指令,并利用实时进度叙述引导用户了解多步后台任务的进展。
Across Google Workspace and Search, our Live models deliver more intuitive, collaborative experiences — especially when tackling your most complex tasks. Try Gemini 3.8 Live Extended Thinking in Google Workspace with Docs Live, Gmail Live, and Keep Live. Get step-by-step, real-time troubleshooting help powered by Gemini 3.8 Live — right inside Search Live. 在 Google Workspace 和搜索中,我们的 Live 模型提供了更直观、更具协作性的体验,尤其是在处理最复杂的任务时。您可以在 Google Workspace 中通过 Docs Live、Gmail Live 和 Keep Live 体验 Gemini 3.8 Live Extended Thinking。在 Search Live 中,您还可以获得由 Gemini 3.8 Live 驱动的实时分步故障排除帮助。
Empowering the developer and enterprise voice ecosystem 赋能开发者与企业语音生态系统
By using the Gemini Live API, developer platforms such as Agora, Fishjam, LangChain, LiveKit, Pipecat, Vercel, and Vision Agents enable developers to build… 通过使用 Gemini Live API,Agora、Fishjam、LangChain、LiveKit、Pipecat、Vercel 和 Vision Agents 等开发者平台使开发者能够构建……