2026年9月17日更新
今天,我们推出两款新模型,将近实时推理能力推向新高度,从而更有效地赋能语音智能体,让与人工智能的对话体验更加直观、智能。
Gemini 3.8 Live:专为大规模部署和成本效率而设计,融合对话智能、流畅对话与视觉 grounding(视觉定位)能力。
Gemini 3.8 Live Extended Thinking:专为高复杂度任务打造,具备更强的智能水平和多步推理能力。
对于开发者和企业而言,这些模型提供了构建可靠、可投入生产环境的语音智能体的核心组件。同时,它们也让通过 Gemini App、Google Workspace 和 Search 与 Gemini 进行语音交互变得更加流畅和协作式——助你仅凭语音即可应对复杂任务。
体验更流畅、更智能的对话
Gemini 3.8 Live Extended Thinking 提供企业级任务完成能力和智能水平,在 Artificial Analysis 的“语音到语音质量指数”(Speech to Speech Quality Index)中位列第一(得分82.6),并在代理式任务完成方面表现领先:在 τ-Voice 基准测试中达到68.6%,在 Sierra 的 τ-Voice-banking 基准测试中达到35.1%。它还具备强大的推理能力,在 Big Bench Audio 测试中得分97.7%,同时与其他前沿模型相比保持了极具竞争力的价格定位。
Gemini 3.8 Live 获得了用户的高度青睐,在 Speech Agent Arena 中位列第二。除优异性能外,它依然保持极高的成本效益——为开发者和企业提供了一款强大且高效、专为大规模应用而设计的模型。
Updated September 17, 2026
Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent.
Gemini 3.8 Live : Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
Gemini 3.8 Live Extended Thinking : Built for high-complexity tasks, with increased intelligence and multi-step reasoning.
For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice.
Experience more fluid, intelligent conversations
Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ -Voice and 35.1% on Sierra’s τ -Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models.
Gemini 3.8 Live has shown a high preference among users, securing a second place in the Speech Agent Arena . In addition to this performance, it remains highly cost-effective — providing developers and enterprises with a capable and efficient model built for scale.
首次收录 · 2026-09-28 · 9.4 分