Gemini 3.8 Live and 3.8 Live Extended Thinking

原始链接: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/

Google 推出了两款全新模型:**Gemini 3.8 Live** 和 **Gemini 3.8 Live Extended Thinking**,旨在增强人工智能语音代理的直觉、智能和响应能力。 * **Gemini 3.8 Live:** 该模型针对规模化和成本效益进行了优化,侧重于流畅的对话和视觉基础,非常适合日常对话应用。它目前备受推崇,在“语音代理竞技场”(Speech Agent Arena)中排名第二。 * **Gemini 3.8 Live Extended Thinking:** 该模型专为高复杂度、多步骤推理而设计,可提供企业级性能。在保持具有竞争力的价格的同时,它目前在语音转语音质量和代理任务完成度方面处于行业基准领先地位。 这两款模型共同为开发人员和企业提供了构建可靠、可投入生产的语音代理的强大工具。通过将这些进展整合到 Google Workspace、搜索和 Gemini 应用中,Google 正在打造一种更加协作、语音驱动的体验,从而简化复杂任务的执行。

关于 Gemini 3.8 及其“实时”功能的最新 Hacker News 讨论显示,社区内部存在严重分歧。 许多用户称赞 Gemini 的语言多功能性,特别是它处理南非语、加泰罗尼亚语和绍纳语等小众语言的能力,以及在实时语音模式下自然、流畅的对话感。对于日常用户和深入调研需求而言,Gemini 因其速度和对话语调而备受青睐,许多人认为它更具“人情味”,不像 Claude 或 OpenAI 模型那样容易出现教条式的“AI 废话”或过度的委婉推托。 然而,相当一部分核心用户和开发者仍持批评态度。批评者认为,虽然 Gemini 在日常互动中表现出色,但在编程、逻辑推理和复杂任务执行方面,它仍落后于 Claude (Sonnet) 或 OpenAI 的最新模型。一些用户对谷歌在 Workspace 和个人账户之间不一致的模型版本发布,以及偶尔出现的“语境健忘”表示不满。 归根结底,这场辩论的核心在于价值判断:有些人认为谷歌因在原始编程基准测试中未能领先而“搞砸了”人工智能竞赛;而另一些人则认为,谷歌通过将高效、快速的 AI 集成到人们现有的工具中,成功占领了庞大的日常用户市场。
相关文章

原文

Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent.

  • Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
  • Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with increased intelligence and multi-step reasoning.

For developers and enterprises, these models deliver the building blocks for reliable, production-ready voice agents. They also make speaking with Gemini across the Gemini app, Google Workspace, and Search more fluid and collaborative — helping you tackle complex tasks using just your voice.

Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models.

Gemini 3.8 Live has shown a high preference among users, securing a second place in the Speech Agent Arena. In addition to this performance, it remains highly cost-effective — providing developers and enterprises with a capable and efficient model built for scale.

联系我们 contact @ memedata.com