跳到正文
Google DeepMind·· 2025-12-13精选AI 评分78

Google DeepMind 发布更新版 Gemini 2.5 Flash Native Audio 原生音频模型

Improved Gemini audio models for powerful voice experiences

AI 导读

Google DeepMind 发布更新版 Gemini 2.5 Flash Native Audio 模型,强化了实时语音智能体的函数调用、指令遵循和多轮对话能力,并带来流式语音双向互译支持。

推荐理由

原生音频模型提升了复杂函数调用与上下文检索能力,使开发者能够在端到端语音交互中直接嵌入实时外部数据。

来源:Google DeepMind · deepmind.google