Google DeepMind·· 2025-12-13精选AI 评分78
Google DeepMind 发布更新版 Gemini 2.5 Flash Native Audio 原生音频模型
Improved Gemini audio models for powerful voice experiences
AI 导读
Google DeepMind 发布更新版 Gemini 2.5 Flash Native Audio 模型,强化了实时语音智能体的函数调用、指令遵循和多轮对话能力,并带来流式语音双向互译支持。
推荐理由
原生音频模型提升了复杂函数调用与上下文检索能力,使开发者能够在端到端语音交互中直接嵌入实时外部数据。
来源:Google DeepMind · deepmind.google