Gemini API Changelog·· 2025-05-20精选AI 评分66
Google 发布 Gemini 2.5 Flash 预览版及多款语音、音乐模型
May 20, 2025
AI 导读
Google 在 Gemini API 更新中发布 gemini-2.5-flash-preview-05-20,一款主打性价比与自适应思考的预览模型,同时发布 gemini-2.5-pro-preview-tts 和 gemini-2.5-flash-preview-tts,支持单人或双人语音生成。
推荐理由
Gemini API 一次更新集中放出多款预览模型与工具能力,可据此判断多模态与实时语音的可用边界。
正文 · 原文
API updates:
- Launched support for custom video preprocessing using clipping intervals and configurable frame rate sampling.
- Launched multi-tool use, which supports configuring
code execution and
Grounding with Google Search on the same
generateContentrequest. - Launched support for asynchronous function calls in the Live API.
- Launched an experimental URL context tool for providing URLs as additional context to prompts.
Model updates:
- Released
gemini-2.5-flash-preview-05-20, a Gemini preview model optimized for price-performance and adaptive thinking. To learn more, see Gemini 2.5 Flash Preview and Thinking. - Released the
gemini-2.5-pro-preview-ttsandgemini-2.5-flash-preview-ttsmodels, which are capable of generating speech with one or two speakers. - Released the
lyria-realtime-expmodel, which generates music in real time. - Released
gemini-2.5-flash-preview-native-audio-dialogandgemini-2.5-flash-exp-native-audio-thinking-dialog, new Gemini models for the Live API with native audio output capabilities. To learn more, see the Live API guide and Gemini 2.5 Flash Native Audio. - Released
gemma-3n-e4b-itpreview, available on AI Studio and through the Gemini API, as part of the Gemma 3n launch.
来源:Gemini API Changelog · ai.google.dev