Gemini 2.5 Flash-Lite 稳定版发布,主打低成本高性能
原标题:Gemini 2.5 Flash-Lite is now ready for scaled production use
AI 摘要
Google DeepMind 发布了 Gemini 2.5 Flash-Lite 的稳定版,这是 Gemini 2.5 系列中速度最快、成本最低的模型,输入每百万 token 0.10 美元,输出每百万 token 0.40 美元。该模型支持原生推理能力、100 万 token 上下文窗口,并已应用于 Satlyt、HeyGen 等客户场景。稳定版现已通过 Google AI Studio 和 Vertex AI 提供,预览版别名将于 8 月 25 日移除。
正文节选
Today, we’re releasing the stable version of Gemini 2.5 Flash-Lite, our fastest and lowest cost ($0.10 input per 1M, $0.40 output per 1M) model in the Gemini 2.5 model family. We built 2.5 Flash-Lite to push the frontier of intelligence per dollar, with native reasoning capabilities that can be optionally toggled on for more demanding use cases. Building on the momentum of 2.5 Pro and 2.5 Flash, this model rounds out our set of 2.5 models that are ready for scaled production use. Our most cost-e