谷歌发布 Gemini 3.1 Flash-Lite:高性价比低延迟模型
原标题:Gemini 3.1 Flash-Lite: Built for intelligence at scale
AI 摘要
Google DeepMind 发布了 Gemini 3.1 Flash-Lite,这是 Gemini 3 系列中速度最快、成本效益最高的模型,专为大规模高吞吐量开发者工作负载设计。该模型定价为每百万输入 token 0.25 美元、每百万输出 token 1.5 美元,性能优于 2.5 Flash,输出速度提升 45%,首 token 时间快 2.5 倍。它已在 Google AI Studio 和 Vertex AI 上以预览版形式向开发者提供,并支持思考级别控制,适用于翻译、内容审核等成本敏感型任务以及复杂推理场景。
正文节选
Gemini 3.1 Flash-Lite: Built for intelligence at scale Today, we're introducing Gemini 3.1 Flash-Lite, our fastest and most cost-efficient Gemini 3 series model. Built for high-volume developer workloads at scale, 3.1 Flash-Lite delivers high quality for its price and model tier. Starting today, 3.1 Flash-Lite is rolling out in preview to developers via the Gemini API in Google AI Studio and for enterprises via Vertex AI. Cost-efficiency without compromise Priced at just $0.25/1M input tokens an