跳到正文
原文
Noam Shazeer· @NoamShazeer · X·· 2026-03-04AI 评分61
AI 导读

Google 发布 Gemini 3.1 Flash-Lite,定位最快最高效、面向高负载工作负载的模型,官方称在推理、可靠性和可扩展性上超过 2.5 Flash 且成本更低。

正文

📢Introducing Gemini 3.1 Flash-Lite, our fastest and most efficient model, built for high-volume workloads. It outperforms 2.5 Flash in reasoning, reliability, and scalability at a lower cost.

This model also introduces thinking levels. You can adjust compute by complexity of the task, burning zero thinking overhead on high-volume tasks, while reasoning through the complex edge cases.

Maximum intelligence, minimal latency.

Read more: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-1-flash-lite

来源:Noam Shazeer · x.com