Noam Shazeer· @NoamShazeer · X·· 2026-03-04AI 评分61
AI 导读
Google 发布 Gemini 3.1 Flash-Lite,定位最快最高效、面向高负载工作负载的模型,官方称在推理、可靠性和可扩展性上超过 2.5 Flash 且成本更低。
正文
📢Introducing Gemini 3.1 Flash-Lite, our fastest and most efficient model, built for high-volume workloads. It outperforms 2.5 Flash in reasoning, reliability, and scalability at a lower cost.
This model also introduces thinking levels. You can adjust compute by complexity of the task, burning zero thinking overhead on high-volume tasks, while reasoning through the complex edge cases.
Maximum intelligence, minimal latency.
Read more: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-1-flash-lite
来源:Noam Shazeer · x.com