跳到正文
原文
🚨 AI News | TestingCatalog· @testingcatalog · X·· 7 小时前AI 评分47
AI 导读

突发 🔥:Google 宣布 Gemini 4 Argon,一款新的前沿模型,面向"跨真实世界软件工程、法律和金融等企业知识工作、以及网络防御的复杂工作流"。 在 DeepSWE v1.1 上取得 77.9% 的分数,创下新 SOTA。在众多基准测试上表现优于 GPT-6 Astra、Opus 5.5 和 Fable 5.1。 即将推出,首先面向付费 API 客户和 Google AI Ultra 订阅用户。 很快!👀

正文

BREAKING 🔥: Google announced Gemini 4 Argon, a new frontier model for "complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cyber defense."

77.9% score on DeepSWE v1.1 is a new SOTA. It performs better then GPT-6 Astra, Opus 5.5 and Fable 5.1 across many benchmarks.

Rolling out soon starting with paid API customers and Google AI Ultra subscribers.

Soon! 👀

引用Sundar Pichai@sundarpichai
Lots of discussion out there about our next model(!), so I wanted to give an early look as soon as possible. Introducing Gemini 4 Argon! It shows frontier performance in complex workflows, cyber defense and software engineering. Teams are using it extensively at Google, from coding to quantum computing, great feedback. Here’s a look at the benchmarks:
在 X 查看被引用的帖子

来源:🚨 AI News | TestingCatalog · x.com