🚨 AI News | TestingCatalog· @testingcatalog · X·· 7 小时前AI 评分47
AI 导读
突发 🔥:Google 宣布 Gemini 4 Argon,一款新的前沿模型,面向"跨真实世界软件工程、法律和金融等企业知识工作、以及网络防御的复杂工作流"。 在 DeepSWE v1.1 上取得 77.9% 的分数,创下新 SOTA。在众多基准测试上表现优于 GPT-6 Astra、Opus 5.5 和 Fable 5.1。 即将推出,首先面向付费 API 客户和 Google AI Ultra 订阅用户。 很快!👀
正文
BREAKING 🔥: Google announced Gemini 4 Argon, a new frontier model for "complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cyber defense."
77.9% score on DeepSWE v1.1 is a new SOTA. It performs better then GPT-6 Astra, Opus 5.5 and Fable 5.1 across many benchmarks.
Rolling out soon starting with paid API customers and Google AI Ultra subscribers.
Soon! 👀
Lots of discussion out there about our next model(!), so I wanted to give an early look as soon as possible. Introducing Gemini 4 Argon! It shows frontier performance in complex workflows, cyber defense and software engineering. Teams are using it extensively at Google, from coding to quantum computing, great feedback. Here’s a look at the benchmarks:在 X 查看被引用的帖子
来源:🚨 AI News | TestingCatalog · x.com