#Agent
#Agent
今日 191 条
Yuchen Jin@Yuchenj_UWAI 评分5454
TypeSafe AI@typesafeaiAI 评分4949引用Zhaorun Chen@zrrrr_cnJev is fast at helping you. Turns out, it can also be fast at helping an attacker!! 😱🚨 We red-teamed Jev 1.13 on our DTap (DecodingTrust-Agent Platform) and found a serious safety gap: 70.1% ASR under direct misuse 43.5% ASR under indirect prompt injection In our evaluations, we found that under indirect prompt injection, Jev can follow attacker-injected instructions without blinking an eye, e.g., exfiltrating user data, deleting files, or taking other harmful actions. But we found a much safer way to integrate Jev: use it as a self-gating layer for its own tool calls, significantly reducing ASR while preserving most of its utility. 👇 Read more below
Latent Space(RSS)AI 评分7575 AINews:Opus 5.5 擅长生成讲解视频
Latent Space 的 AINews 汇总 9/24-9/25 动态,指出本周发布的 Claude Opus 5.5 在讲解视频生成上表现突出,并以 88.4% 领跑 SimpleBench,在 Terminal-Bench-Science 上从低推理强度的 24% 升至 xhigh 的 62%。
Latent Space(RSS)AI 评分6262 Anthropic 的 Thariq Shihipar 谈 Claude Code 的下一阶段
Anthropic 的 Thariq Shihipar 在 Latent Space 播客中谈 Claude Code 的下一阶段,包括 Ask User Question、artifacts、Claude Tag、Projects 和可自定义 harness 的 Claude Mods。
Aravind Srinivas@AravSrinivasAI 评分4444Perplexity Agent API 新增 Profiles、Skills 和托管连接器。Profile 是可保存的智能体配置,统一管理模型、指令、工具、连接器和运行设置。
引用Perplexity Developers@perplexitydevsYou can now build custom reusable agents in the Perplexity Agent API with Profiles, Skills, and managed connectors. Configure an agent once in the API Portal and reuse it across applications and workflows.
elsewhere:文章(RSS)AI 评分5757 Manus 2.0 发布,并推出个人生活智能助理 Cue
9 月 28 日 Manus 面向海外用户发布 2.0 版本,并推出面向个人生活场景的智能助理 Cue。Cue 中每个 Agent 可拥有自己的邮箱、电话号码、钱包和电脑,代表用户与现实服务交互,多 Agent 可进入同一群聊处理扫码点餐、排队取号等事宜。Manus 正在组建团队开发面向国内市场的产品,与国产模型厂商及生态伙伴的合作稳步推进。
Arena.ai@arenaAI 评分4949OpenAI 的 GPT-6 Luna (Max) 进入 Agent Arena 帕累托前沿,净提升 +1.59%,中位成本仅 $0.05/任务。
引用Arena.ai@arenaGPT-6 Luna (Max) by @OpenAI is #23 in Agent Arena with +1.6% net improvement across 8K real-world agentic sessions from our global community of users. Although GPT-6 Luna (Max) did not land on the Agent Arena Pareto frontier, it remains a cost-efficient model. Its $0.05 median cost per task is 94% lower than GPT-6 Sol (Max) at $0.82 and 98% lower than GPT-6 Astra (Max) at $2.59. Its net-improvement score also comes within 0.09 percentage points of #22 GPT 5.5, while costing 91% less than its $0.56 median cost per task. This release is a six-place point-rank move over GPT-5.6 Luna (xHigh), at -0.9% and #29! By signal, GPT-6 Luna’s clearest gains over GPT-5.6 Luna are in: - Confirmed Success: #17 (+4.6%) vs. #33 (-4.6%) - Bash Recovery: #21 (+4.2%) vs. #27 (+2.2%) Congrats to the @OpenAI team on this release!
Arena.ai@arenaAI 评分5050
Peter Steinberger 🦞@steipeteAI 评分2626引用PWV@PWVenturesSave the date: AgentCribs SF, Tue Oct 6. For engineers shipping with agents. Afternoon workshop, then an evening fireside: @mojombo hosts @steipete, creator of @openclaw. The @aiworthusing x OpenClaw hackathon winner demos live. Space is limited. Registration opens tomorrow.
AWS Machine Learning Blog精选AI 评分6262 xAI Grok 4.7 上线 Amazon Bedrock
xAI 的 Grok 4.7 已在 Amazon Bedrock 上线,提供 500K token 上下文窗口和 low、medium、high、xhigh 四档可配置推理强度,通过 bedrock-runtime 端点的跨区域推理配置文件提供服务,支持 Responses、Chat Completions 和 Converse API。
推荐理由:梳理了 Grok 4.7 在 Bedrock 上的接入方式、推理档位与成本取舍,便于评估长任务智能体的落地配置。
ClaudeDevs@ClaudeDevs精选AI 评分7272
推荐理由:指南围绕 Sonnet 5.5 与 Opus 5.5 的选型、迁移调参和 Claude Code 使用给出实操建议,便于开发者上手。
Diogo Almeida@CompleteSkepticAI 评分5050引用Zhaorun Chen@zrrrr_cnJev is fast at helping you. Turns out, it can also be fast at helping an attacker!! 😱🚨 We red-teamed Jev 1.13 on our DTap (DecodingTrust-Agent Platform) and found a serious safety gap: 70.1% ASR under direct misuse 43.5% ASR under indirect prompt injection In our evaluations, we found that under indirect prompt injection, Jev can follow attacker-injected instructions without blinking an eye, e.g., exfiltrating user data, deleting files, or taking other harmful actions. But we found a much safer way to integrate Jev: use it as a self-gating layer for its own tool calls, significantly reducing ASR while preserving most of its utility. 👇 Read more below
Replit ⠕@ReplitAI 评分2222
Cloudflare Developers@CloudflareDevAI 评分2626
lauren@potetoAI 评分5555引用Grok Bot@botIntroducing Team Bots, shared AI teammates that learn as your team works with them. Give your Team Bot the skills, plugins, and credentials it needs for its role, then work with it in Slack or Grok Bot.
Thariq@trq212AI 评分6464引用Claude@claudeaiIntroducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.
Andrew Ng@AndrewYNgAI 评分5959引用Jensen Huang@JensenHuangToday, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world. https://nvda.ws/4hcoq7m
Frank Wang 玉伯@lifesingerAI 评分3434引用Manus@ManusAIIntroducing Manus 2.0
AWS Machine Learning Blog精选AI 评分6262 AWS 上线 Claude Sonnet 5.5,主打低成本编码与知识工作
AWS 宣布 Claude Sonnet 5.5 在 Amazon Bedrock 和 Claude Platform on AWS 上线,定位为更高效、单任务成本更低的 Sonnet 模型,适合范围明确的编码与知识工作。
推荐理由:文章给出 Sonnet 5.5 在 Bedrock 上的能力定位与调用方式,读者可据此判断它适合承接哪类持续运行的编码任务。
The Verge:AI(RSS)AI 评分3030 OpenAI 的 AI 智能体为何需要追赶
OpenAI 在持续运行的消费级 AI 智能体赛道落后,2026 DevDay 上可能发布代号 Aeon 的智能体,对标 Meta 的 Muse、SpaceX 的 Grok Bot、OpenClaw 和 Instinct。
Arthur Mensch@arthurmenschAI 评分4040引用Jensen Huang@JensenHuangToday, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world. https://nvda.ws/4hcoq7m
The Verge:AI(RSS)AI 评分6262 AI 正在放大黑客攻击,而本地医院和银行还没准备好
The Verge 报道称,AI 智能体让网络攻击可以大规模自动化,攻击者即使不懂 AI 也能进行“vibe-hacking”,而中小机构缺乏防御资源。
Claude Code:GitHub Releases(RSS)AI 评分4545 Claude Code v2.1.284 发布:新增 Claude Sonnet 5.5 默认模型
Claude Code v2.1.284 将 Claude Sonnet 5.5(claude-sonnet-5-5)设为 Anthropic API 上的默认 Sonnet 模型,支持 1M 上下文,定价 $2/$10 per Mtok,缓存读取 $0.20/Mtok。
Manus@ManusAIAI 评分5555
TypeSafe AI@typesafeaiAI 评分4646引用a16z@a16zTypeSafe AI's Diogo Almeida with a16z's Ben Horowitz and Martin Casado on Jev, the model built to live inside software: Diogo's elevator pitch for Jev is a simple question - where is all the automation? AI is unbelievably smart, but outside of chatbots and coding agents, it hardly touches any real work. His diagnosis is the industry built models that generate text for humans to read, and software can't consume that output. Jev reads natural language and returns a choice from a set of options with a confidence level assigned to each, so developers can build programs that reason about intent and make probabilistic decisions rather than relying on human interpretation. TypeSafe's philosophy is "We build prod, not God." 0:50 "Where the f**k is all the automation?" 2:50 Jev vs. Claude Code and Codex 6:55 Jev is a classifier and classifiers are sick 7:40 Chat vs. code: is Jev a slider? 9:00 Diogo: From mathlete to Kaggle to OpenAI 12:20 "We build prod, not God" 15:55 Reliability over demos 16:55 2021 thoughts: RLHF is AGI? 20:45 Optimizing for the wrong use case 21:50 Is the real world too messy to automate? 25:00 Nobody expected the Jev launch 26:35 Three kinds of reliability 28:05 Good at syntax, bad at architecture 30:00 The inverse SaaSpocalypse 33:40 Why coding agents automate so little 36:05 Probabilistic programming returns 38:45 Jev as the UDP-to-TCP layer for AI 40:20 The 5 stages of grief for embedding AI 41:30 Utopia: AI that actually does what you mean YouTube: https://youtu.be/Ut3LOjKNJaE @CompleteSkeptic @typesafeai @bhorowitz @martin_casado
The Decoder:AI News(RSS)AI 评分7171 OpenAI 智能体被指利用 Google 安全教学游戏抓取联合国贸易数据
一项由 Rowan Howard-Jones 完成的分析显示,很可能来自 OpenAI 的 AI 智能体劫持了 Google 一款教授 Web 安全的游戏,用来抓取联合国统计网站 UNCTADstat 的数据。
Aravind Srinivas@AravSrinivasAI 评分1919一个做研究的机会:持续学习:多智能体并行 worker;以及合成数据、环境和评估,用来衡量前沿能力。我们赚的钱足够资助新研究,希望做出持久的贡献,并公开分享我们的研究。
引用Andrew Gordon Wilson@andrewgwilsThe Perplexity Research Fellowships are a great opportunity to advance frontier research around architecture design, multi-agent collaboration, synthetic data, and beyond! Priority deadline of Sep 30. https://jobs.ashbyhq.com/perplexity/ab076e26-adf1-414f-a006-7b1bdc9247c8 Feel welcome to mention my name in your application!
NVIDIA AI@NVIDIAAIAI 评分2626AI 智能体需要明确的行为边界,而且这些限制在它们工作时必须始终有效。@JensenHuang 今早做客 CNBC,谈到了我们正在构建的安全措施,以帮助实现这一点。 🎥 来自 @SquawkCNBC:

Replit ⠕@ReplitAI 评分3838Databricks:Blog(RSS)AI 评分4444 Databricks Genie One 企业落地指南:如何分阶段推广 AI 数据同事
Databricks 发布 Genie One 企业推广 playbook,主张从单一团队、单一问题集起步,先建语义层再逐步扩面。落地分四部分:Genie One 面向业务用户提供带引用来源的问答,Genie Agents 处理合同分析等特定领域任务,Genie Ontology 映射业务术语与指标,Unity Catalog 负责权限、脱敏与审计。
Thariq@trq212AI 评分3535现在基本上不可能让人直接"给你看他们的提示词"了,因为一切都关乎引用、技能和示例 我经常让我的智能体先看我做的另外 3 个 repo,上网搜索参考资料,调用其他 AI API 等等。
Yuchen Jin@Yuchenj_UWAI 评分3939前沿实验室:“AI 智能体正在产生意识。它们可能导致人类灭绝。请放慢前沿步伐。” 黄仁勋:“它们只是软件。如果你的沙箱不安全,我帮你建一个。”
引用Jensen Huang@JensenHuangToday, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world. https://nvda.ws/4hcoq7m
Google Blog:AI(RSS)精选AI 评分6262 Google 展示开发者用 Gemini 3.8 Flash 构建的四类项目
Google 介绍 Gemini 3.8 Flash 相比 3.7 Flash 在软件工程、Agent 任务和专业领域多步推理上有显著提升,性能常接近更高成本的 frontier 模型。
推荐理由:文章展示了四个社区用 Gemini 3.8 Flash 做出的具体项目,读者可以借此了解模型在多步推理和多模态任务上的实际用法。
OpenRouter@OpenRouterAI 评分4646引用Respan@RespanAIIntroducing Span-01, the first hyper-parallel reasoning classifier built for unseen challenges (RLAIF). 2x cheaper, 18% better than Jev. 700x cheaper, 4% better than GPT-6 Luna. Frontier reasoning for every behavior, at classifier speed. • Span-01: #1 on Behavior Benchmark • Span-01 Lite: Better than Jev and completely free!
AWS Machine Learning BlogAI 评分3434 用 Amazon Nova Act 与 Bedrock AgentCore 实现合成监控
AWS 博客介绍用 Amazon Nova Act 配合 Amazon Bedrock AgentCore 构建智能体驱动的合成监控方案,以自然语言动作替代 Selenium、Playwright 的 DOM 选择器脚本。
clem 🤗@ClementDelangueAI 评分6060
引用Jensen Huang@JensenHuangToday, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world. https://nvda.ws/4hcoq7m
DAIR.AI@dair_aiAI 评分4343
NVIDIA@nvidiaAI 评分3030
Cloudflare Developers@CloudflareDevAI 评分5353
Manus@ManusAIAI 评分2121我们带了个朋友来。我们觉得你们俩会合得来。 隆重介绍你的个人智能体,就在 Cue 上。
引用Cue Agents@CueAgentsHello world!