跳到正文

全部动态

今日 64 条
9月30日周三
  1. Dongxi 东锡 NLP47

    一条推文设想中大型公司 AI Engineer 的完整工作流:从读 Jira Ticket、分析 LangSmith trace、写代码、配环境跑本地代码到准备和修 PR,几乎每步都可交给 Agent,人只在第二天 standup 上表演 Agent 总结的内容。推文最后反问:以上所有环节,哪个是 dots 做不了的?背景显示 dots 由 GPT-6 Astra 驱动,定位为常驻在线的智能体。

    引用OpenAI@OpenAI

    Introducing dots, powered by GPT-6 Astra. Remarkably capable, always-on agents built to handle everything.

  2. Sherwin Wu56

    Sherwin Wu 转引 @OpenAIDevs 消息并评论:当模型推理基本即时后,期待看到人们接下来抱怨什么新瓶颈。其引用的内容称 Codex 中的 Ultrafast 是迄今最快的 Astra 构建方式,比 Astra Standard 快至 8 倍、比 Astra Fast 快至 4 倍。作者称这是一种 Divinely discontent(神圣的不满足)。

    引用OpenAI Developers@OpenAIDevs

    Ultrafast is our fastest way to build with Astra yet: in Codex, it runs up to 8x faster than Astra Standard and 4x faster than Astra Fast. Bring your ideas to life as fast as you can type them.

  3. Nathan Lambert25

    像这样的人的问题不是他们笨什么的,Timnit 拥有顶尖科学家的全部技能,问题在于他们所处的信息生态和同侪群体不鼓励对思想进行拷问。回音室是清晰思考的慢性死亡。

    引用Alec Stapp@AlecStapp

    Timnit Gebru doubles down on the "stochastic parrots" framing, saying you "cannot expect LLMs to be factual." As evidence to support this, she cites errors in... Google AI Overviews. We need to start a GoFundMe to pay for these people to have access to Opus 5.5 and Astra.

9月29日周二
  1. Ethan Mollick47

    AI对就业的影响目前仍相当不明确。我并不觉得这很意外,因为雇主们还在摸索个人层面的采用,更不用说在复杂组织中如何使用AI了。随着能胜任实际工作的智能体普及,情况可能会开始加速变化。

    引用Jacob Schaal@FutureEconJacob

    Has AI hit the labor market yet? @alexolegimas and my verdict after ~20 papers: not yet in aggregate. Unemployment and layoffs show almost nothing. But AI may already be cutting junior hiring in exposed white-collar jobs. Remote work may explain part of that decline. A 🧵

  2. Aravind Srinivas38

    AI 安全是一个工程问题

    引用Perplexity@perplexity_ai

    Agent governance is an engineering problem. We’ve built safeguards into Perplexity’s infrastructure, harnesses, and tools, and put them to work across our products. Today we're sharing how we engineer safer agents: https://www.perplexity.ai/hub/blog/how-we-engineer-safer-agents

  3. Ars Technica:AI(RSS)22

    Mozilla Firefox 157 重新设计界面,负责人谈如何从 Chrome 争夺用户

    Mozilla 随 Firefox 157 在桌面和移动端推出界面重新设计,希望借此吸引隐私意识极客和开源倡导者之外的更广泛用户。Firefox 负责人 Ajit Varma 表示团队正借助 AI 工具提升开发速度,并恢复紧凑模式、增加自定义选项,让浏览器在体验上区别于基于 Chromium 的竞品。

  4. MIT Technology Review · AI12

    HPE:如何让 AI 从费用变成资产

    HPE 提出企业 AI 从按 token 消费转向自建容量的判断框架:当需求稳定、可预测且规模足够时,拥有算力可能比逐次购买更经济。文中引用 Deloitte 2026 企业 AI 状况报告称,2025 年员工 AI 使用率上升 5%,至少 40% 的 AI 项目进入生产的公司比例预计半年内翻倍。企业需先回答三个问题:需求是否稳定、在什么使用水平下自建更划算、能否通过采用与治理让容量保持高产。

  5. Ethan Mollick26

    嘿,Claude:“我让早期的 Claude 把《西线无战事》里的鱿鱼移除掉。现在你都能做电影之类的事了。我需要你展示一下你进步了多少” & 我把下面这条推文粘贴了进去 有些巧妙的东西,模型现在真的有幽默感了

    引用Ethan Mollick@emollick

    👀Claude handles an insane request: “Remove the squid” “The document appears to be the full text of the novel "All Quiet on the Western Front" by Erich Maria Remarque. It doesn't contain any mention of squid that I can see.” “Figure out a way to remove the 🦑​​​​​​​​​​​​​​​​“

  6. TypeSafe AI49

    Jev 不好用?它是为可组合性而生的! 需要更多 Jev!

    引用Zhaorun Chen@zrrrr_cn

    Jev is fast at helping you. Turns out, it can also be fast at helping an attacker!! 😱🚨 We red-teamed Jev 1.13 on our DTap (DecodingTrust-Agent Platform) and found a serious safety gap: 70.1% ASR under direct misuse 43.5% ASR under indirect prompt injection In our evaluations, we found that under indirect prompt injection, Jev can follow attacker-injected instructions without blinking an eye, e.g., exfiltrating user data, deleting files, or taking other harmful actions. But we found a much safer way to integrate Jev: use it as a self-gating layer for its own tool calls, significantly reducing ASR while preserving most of its utility. 👇 Read more below