Claude Code 现在会在你任务进行到一半触及 5 小时限额时,尝试找到一个优雅的停止点,而不是在编辑中途直接中断。它会从你的每周限额中提取一小笔固定额度,用来尽量收尾手头的工作。
#编码
#编码
今日 54 条
ClaudeDevs@ClaudeDevsAI 评分3737
OpenAI:官网动态(RSS · 排除企业/客户案例)AI 评分4040 Proaction 借助 Codex 提升销售 60%、每月节省 75+ 小时
Proaction 使用 Codex 后销售转化率提升 60%,每月节省 40–60 小时工程时间和 25–33 小时创始人时间。其联合创始人 Colin Knudsen 每月用 Codex 在 30–45 分钟内构建 4–6 个定制交互演示,并借助 Granola、Gmail、Slack、Linear、GitHub、HubSpot 等插件统一处理日常工作。
Mustafa Suleyman@mustafasuleymanAI 评分2020Copilot 团队做得太棒了。Autopilot 非常酷。快去看看!
引用Jacob Andreou@jacobandreouHome. Code. Autopilot. https://x.com/i/article/2103304083172167680
Noah Zweben@noahzwebenAI 评分6464
引用Noah Zweben@noahzwebenRolling out Claude Code Remote Control to Pro users - because they deserve to use the bathroom too . (Team and Enterprise coming soon). 🧻 Rolling out to 10% and ramping 1. Update to claude v2.1.58+ 2. Try log-out and log-in to get fresh flag values. 3. /remote-control
Claude@claudeaiAI 评分5555引用Ryan Sael@RyanSaelI asked Opus 5.5 to explain camera focus by building an interactive lens lab Here's what it came up with after 1 hour 26 minutes in one shot, $25.66 API cost https://lens.lab.sael.net Move the focus ring and you can see the glass elements shift the sharp plane through the scene
DAIR.AI@dair_aiAI 评分5151
karminski-牙医@karminski3AI 评分4343
Microsoft:Official Blog(RSS)AI 评分6363 微软发布新版 Copilot:新增 Home、Code 和 Autopilot 三大能力
微软宣布重塑 Microsoft Copilot,推出三大新能力:Home 作为统一入口整合 Chat 和 Cowork 并内置 Office;Code 让非开发者用自然语言构建应用,与 GitHub Copilot 同源技术并在沙箱中运行;Autopilot(原名 Scout)是可在租户内持续自主工作的智能体。
François Chollet@fcholletAI 评分4444引用Simon Willison@simonwThe more time I spend working with coding agents, the more convinced I am that they make software engineering even harder We can do amazing things with them, but unlocking their full potential requires extraordinary discipline and knowledge
Simon Willison 博客AI 评分2828 Simon Willison:编码智能体让软件工程变得更难
Simon Willison 在 9 月 24 日的笔记中表示,与编码智能体协作越多,他越确信这些工具让软件工程变得更难。他认为智能体能带来惊人成果,但释放其全部潜力需要极高的纪律性和知识储备。
Simon Willison 博客AI 评分1515 commit-rewriter 0.2 发布
Simon Willison 发布 commit-rewriter 0.2。该工具与 git 相关,具体功能与更新细节原文未展开说明。
Claude Code:GitHub Releases(RSS)AI 评分3737 Claude Code v2.1.282 发布
Claude Code v2.1.282 新增 maxProseWidth 设置,可限制宽终端中 Claude 正文宽度而表格与代码块保持全宽,并新增启动提示及 /status、claude doctor 中列出被忽略或关闭遥测的变量。
Databricks:Blog(RSS)AI 评分4747 Databricks 推出 Unity Gateway CLI,规模化部署与管理编码智能体
Databricks 发布 Unity Gateway CLI,让管理员集中配置编码智能体的默认模型、MCP 服务器、技能、Smart Routing 和支出策略,开发者用 `ug claude`、`ug codex` 等命令即可启动已批准的智能体。
Pragmatic Engineer(RSS)AI 评分5050 RoR 创始人 DHH 再掀“手写代码之死”争论
Ruby on Rails 创始人 David Heinemeier Hansson 在 Rails World 主题演讲中宣布,至少在 37signals,专业工作手写代码的时代已经结束。这一表态再度引发“手写代码之死”的争论,文章还提及 Amazon 和 Meta 在招聘方面遇到的困难。
ViggleAI@ViggleAIAI 评分3434Claude Opus 5.5 在浏览器中一次性构建出 Minecraft……结果太疯狂了! 这是 Claude Opus 5.5 x @Viggle_PINOC MCP 用于游戏开发的又一个例子。
引用PINOC@Viggle_PINOCClaude Opus 5.5 with PINOC MCP one shotted this: A playable Minecraft recreation Link, prompt and more examples below 🧵
GitHub BlogAI 评分2424 GitHub Copilot 应用如何渲染超大 pull request
GitHub Copilot 应用重建了 pull request 视图,用一个含 2200 个文件、超 100 万行改动和 400 多条行内评论的开源 PR 做压力测试。其做法是把文档高度拆成确定性的代码几何与动态评论块两套几何:代码行高提前精确计算,评论高度则按需测量、修正幅度小且锚定在用户当前查看位置,从而避免滚动跳动。
Pragmatic Engineer(RSS)AI 评分6060 Gergely Orosz 访谈 GitHub Next 设计工程师 Maggie Appleton:AI 时代的设计工程实践
Pragmatic Engineer 播客访谈 GitHub Next 研究工程师 Maggie Appleton,探讨设计师如何与 AI 协作。
karminski-牙医@karminski3AI 评分2929
MeshyAI@MeshyAIAI 评分2323Latent Space(RSS)精选AI 评分8484 Claude Opus 5.5 发布成新默认模型,OpenAI 同日推 GPT-6 Sol 和 Luna 降价应战
AINews 汇总:Anthropic 发布 Claude Opus 5.5,称多数任务达 Fable 5.1 水平、比 Opus 5 便宜 40% 且快 30%,成为 Claude Code 和应用新默认模型;OpenAI 数小时后推出 GPT-6 Sol($2/$10)和 Luna($0.10/$0.50),价格比 GPT-5.6 前代低约 50%。
推荐理由:除了两家降价发布,原文还汇总了缓存定价拆解、effort 异常和第三方评测口径问题,提供了单一官方稿之外的交叉信息。
Simon Willison 博客AI 评分2323 Simon Willison 与 Jesse Vincent 将于 10 月 14 日在旧金山举办 Agentic Engineering 交流活动
Simon Willison 与 Jesse Vincent 将于 10 月 14 日在旧金山举办一场面向 coding agent 构建者的晚间交流活动,形式为非正式的 show-and-tell。活动欢迎分享尚未公开的尝试、奇怪实验或没有明确市场的不完整项目,不要求做演示,也不是产品推介。
Together AI 研究与产品博客(RSS)AI 评分6262 Together AI 教程:花 17 美元微调自己的 Jev 式分类模型
Together AI 发布教程,演示如何以 Qwen3.5 4B 为基础模型,用 MultiNLI、BoolQ、Banking77 等 6 个数据源共 37,840 条样本微调一个 Jev 式分类模型,训练成本约 $17.0,耗时约 25 分钟。
Simon Willison 博客精选AI 评分8484 Anthropic 发布 Claude Opus 5.5,OpenAI 同日推出 GPT-6 Sol 和 GPT-6 Luna 掀起新一轮价格战
Anthropic 于9月22日发布 Claude Opus 5.5,约一小时后 OpenAI 发布 GPT-6 Sol 和 GPT-6 Luna。GPT-6 Luna 价格降至 $0.10/M 输入、$0.50/M 输出,为 GPT-5.6 Luna 的一半;GPT-6 Sol 同样减半至 $2/$10。
推荐理由:作者用自己实测的价格表和 pelican 测试对比了三款新模型,还发现 Opus 5.5 max 会想满输出上限,可直接参考。
Boris Cherny@bchernyAI 评分4949
karminski-牙医@karminski3AI 评分3131
Greg Brockman@gdb精选AI 评分7979引用OpenAI@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
推荐理由:原文给出了两个新模型的能力来源与 API 降价幅度,读者可据此评估在大规模任务中替代 GPT‑5.6 的成本。
ChatGPT@ChatGPT精选AI 评分7676
推荐理由:官方账号宣布两款新模型当天上线及覆盖的订阅层级,读者可快速确认自己能否在 ChatGPT Work 和 Codex 中用到。
Charlie Holtz@charlieholtzAI 评分4545
Boris Cherny@bcherny精选AI 评分7070引用Claude@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.
推荐理由:作者实测对比 Opus 5.5 与 Fable 5.1 移植 HAProxy 的用时和成本,给出第一手数据供选型参考。
Claude Code:GitHub Releases(RSS)精选AI 评分6565 Claude Code v2.1.280 发布:新增 Claude Opus 5.5 默认模型
Claude Code 发布 v2.1.280,新增 Claude Opus 5.5(claude-opus-5-5)并设为默认 Opus 模型,支持 1M 上下文,价格 $4/$20 per Mtok、缓存读取 $0.20/Mtok;Pro 和 Team Standard 计划默认模型也从 Sonnet 改为 Opus。
推荐理由:原文列出该版本新增 Claude Opus 5.5 默认模型、MCP 描述长度可配置等改动,读者可对照修复清单决定是否升级。
StepFun@StepFun_aiAI 评分6060
karminski-牙医@karminski3AI 评分5454
Kimi.ai@Kimi_MoonshotAI 评分4848elsewhere:文章(RSS)AI 评分6868 阶跃 Step 5 Preview 实测评测:数据可视化与金融分析亮眼,泛化和审美仍有短板
阶跃发布 Step 5 Preview,总参数量 600B、激活参数 27B,有视觉输入,官方称在 Artificial Analysis 上涨 44 分、单任务成本仅为 Claude Opus 5 的 1/8。作者与友人实测发现其在数据可视化、金融分析上表现不错,但泛化性、领域知识和审美偏弱,思考过程过长导致长任务耗时且易中断,且长上下文下安全指令遵循可被绕过。
Xiaomi MiMo@XiaomiMiMoAI 评分4040引用Design Arena@DesignArenaBREAKING: MiMo-V2.6-Pro by @XiaomiMiMo lands at #8 overall (#3 open-weight) on Design Arena with an Elo of 1338. This is an impressive 54-point and 22-position increase from MiMo-V2.5-Pro. MiMo-V2.6-Pro also reaches #4 overall in Website (#2 open-weight) and #6 overall in Agentic Frontend Development (#2 open-weight), showing strong performance across both direct generation and agentic coding. This places @XiaomiMiMo's new model among leading models such as GPT-5.6 Sol by @OpenAI, Claude Opus 5 by @Anthropic, and Kimi K3 by @MoonshotAI. Congratulations to the @XiaomiMiMo team on returning to a top-10 placement on Design Arena!
Xiaomi MiMo@XiaomiMiMoAI 评分5656引用Arena.ai@arenaMiMo-V2.6-Pro just landed @XiaomiMiMo back in the top 10 on Code Arena: WebDev, debuting at ~#10 overall, and ~#3 among open-weights models with an MIT license. It scores 1628 pts (AutoEval), tying Claude Fable 5 (High) and just ahead of Hy4-preview (1624 pts). That's a +153 pt jump from the previous MiMo-V2.5-Pro (1475 → 1628 pts). Among open-weights models, it lands at ~#3. Impressively only 7 pts behind Qwen3.8 Flash Next (#2) and 46 pts behind Kimi K3 Max (#1). Note: this is an early AutoEval score, in which a Reward Model trained on Arena’s human preference data casts automatic votes in place of live votes. We’ll continue to see how scores converge as more live human votes come in. Congrats to the @XiaomiMiMo team on this release!
Xiaomi MiMo@XiaomiMiMo精选AI 评分6666推荐理由:原文给出 Pro 与 Claude Opus 5、GPT-5.6 Sol 的 agent 基准对比和开源范围,可据此评估其相对位置。
Google AI Developers@googleaidevsAI 评分2626
eric zakariasson@ericzakariassonAI 评分7575
小米 MiMo:GitHub 新仓库(模型发布)AI 评分5656 小米 MiMo 开源 mimoagent:百行代码智能体在 SWE-bench Verified 得分超 74%
小米 MiMo 在 GitHub 开源 mimoagent,一个仅约 100 行代码的 AI 智能体,可解决 GitHub issue 或在命令行中辅助用户。项目主打极简设计,无需庞大配置和大型 monorepo,并在 SWE-bench Verified 上取得超过 74% 的分数。仓库地址:https://github.com/XiaomiMiMo/mimoagent