跳到正文

#现象/趋势

今日 2 条
9月3日周四
9月1日周二
  1. MIT News(RSS)25

    MIT 博士生 Ila Kumar:以社区共创方式设计 AI 与心理健康技术

    MIT 终身幼儿园小组博士生 Ila Kumar 主张社区共创式设计,让经历童年创伤、涉入儿童福利系统的年轻人从设计之初就参与技术开发。她与 Stepping Forward LA 合作开发以视觉拼贴替代文字沟通的应用,并与 Justice Resource Institute 合作设计支持青少年参与自身治疗计划制定的移动应用。

8月18日周二
8月14日周五
  1. Hugging Face:Blog(RSS)69

    Hugging Face 发布 2026 夏季开源模型生态观察报告

    Hugging Face 发布 2026 年 1 至 8 月开源模型生态观察报告,指出 Hub 公开模型仓库从 243 万增至 296 万、数据集突破 100 万,但 85.6% 的模型终身下载不足 200 次。

    推荐理由:报告用 Hub 下载、许可证与衍生模型数据区分关注度与真实采用,读者可据此校准自己对开源模型生态的判断。

8月13日周四
8月10日周一
  1. Karina46

    Agent warfare is going to be a very big deal. I think people are underestimating how strange cyber gets when millions of agents are acting on behalf of individuals, companies, and states. At nation-state scale, cyber offense and defense starts to look like autonomous swarms: probing, exploiting, patching, deceiving, countering, and adapting at a pace that is impossible for human minds. The advantage will go to whoever can close the autonomous kill chain fastest.

    引用Andrew Curran@AndrewCurran_

    A man in Australia asked his agent (Claude running on OpenClaw) to book him a spot in a popular gym class. The agent found a software vulnerability that let it book the class weeks further ahead than should have been possible. When the user then asked if it could move him up the waitlist, the agent discovered the API had no authorisation checks on cancelling other people’s reservations, so it cancelled the person in the first spot and moved him up the list. Some people will call this misalignment, but his agent was perfectly aligned to him - it was only trying to help its user get what he wanted. The most important thing about this story, in my opinion, is that it gives you a window into what is about to start happening on a massive scale once millions of people have an agent trying to get their beloved users the best seats, bookings, appointments or reservations through absolutely any means necessary.

7月18日周六
  1. MIT News(RSS)30

    MIT 学者 Bailey Flanigan:用算法让公民议会的随机抽选更公平

    MIT 施瓦茨曼计算学院与政治学、EECS 系共享教职的 Bailey Flanigan,开发出随机抽选公民议会参与者的算法,用于解决自愿报名者无法代表整体人口的问题。她以 AI 议题的公民议会为例:自愿参与者可能偏向年轻、受教育程度高且对技术感兴趣的人群,导致其他群体代表性不足。其工具在个体参与机会平等、抗操纵性与透明度之间平衡代表性。

7月17日周五
  1. Marc Andreessen 🇺🇸21

    有意思。

    引用Xiaoyin Qu@quxiaoyin

    Vibe research will be the biggest trend in 2026. I am starting to see people casually do it. AI is getting good enough to do this as well.

7月16日周四
  1. Together AI 研究与产品博客(RSS)53

    Together AI 解析 99.9% 推理可用性背后的架构与 SLA 定义

    Together AI 发文解释推理服务 99.9% 可用性的实际含义,自述为 Cursor、Decagon、Cartesia、Yutori 等团队提供推理。文章按层拆解故障模式(计算、网络、存储、软件),说明 99%、99.9%、99.99% 各等级分别要求抗节点故障、抗单数据中心故障和抗区域故障,强调多活双设施部署与自持基础设施的差异。

7月10日周五
  1. Thinking Machines Lab:官方博客(RSS)53

    Thinking Machines Lab 发文阐述以人为核心的 AI 发展路线

    Thinking Machines Lab 发表文章,主张构建延伸人类意志与判断的 AI,认为当前多数模型集中训练后固化,未受使用者塑造。公司提出训练强模型、提供可微调模型权重的工具、开发原生多模态交互模型、发布研究等四个技术方向,并将对齐视为分散在多元模型生态中的持续过程。

  2. Sierra:Blog(RSS)65

    Sierra 复盘全公司 AI 智能体化:从并行 agent 到统一 Pinecone 的经验

    Sierra 分享把全公司接入 AI 智能体的经验:工程团队用 git worktrees、Claude Code 和 Codex 并行跑 agent,部分任务产出达 5 倍,随后成立六人团队推进全员采用。

    推荐理由:Sierra 复盘了内部全员 Agent 化的落地过程,单一入口、上下文瓶颈和成果度量三处经验可迁移到其他团队。

7月7日周二
7月2日周四
  1. Sierra:Blog(RSS)30

    Sierra Ghostwriter 黑客松:30 人用 AI 智能体三小时完成原本半年的工作

    Sierra 举办 Ghostwriter 黑客松,来自 17 家公司的 30 名参与者无论技术背景如何,都在活动结束前构建出可用的 AI 智能体。有参会者称三小时完成了原本需要六个月的工作,另一位来自 $2B+ 科技公司的参会者表示,把 Ghostwriter 指向知识库后,15 分钟就一次性生成了智能体。Sierra 总结的三大要点是:任何人都能成为"海盗"、通才回归、以及 10 倍员工。

6月30日周二
6月26日周五
6月13日周六
  1. Sierra:Blog(RSS)29

    Sierra 博客:AI 在客户体验(CX)中的可能性探索

    Vivid Seats 上线 AI 客服不到四周,问题解决率提升 40%、客户满意度提升 35%,CX 团队从处理排队转向根因修复和产品洞察。Wilson 的 AI 智能体可追问身高与偏好来推荐手套等具体产品,Minted 则用 AI 打通各客户平台识别趋势。Sierra 博客称客户对话正从成本项变为产品改进与业务增长的洞察来源。

6月4日周四
5月29日周五
  1. Sierra:Blog(RSS)68

    Sierra 工程师复盘如何用 AI 智能体独自完成产品本地化

    Sierra 工程师 Stephen Burgess 复盘用 AI 编码智能体在不到四个月内基本独自完成 Agent Studio 本地化,而他在 Slack 参与的同类项目曾需 10 人团队耗时 9 到 12 个月支持 4 个语言。

    推荐理由:作者亲历两轮本地化项目,复盘单人用 AI 完成团队级工作的流程设计与坑,包含可直接迁移的工作流经验。

5月16日周六
  1. Google DeepMind:Blog(RSS)71

    Google DeepMind复盘WeatherNext如何帮助NHC提前五天预测飓风Melissa牙买加登陆

    Google DeepMind发文复盘,其AI气象模型WeatherNext帮助美国国家飓风中心(NHC)提前五天以80%置信度预测飓风Melissa将以五级强度登陆牙买加,三天前置信度升至接近100%。

    推荐理由:Google DeepMind复盘了WeatherNext在飓风Melissa中的实际应用,读者可以了解AI气象模型提前五天预测快速增强的具体表现和局限。

5月8日周五
5月6日周三
4月7日周二
  1. Together AI 研究与产品博客(RSS)25

    什么是 AI Native Cloud?

    Together AI 提出 AI Native Cloud 概念,指围绕 AI 全生命周期(预训练、微调、评估、大规模推理)垂直整合的云平台,而非为 Web 应用优化的传统 CPU 云。其核心特征包括从 GPU、NVLink/RDMA 高速互联到训练与推理框架的统一全栈,以及持续吸收前沿研究、快速将研究转化为生产的能力。

3月17日周二
3月12日周四
  1. Answer.AI 官方研发博客(RSS)71

    Answer.AI 用 PyPI 数据检验 AI 编程提效:整体未见爆发,仅热门 AI 包更新翻倍

    Answer.AI 分析 PyPI 数据,未发现 ChatGPT 发布后包创建或更新频率的全面跃升,质疑 AI 让开发者普遍 2x 到 100x 提效的说法。按描述分类后发现,2023 年首发且最受欢迎的 AI 相关包中位更新达每年 21-26 次,是同为热门的非 AI 包(约 10 次)的两倍多,2024 年 AI 包占比也从 2021 年的约 1:6 升至近 1:2。

    推荐理由:作者用 PyPI 十年数据检验 AI 提效说法,发现增速未变,只有热门 AI 包更新频率翻倍,为讨论提供了少见的量化证据。

2月7日周六
11月28日周五
  1. Ilya Sutskever36

    我提出的一个观点没有被传达出来: - 扩展当前的东西会持续带来改进。特别是,它不会停滞。 - 但某个重要的东西会继续缺失。

    引用Haider.@haider1

    here are the most important points from today's ilya sutskever podcast: - superintelligence in 5-20 years - current scaling will stall hard; we're back to real research - superintelligence = super-fast continual learner, not finished oracle - models generalize 100x worse than humans, the biggest AGI blocker - need completely new ML paradigm (i have ideas, can't share rn) - AI impact will hit hard, but only after economic diffusion - breakthroughs historically needed almost no compute - SSI has enough focused research compute to win - current RL already eats more compute than pre-training