Microsoft 发布 Amplifier 项目记忆功能包 amplifier-bundle-memory
Microsoft 在 GitHub 上线 microsoft/amplifier-bundle-memory,为 Amplifier 项目提供记忆功能包。该仓库定位为 Amplifier 的记忆 bundle,目前公开信息仅说明其用途,未披露具体实现细节与可用方式。
Microsoft 在 GitHub 上线 microsoft/amplifier-bundle-memory,为 Amplifier 项目提供记忆功能包。该仓库定位为 Amplifier 的记忆 bundle,目前公开信息仅说明其用途,未披露具体实现细节与可用方式。
vLLM 推出 TT Plugin,通过 out-of-tree 平台插件机制将 Tenstorrent 加速器接入 vLLM,安装后只要 ttnn 可导入即自动注册为 vLLM 平台,OpenAI 兼容 API 与请求格式保持不变。
Microsoft 在 GitHub 新建了 microsoft/foundry-iq-v2 仓库,目前访问该仓库需先完成仓库设置。
browser-use 在 GitHub 发布新仓库 browser-use-pi,一个小型 TypeScript 网页智能体,通过真实浏览器评测进行爬山法优化。
GitHub 推出 Project HydraFusion 研究预览,通过运行时多模型编排提供前沿级智能,所有 GitHub Copilot 计划用户可通过 Copilot CLI 的 /experimental 使用。
推荐理由:官方详解了 HydraFusion 三种执行模式与基准成本质量数据,读者可据此判断多模型编排是否值得在 Copilot CLI 中试用。
GLOBAL BUILD is live 🌍🔥 GLM-5.3-Flash, FREE in ZCode. 10 hours a day. Sep 3 – 18. Who gets it 👑 Coding Plan members — free every day, all 15 days. Our VIPs go first. 🥚 New users — 100M free tokens on sign-up. One-time, valid until the window closes When it opens 🇺🇸 8 AM PT · 11 AM ET 🇬🇧 4 PM London 🇨🇳 11 PM Beijing How to claim 1️⃣ Open ZCode → tap the card in the bottom-left corner 👀 2️⃣ Not there? Restart the app 😏 http://zcode.z.ai
Grok Bot for Enterprise is available today. It’s free for all Grok and Cursor enterprise customers for the next two weeks. https://x.ai/news/grok-bot-for-enterprise
Hermes Desktop now sets up local models in one click. It automatically reads your hardware, picks the best model for you, then downloads it and configures the runtime.
Midjourney 向 alpha.midjourney.com 推送大规模更新,最大变化是新的 v8.2 编辑模型上线,编辑器直接内置于 lightbox,可用自然语言指令编辑、附加最多 4 张参考图并集中查看本次会话的编辑。
browser-use 在 GitHub 上线新仓库 agency,提供 Agency card feed:智能体完成任务后,由 Magnus 一键完成决策。该仓库目前仅披露这一交互流程,未公布模型、参数或可用性细节。
wrote down some of the design thinking behind Grok Bot. persistent roles, clear state, scoped context, coordinated teams — an interface designed to move you from operating AI to delegating work. https://x.ai/news/designing-grok-bot
我们的短视频概览国际扩展已在网页端向所有用户100%全量推送(移动端即将推出!) 70+ 种新语言和 3 种新英文变体,满足你所有的短视频需求。 你觉得怎么样?(我们自己挺满意的。)
NVIDIA 推出 PAIR 虚拟推理路由器,可将本地网络中的可用算力扩展给 AI 智能体使用。该方案面向多智能体协作场景:主智能体把复杂任务拆解后分派给专用子智能体,用户也会同时运行多个智能体会话。这种广度优先的方式可提升任务完成速度。
Hugging Face 发布开源工具 funes,为 Claude Code、Codex、pi、Hermes 等 coding agent 提供基于本机已有会话日志的持久记忆层。
推荐理由:官方介绍 funes 如何把本机已有 agent 会话变成可检索的本地记忆,给出安装方式、跨 agent 共享和成本对比。
Fable 5.1 makes Claude Tag even more useful. Here it builds a last-minute leadership deck from a metrics spreadsheet and other data across Slack, spots a vendor report that disagrees with the numbers, and flags it before moving on. Claude Tag is available in Slack on Team and Enterprise plans.
LongCat-2.0 is now free in Command Code. 1.6T params. 1M context. 48B activate. Available on all plans to all subscribers. npm i -g command-code 🐐
Google 推出 Fairwind Program,面向 Google Cloud 客户、政府机构和网络安全合作伙伴提供受限访问,首批开放 Gemini 3.8 Flash Cyber 与 CodeMender 组合,用于自主发现、验证并修复漏洞。
推荐理由:原文给出受限访问计划的能力组合与准入对象,读者可据此判断前沿模型在漏洞修复环节的落地方式。
🎬Video generation faster than playback! 🚀MiniMax H3 on vLLM-Omni + FastVideo's FastH3: a complete 10.1s MP4 - video AND synchronized audio - rendered in 8.7s!⚡️ Thanks to @MiniMax_AI for the great Minimax H3 release, the FastVideo team @haoailab for open-sourcing FastH3 and helping on the serving integration, and @NVIDIAAI for the continued sponsorship and joint optimization efforts!
Google DeepMind 推出 agentic video understanding,覆盖 Gemini 3.7 Flash、3.6 Flash 和 3.5 Flash-Lite,通过智能体循环动态调用原生视频工具按需检索画面、音频和字幕,而非固定帧率静态处理。
推荐理由:原文给出了具体降本增效数字、适用模型和接入方式,开发者可据此评估是否切换视频分析流程。
不止是双臂。 具身 AI 的起点。 LEMO by Galaxea —— 一个平台,开箱即用,即刻开始构建。 🦾 让创造更快发生。
vLLM-Omni 团队发布 MiniMax H3 服务优化详解,先对完整管线做无损系统级优化,相比 Diffusers 将完整响应延迟降低 30.8%(1.445x)。
Hugging Face WebAI 团队发布 @huggingface/kernels 库和 207 个 Apache-2.0 许可的 WebGPU 内核,每个内核以独立版本化仓库形式发布在 Hub 上,包含 manifest、正确性测试、基准案例和 WGSL 着色器模板。
NVIDIA 发布指南,介绍如何在 Claude Science 中运行 BioNeMo NIM 微服务完成蛋白质结构预测。该方案面向可读取论文、提出假设并调用模型的 AI 科学家,用于判断后续实验优先级。正文指出,科研比软件工程更依赖反复评估证据与修正假设,编码智能体已在生产代码中验证价值。
NVIDIA 介绍用 Omniverse NuRec 解决同一套感知软件换装到 SUV、轿车等不同车型后传感器布局、标定、视场与遮挡变化导致感知结果改变的问题。原文未披露具体版本号、参数规模或性能数字。
We have a huge news to share today! Today we are unveiling the first truly accessible RL robot - welcome Microduck A 25 cm tiny open-source biped with 15 actuators and packed with sensors (camera, speaker, LiDAR, NFC, bluetooth, wifi, etc) that you train yourself with reinforcement learning. It's also playable out of the box with more than half a dozen fun and playful pre-trained policies to have it walk, sit, crouch, roller-skate, pick up objects with its articulated beak, and recover on its own. And all for less than $400. See all the details, play with the simulator and order it at: https://pollen-robotics.com/microduck/ (video with sound on 🔊)
At Reactor we ❤️ open-source. We teamed up with @haoailab to ship an infinite live stream powered by FastVideo’s FastH3: https://twitch.tv/dereactor And we’re open-sourcing everything!
Nous Research 推出 Hermes Agent 的 BackSearch 插件,可在冻结的新闻存档上做时间点(point-in-time)网页搜索与抓取。该插件归属 General Reasoning 方向,仓库名为 NousResearch/hermes-plugin-backsearch。
Midjourney 更新了 V8.2 编辑模型,图像质量得到提升。官方建议过去 24 小时内遇到问题的用户重新尝试,并继续反馈意见。更多更新即将推出。
我们更新了 V8.2 编辑模型,图像质量更好了。如果你在过去 24 小时内遇到任何问题,请再试一次,并继续给我们反馈。谢谢!更多更新即将到来。
NVIDIA TensorRT Model Connect 提供开源参考实现集合,让开发者用两条命令把开源模型从 checkpoint 直接跑成 TensorRT 推理。它面向原生 C++ 应用,省去针对具体模型的转换、预处理、后处理与运行时代码编写。
Midjourney 开始让所有用户测试首个 V8.2 图像编辑模型。该模型支持用指令编辑图像、以最多 4 张图像参考生成新图(替代 omni-reference)、局部重绘(inpainting)与扩图(outpainting),并可用 personalization、moodboards 和 srefs。
推荐理由:官方宣布 V8.2 图像编辑模型开放测试,列出指令编辑、多图参考、局部重绘等能力变化和具体入口。
Google 在 Earth AI 之下推出实验性研究能力行星预测引擎(PPE),从自然语言查询出发自主完成数据发现、特征工程、模型训练与评估的全流程。