NVIDIA Topograph:拓扑感知的 GPU 工作负载调度
NVIDIA 推出 Topograph,用于拓扑感知的 GPU 工作负载调度。AI 工厂是功率受限系统,工作负载放置不当会割裂拓扑域、迫使流量走共享链路,降低吞吐并推高作业成本,GPU 还会在等待数据时白白消耗已分配功率。
NVIDIA 推出 Topograph,用于拓扑感知的 GPU 工作负载调度。AI 工厂是功率受限系统,工作负载放置不当会割裂拓扑域、迫使流量走共享链路,降低吞吐并推高作业成本,GPU 还会在等待数据时白白消耗已分配功率。
Simon Willison 发布 llm-anthropic 0.29,这是 llm 工具接入 Anthropic 模型的插件更新。原文未披露该版本的具体功能、参数或可用性变化。
Claude Code 发布 v2.1.280,新增 Claude Opus 5.5(claude-opus-5-5)并设为默认 Opus 模型,支持 1M 上下文,价格 $4/$20 per Mtok、缓存读取 $0.20/Mtok;Pro 和 Team Standard 计划默认模型也从 Sonnet 改为 Opus。
推荐理由:原文列出该版本新增 Claude Opus 5.5 默认模型、MCP 描述长度可配置等改动,读者可对照修复清单决定是否升级。
.@Grok in your Tesla can now do meaningful work for you With Connectors, you can manage your inbox, clean up your calendar, or talk through existing files/chat/tasks – all hands-free
Simon Willison 发布 LLM 插件 llm-typesafe 0.1a0,为 LLM 命令行工具接入 TypeSafe AI 的新模型 Jev。
🎨 Hy Image3.5 preview 已在 Miora 上线。编辑时保留已经好用的部分——同一画布,你的品牌规则已被记住。免费两周。
Hy Image3.5 preview is live on Miora. FREE through October 7. Campaign visuals. Product scenes. Storyboards. Game concepts. Up to 2K. Text-to-image and image-to-image. Try it: https://miora.design/
NVIDIA 发布 DLSS 5,引入 DLSS 3D-Guided Neural Rendering 和细粒度控制,帮助开发者在保留艺术意图的同时加入逼真光照与材质细节。同步更新 NVIDIA ACE、RTX Mega Geometry 2.0 和 RTX Kit,覆盖角色 AI、高密度几何与渲染工作流。
Apple 宣布新款 Mac mini 和 Mac Studio 今日上市。Mac mini 搭载 M6(4x AI 性能、2x 图形和存储、40% CPU 性能提升。
推荐理由:官方公告给出两款新 Mac 的芯片、AI 性能倍数、统一内存上限和定价,读者可以据此评估端侧 AI 与升级选择。
NVIDIA 在加拿大多伦多 ROSCon 大会上发布 Isaac ROS 5.0,为基于 ROS 的机器人开发引入智能体工作流和新的平台支持。
推荐理由:Isaac ROS 5.0 把智能体工作流引入机器人开发,并给出从 Jetson Orin 到 Thor 的部署路径,可据此判断机器人开发方式的变化。
Nous Research 在 GitHub 发布 hermes-nvidia 插件,为 Hermes 提供 NVIDIA App 和 NVIDIA Broadcast 的 MCP server 与技能,目前仅支持 Windows。该插件为便携式 Agent Plugin,功能开关位于 extensions["com.nousresearch.hermes"] 下。
Meshy 🤝 chrona 让你的 3D 创意尽情释放 🤩
What if Tech Week wasn’t a calendar, but a city you could actually enter and play? We made a SF TECH WEEK CITY — rebuilding San Francisco as a 3D world with Tech Week events placed at their real venues. RSVP activity, trending events, and busy areas become part of the world itself. The birds flying above the city represent real community members, so you can see where people are heading and explore alongside them. Built with AI coding and published on Chrona. A small experiment in turning real-world data into a place worth entering. @MeshyAI @ChronaWorld #MeshyxChrona #TechWeek
Microsoft 在 GitHub 新建 microsoft/synthetic_proposals 仓库,定位为 AI Night Scientist 的配套服务。仓库名称与来源信息显示其与合成提案(synthetic proposals)相关,目前公开信息仅有仓库名与一句简介,未披露具体功能、模型或可用性细节。
Together AI 上线 endpoints rollout 功能,支持 Canary、蓝绿、滚动三种策略在同一端点的两个部署之间迁移流量,实现生产模型不停机升级。
vLLM 官方宣布 vllm-metal 首个正式版本 v0.28.0,将上游 vLLM 的 V1 调度器、分页 KV cache 和 OpenAI 兼容服务带到 Apple Silicon,模型执行由 MLX 和 Metal 完成。
推荐理由:官方详解 vllm-metal 的架构与跨引擎实测数据,读者可据此评估 Apple Silicon 上多并发推理是否值得切换。
Hugging Face 为 transformers 加入 GGUF 模型支持,用户可从 Hub 选取 GGUF 检查点,通过 from_pretrained 传入 gguf_file 在本地生成。
推荐理由:Hugging Face 把 GGUF 量化模型接入 transformers,读者可据此判断本地推理工作流会如何变化。
NVIDIA 在 Dynamo-Triton 中集成 TensorRT 多设备推理,让单个 TensorRT 网络借助 NCCL 分布式集合通信跨多 GPU 执行,同时保留 TensorRT 推理优化。该能力自 TensorRT 11.0 起获得完整支持,用于应对生成式 AI 超出单 GPU 的算力与显存需求。
We put Grok 4.7 from @SpaceXAI to work on a $2 million insurance claim where one deductible error alone changes the calculation by $143,000. In this Box Agent preview, @grok reconciles the claim against the policy and supporting records. It catches an $82,000 duplicate invoice and a missing $64,000 supplier credit. It also explains why the Business Income waiting period doesn’t apply to Extra Expense. The result is a cited claims review for the adjuster, with final coverage and payment decisions left to the insurer. Explore Box AI Studio to build custom agents for your own document-heavy workflows.
水下拍摄通常需要大量专业设备,搭建水下布景格外昂贵,而绿幕打光在水下还有额外的独特难题需要克服。 混合制作让水下拍摄变得容易得多。 Made with Luma.
发布还在继续!用交互式学习概览提升你的学习——现已面向所有用户开放!📓✨ 在 Reports 下,将来源摘要与 studio 产物无缝整合为一个交互式中心。非常适合学习或一站式深入了解一个新主题。
Googlebook is officially here! I’ve been so excited to share the details with the world. Today, many of us rely heavily on laptops to get work done, but we think there is an opportunity to rethink the category to address the needs of people today. So we brought the best of ChromeOS and Android to create a new platform for laptops. Our initial focus with Googlebook is to deliver an amazing laptop that feels awesome for Android phone users. Here’s what to know about Googlebook 🧵👇
从温暖的饱和感到更具实验性的风格,描述你想要的效果,就能在你的音轨中听到 ✨
NVIDIA Earth-2 面向天气敏感行业,将能源公司的风光资产测量、应急管理的雷达与本地传感器、卫星对地观测等更早、更本地化的观测数据用于天气决策。这些数据帮助各行业组织理解并管理物理风险。
Higgsfield AI 借助 GPT-6 Astra 在一天内上线新视频功能,让小型企业更轻松地制作视频广告,并更快将新创意工具推向市场。
小米 MiMo 在 GitHub 开源 mimoagent,一个仅约 100 行代码的 AI 智能体,可解决 GitHub issue 或在命令行中辅助用户。项目主打极简设计,无需庞大配置和大型 monorepo,并在 SWE-bench Verified 上取得超过 74% 的分数。仓库地址:https://github.com/XiaomiMiMo/mimoagent
OpenAI Academy 新增面向员工、开发者、管理者、教育者和学生的学习路径,用于培养和展示实用 AI 技能。
感谢 @sgl_project 的 day-0 支持!🙌 SGLang-Diffusion 现已支持 Qwen-Image-2.1:文生图、多图编辑,以及透明 RGBA 输出。快来试试!🎨
Day-0 support for @Alibaba_Qwen’s Qwen-Image 2.1 is here in SGLang-Diffusion! 🖥️ Native precision on a single RTX 4090 24GB with CPU offload - 1024×1024 generation in 18.7s and image editing in 21.7s with 22.7 GiB peak GPU memory during requests. - On an RTX PRO 6000 96GB: 8.0s generation and 9.6s editing. 🎨 Text-to-image, multi-image editing, and transparent RGBA output—all with one checkpoint. ⚡ Native inference with TP/SP, LoRA, and OpenAI-compatible APIs. 40 denoising steps, one image per request, warmed HTTP latency including PNG output. No quantization. Cookbook and GPU-specific commands below 👇