这是几个月前我还在 Google 时录制的,真的是一次非常有趣的对话!
How does this only have 21,000 views in 8 days? Chat with @JeffDean (then Google) and Bill Jia about large scale AI models. https://youtu.be/BVQSWeK2Nrw?si=MvMXdAEqjQE4HBOb
这是几个月前我还在 Google 时录制的,真的是一次非常有趣的对话!
How does this only have 21,000 views in 8 days? Chat with @JeffDean (then Google) and Bill Jia about large scale AI models. https://youtu.be/BVQSWeK2Nrw?si=MvMXdAEqjQE4HBOb
Google Research 发布 AI video co-director,一个构建在 Gemini 和 Veo 之上的多智能体编排框架,用于生成连贯的多镜头长视频,并原生继承 SynthID 水印等安全机制。
推荐理由:Google 把长视频生成拆成四个可组合的智能体框架,并给出三套自建基准与量化结果,便于对照现有方案。
Chrome 将 Gemini 的媒体理解能力在桌面端扩展到播客及 YouTube 以外的视频,播放后即可提取要点、定位信息或解释难点。Gemini 还能基于标签页内容生成互动测验,已在美印桌面端以英文推出,未来数月扩展至更多地区和移动端。此外 Chrome 新增跨设备发送标签页时保留滚动位置和未填完表单,并支持分屏视图、标签组、沉浸式阅读模式和朗读功能。
Google 在 Google Health 应用中推出 Health Guardian 功能,Pixel Watch 用户可设置血压趋势、胰岛素抵抗趋势和睡眠呼吸质量指标,本周开始启用,首批趋势报告将于 10 月初送达。
Google Photos 上线 5 项更新,其中 Photos in Gemini Spark 支持用一句提示词挑选、提亮并分享家庭合照,面向美国符合条件的 Google AI Pro 和 Ultra 订阅用户。
Google DeepMind 发布 Gemini 3.8 Live with Live Avatar,把近实时视频生成与语音对话模型结合,让对话 AI 具备动态视觉形象,支持精准唇形同步、自然表情和流畅轮次切换。
推荐理由:官方披露了实时视频与语音耦合的对话能力、异步工具调用和 97 种语言支持,可据此判断企业级数字人交互的落地边界。
谷歌 Project Suncatcher 的首颗轨道 AI 数据中心试验卫星将于 10 月 1 日发射,用于验证其 AI 卫星星座设想。这颗代号 MVP 的卫星约冰箱大小,搭载 4 块谷歌自研 TPU,太阳能板供电约 1 千瓦,卫星本体来自 Planet Labs。谷歌原计划 2027 年发射两颗定制卫星,为加速进度改为把自家芯片集成到 Planet Labs 已建好的卫星上。
Google 为 Demand Gen 推出九月更新,在 YouTube 广告中上线 Business Agent,观众可在视频广告旁与对话式 AI 互动,通过商品 feed 即时获取产品或品牌信息。
Google 推出 Gemini 3.8 Live with Live Avatar,将低延迟流式视频与实时对话能力原生结合,让企业智能体在对话中"能听、能看、能说",并具备精准唇形同步与自然表情。
Google Cloud 宣布 AlloyDB 推出面向 AI 智能体的 PostgreSQL(预览版),可在数秒内动态创建沙箱化数据库实例,实现与生产工作负载的完整隔离。该架构基于 Colossus 统一存储层,支持亚毫秒级 I/O、每秒超过 3 百万次查询和 terabit 级聚合扫描吞吐,可扩展至数千个无服务器实例,并在智能体任务结束后自动缩容至零,仅按活跃推理循环计费。
Google Cloud 发布 AlloyDB 智能体数据库架构并开放 AlloyDB PostgreSQL for agents 预览,提出隔离、亚毫秒延迟、弹性扩展三大原则。
我们与 @GoogleDeepMind、@emblebi 及研究伙伴合作,将 2,800 多种病毒的 AI 预测蛋白质复合物结构公开开放。 这为科学家提前应对潜在疫情提供了先机。
NVIDIA 联合 Google DeepMind、EMBL-EBI 等机构,通过 AlphaFold Database 开放了 2800 多种病毒的蛋白复合物预测 3D 结构,供全球科学家免费使用。
推荐理由:读者可了解开放病毒蛋白结构数据集如何降低结构预测门槛,以及配套开源流程的复用方式。
Remedy Entertainment 的 CONTROL Resonant 在发售当日加入 GeForce NOW,Ultimate 会员可以 GeForce RTX 5080 级云端性能串流,支持 DLSS 4、光线追踪、NVIDIA Reflex 和最高 5K HDR,无需 100GB 本地安装。
Google 宣布 Project Suncatcher 将通过 SpaceX Transporter-18 拼车任务发射首颗原型卫星,测试 Google TPU 能否在太空运行,并与 Planet 合作开发。
推荐理由:原文给出 Project Suncatcher 首次在轨测试安排,以及振动、辐射、散热和星间激光互联的实测进展。
Google 宣布一项承诺,将帮助 25,000 名退伍军人和军属在技能行业建立职业,Google.org 的资金与支持将提供给 Hiring Our Heroes、Student Veterans of America(SVA)和 Home Builders Institute(HBI),由它们提供学徒前实操培训、职业辅导和直接就业安置。
Google Arts & Culture Lab 推出 Gemini 辅助的 The Talking Museum,可近距离查看近 200 件 3D 文物并提问。
YouTube 在 Made on YouTube 活动上公布自定义信息流 Custom Feeds,并预告面向观众和创作者的更多 AI 功能。用户输入描述即可生成信息流,通过修改描述、给推荐视频打分来调优,并可保存多个,该功能即将面向美国 web、移动端和 TV 用户推出。今年晚些时候 YouTube 还将支持评论发 GIF,并为私信加入群聊,群聊初期仅限美国、英国、新加坡、巴西及部分欧洲国家。
Google 宣布所有 Google 或 Google Workspace 账号均可通过 Google Vids 使用最新的 Gemini Omni 1.1 Flash 模型免费生成高质量视频,入口为 vids.new 并选择 "Create AI videos"。
推荐理由:原文来自产品经理宣布,交代了免费开放入口、具体模型和新控制功能,读者可据此判断是否改变自己的视频制作流程。
Android Enterprise 公布 6 项更新,涵盖 Gemini 多步骤跨应用自动化、工作资料隔离与 IT 集中管控,以及 XREAL Aura 有线眼镜等 XR 设备统一管理。
Google Beam 正式扩展至美国、加拿大、英国、法国、德国和日本六国,由 18 家渠道合作伙伴支持部署,旗舰硬件为 HP Dimension with Google Beam。
Google 发布两个 Gemini 文本转语音模型 gemini-3.8-flash-tts 和 gemini-3.8-flash-lite-tts,提供超过 2,000 个声音,并支持用 30 秒音频样本创建自定义语音。
Now Gemini can connect with 13 new apps like @adobe, @squarespace, @onepeloton, and more. Instead of switching between tabs, you can now grow your business, design assets, and plan your workouts all in Gemini. 🧵
Google DeepMind 公布 Private AI Compute 架构更新,将持久化的服务端记忆引入该平台,使 AI 助手能跨设备保留上下文。数据被密封在加密存储中,解锁密钥仅保存在用户个人设备上,模型访问时通过端到端加密通道连接云端安全隔离区,在隔离内存中临时解密、保存新上下文后立即重新加密。
推荐理由:Google 公开了 Private AI Compute 的持久化服务端记忆方案,读者可了解云端记忆如何在设备持钥前提下实现。
Google DeepMind 推出 Gemini 3.8 Flash TTS 和 Gemini 3.8 Flash-Lite TTS 两款文本转语音模型,前者面向角色设计与深度创作控制,后者面向高并发、低成本的配音与语音智能体场景。
推荐理由:两款 TTS 模型给出语音克隆、逐行导演与多语言覆盖的具体能力,可据此判断语音生成工作流的变化。
Latent Space 发布对 John Platt 的访谈,介绍其团队在 Google 的 Empirical Research Assistance(ERA)项目。
More than 25 million people are using Google Flow every month to dream up new ideas, create stories, and build cool things. Thank you. We’re continuing the 50 additional daily credits for all users. Dive in and keep creating.
发布还在继续!用交互式学习概览提升你的学习——现已面向所有用户开放!📓✨ 在 Reports 下,将来源摘要与 studio 产物无缝整合为一个交互式中心。非常适合学习或一站式深入了解一个新主题。
I had the great honor and pleasure of sitting down with @JeffDean for his first public talk since leaving Google, where he spent an extraordinary 27 years. Few people have shaped modern computing and AI as profoundly - from MapReduce and Bigtable to TensorFlow, Mixture-of-Experts, TPUs, and Gemini. Our conversation covered some of the biggest questions shaping the future of AI: • How do you recognize a foundational idea before everyone else does? • How do you choose a research problem worth spending 5 years on? • What can coding teach us about building better reasoning models? • What might recursive self-improvement (RSI) actually look like? • What happens when the scientific discovery loop itself becomes increasingly automated? (and how is Jeff’s new startup going to contribute in this space?) • As AI becomes increasingly autonomous, how do we keep it safe and secure? • What should the next generation of researchers be working on? Here are some key insights and highlights for anyone building the future of AI. 🧵1/8
Googlebook is officially here! I’ve been so excited to share the details with the world. Today, many of us rely heavily on laptops to get work done, but we think there is an opportunity to rethink the category to address the needs of people today. So we brought the best of ChromeOS and Android to create a new platform for laptops. Our initial focus with Googlebook is to deliver an amazing laptop that feels awesome for Android phone users. Here’s what to know about Googlebook 🧵👇