跳到正文
原文
elvis· @omarsar0 · X·· 4 小时前AI 评分48
AI 导读

我认为更快的推理速度是编码智能体的下一个重大突破。 Volantis 正在利用光学技术为每颗芯片提供远超以往的内存和更高的内存带宽。他们的目标是在超过 10T 参数的模型上实现每用户每秒高达 10,000 tokens。这太疯狂了! 以那种速度,今天需要数小时的编码智能体可以在几分钟内完成。 绝对是我最近见过的最令人兴奋的融资之一。

正文

I think faster inference is one of the next big unlocks for coding agents.

Volantis is using optics to give each chip far more memory and much higher memory bandwidth. They're targeting up to 10,000 tokens per second per user on models over 10T parameters. That's crazy!

At that speed, a coding agent that takes hours today could finish in minutes.

Definitely one of the more exciting raises I have seen recently.

引用Tapa Ghosh@semiDL
Excited to announce Volantis's $88M Series A. We are solving Al's memory bottleneck by using optics, enabling chips with huge amounts of fast & cheap memory. By boosting both the memory bandwidth and capacity per chip by orders of magnitude, we enable ultra-fast inference (up to 10,000 tps/user) for large models (>10T) - with low $/tok to boot. Initially, this will enable insanely fast agents - think coding agents that finish in minutes or even seconds instead of hours. More excitingly, optics is a fundamentally scalable way to increase memory systems. Not 2X/year, but by orders of magnitude across new generations. This will enable a structurally new Al industry, including restarting scaling laws, holding entire repos in context windows & more. Our team has pioneered many core semiconductor technologies: the 1st CoWoS product, early HBM, the 1st silicon photonics CPO systems, the 1st high volume tunable VCSELs, the 1st processors to directly communicate using light & more. We’ve already sent data >10× farther than equally tiny electrical wires inside a chip package. Our next iteration is already taped out and targets world-record bandwidth density over relevant distances, read more: https://volantissemi.ai/news-insights/our-88m-series-a-demolishing-the-memory-wall-with-photonics-post
在 X 查看被引用的帖子

来源:elvis · x.com