elvis· @omarsar0 · X·· 2 小时前AI 评分52
AI 导读
Volantis 宣布完成 8800 万美元 A 轮融资,用光学技术为每颗芯片提供更大容量和更高带宽的内存,目标是在超过 10T 参数的大模型上实现每位用户最高 10,000 tokens 每秒的推理速度。作者 Elvis Saravia 认为更快推理是编码智能体的下一个重大解锁,届时耗时数小时的编码智能体任务可能几分钟内完成。团队有 CoWoS、HBM、硅光 CPO 等半导体技术背景,下一代产品已流片。
正文
I think faster inference is one of the next big unlocks for coding agents.
Volantis is using optics to give each chip far more memory and much higher memory bandwidth. They're targeting up to 10,000 tokens per second per user on models over 10T parameters. That's crazy!
At that speed, a coding agent that takes hours today could finish in minutes.
Definitely one of the more exciting raises I have seen recently.
Excited to announce Volantis's $88M Series A. We are solving Al's memory bottleneck by using optics, enabling chips with huge amounts of fast & cheap memory. By boosting both the memory bandwidth and capacity per chip by orders of magnitude, we enable ultra-fast inference (up to 10,000 tps/user) for large models (>10T) - with low $/tok to boot. Initially, this will enable insanely fast agents - think coding agents that finish in minutes or even seconds instead of hours. More excitingly, optics is a fundamentally scalable way to increase memory systems. Not 2X/year, but by orders of magnitude across new generations. This will enable a structurally new Al industry, including restarting scaling laws, holding entire repos in context windows & more. Our team has pioneered many core semiconductor technologies: the 1st CoWoS product, early HBM, the 1st silicon photonics CPO systems, the 1st high volume tunable VCSELs, the 1st processors to directly communicate using light & more. We’ve already sent data >10× farther than equally tiny electrical wires inside a chip package. Our next iteration is already taped out and targets world-record bandwidth density over relevant distances, read more: https://volantissemi.ai/news-insights/our-88m-series-a-demolishing-the-memory-wall-with-photonics-post在 X 查看被引用的帖子
来源:elvis · x.com