产品精选76°

Volantis获8800万美元A轮融资

精选理由

Volantis用光学技术让大模型推理速度提升百倍,编程助手从几小时缩短到几分钟。

Volantis公司利用光学技术解决AI内存瓶颈问题,使芯片内存容量和带宽提升数量级。该公司目标实现每用户每秒10,000个token的处理速度,支持超过10T参数的大模型。这项技术可使原本需要数小时的编程助手任务缩短至几分钟完成。

原文 · elvis

I think faster inference is one of the next big unlocks for coding agents. Volantis is using optics to give each chip far more memory and much higher memory bandwidth. They're targeting up to 10,000 tokens per second per user on models over 10T parameters. That's crazy! At that speed, a coding agent that takes hours today could finish in minutes. Definitely one of the more exciting raises I have seen recently. Tapa Ghosh @semiDL Excited to announce Volantis's $88M Series A. We are solving Al's memory bottleneck by using optics, enabling chips with huge amounts of fast & cheap memory. By boosting both the memory bandwidth and capacity per chip by orders of magnitude, we enable ultra-fast inference (up to 10,000 tps/user) for large models (>10T) - with low $/tok to boot. Initially, this will enable insanely fast agents - think coding agents that finish in minutes or even seconds instead of hours. More excitingly, optics is a fundamentally scalable way to increase memory systems. Not 2X/year, but by orders of magnitude across new generations. This will enable a structurally new Al industry, including restarting scaling laws, holding entire repos in context windows & more. Our team has pioneered many core semiconductor technologies: the 1st CoWoS product, early HBM, the 1st silicon photonics CPO systems, the 1st high volume tunable VCSELs, the 1st processors to directly communicate using light & more. We’ve already sent data >10× farther than equally tiny electrical wires inside a chip package. Our next iteration is already taped out and targets world-record bandwidth density over relevant distances, read more: volantissemi.ai/news-insights/… 🔗 View Quoted Tweet 💬 10 🔄 1 ❤️ 13 👀 2018 📊 8 ⚡