Etched做了专为AI推理优化的芯片,比通用硬件在吞吐、延迟、能效上都做到了目前最好。今年夏天就发货,关注成本的人可以看看。
AI推理成本高昂的原因之一是多数工作负载运行在预LLM时代的通用硬件上。Etched从隐身模式亮相,其系统专为现代推理从零设计。公司已获得超过10亿美元客户合同和8亿美元融资。早期测试显示,在推理工作负载上吞吐量、延迟和能效均达到SOTA。首批机架将于今年夏季发货。
AI is expensive to run partly because most workloads today run on generic hardware designed pre-LLMs...
AI is expensive to run partly because most workloads today run on generic hardware designed pre-LLMs. Etched is the first system designed from the ground up for modern inference. Etched @Etched We're coming out of stealth. We've built our first racks after a successful A0 tapeout, $1B+ in customer contracts, and $800m raised. Early customer tests show us achieving SOTA throughput, latency, and power efficiency on inference workloads. Our first racks ship this summer. 🔗 View Quoted Tweet 💬 22 🔄 23 ❤️ 421 👀 42740 📊 60 ⚡