Frontier Inference Clusters
We co-design chips, racks, software, and manufacturing methods so frontier models can run with best-in-class throughput, latency, cost, and power efficiency for both prefill and decode workloads.
Earlier this year our A0 silicon came back from TSMC N4P, and today we are busy validating our first rack-scale product with customers to fulfill $1B in demand.
We’re a team of 400+ engineers from NVIDIA, Google TPUs, Broadcom, SK Hynix, TSMC, and more. We’ve raised $800M across four unannounced financings, including a strategic investment from VentureTech Alliance. We’re excited to deepen our partnership with the world’s leading semiconductor manufacturer.