General Compute signed a multi-year agreement to integrate Cerebras' wafer-scale AI processors into its GPU cloud platform offerings.
General Compute, an AI cloud startup, has signed a multi-year agreement to deploy Cerebras' wafer-scale AI hardware. The company will offer access to Cerebras' technology via its cloud platform starting in Q1 2027.
Under the arrangement, General Compute will finance the systems and sell the result as inference. The company describes this as its "largest single hardware commitment" and the first major deployment since raising $400 million in debt financing in July 2026. The value of the agreement and scale of deployment have not been disclosed.
Sean Lie, Cerebras CTO and co-founder, said: "In AI, speed is productivity. An agent that takes hundreds of steps to finish a task is only as fast as its slowest step. Working with General Compute puts Cerebras speed in front of the developers building these agents, on a platform they already trust."
General Compute CEO Finn Puklowski added: "The chips that win inference are not going to come from one vendor, and most customers cannot put a wafer-scale system on their own balance sheet. That is the gap we exist to close. We buy the hardware, and our customers get Cerebras speed on a contract they can actually sign. Agentic coding is where that speed is worth the most right now, so that is where we are starting."
The deployment will incorporate Nvidia GPUs for prefill, which according to Puklowski "gives a major step up in reducing the cost of delivering inference—meaning more intelligence per dollar." General Compute's broader cloud platform also includes chips from AMD and SambaNova, with SambaNova GN50 chips handling decoding calculations and AMD MI300X graphics cards managing the remaining inference workload phases.
General Compute was founded by Puklowski and CTO Jason Goodison. The company raised $15 million in seed funding in May 2026, with a whitepaper from that time noting it has "options on 15 megawatts of air-cooled power, which is enough to support the Q4 2026 architecture and the growth beyond it, at colocation facilities that exist today."
This is Cerebras' second major cloud platform agreement this week. AI cloud startup Gimlet Cloud announced it would be deploying 100MW of Cerebras' wafer-scale compute for inference workloads, with the first data center expected to come online later this year.