The agreement includes renewal options and ranks among the company’s largest customer commitments announced to date. The customer’s platform helps companies deploy large language models, vision models, speech models and other AI applications with low latency and high reliability and runs models for customers specializing in LLMs, image and video generation. The dedicated Blackwell capacity will give the platform high-performance compute to serve production inference workloads as demand from its customers scales.
Capacity under the agreement will be served from QumulusAI’s U.S. data center footprint and is expected to be ready for customer use in the third quarter of 2026. The company’s demand-led deployment model places capacity into available pockets of power across a distributed network of colocation and owned facilities, enabling it to bring GPU capacity online in months, not years. The agreement adds more than $71 million in contracted, multiyear commitments to QumulusAI’s book of business.
Login to comment