chiprook
← AI
AIOctober 3, 2026, 07:30

OpenAI and Cerebras confirm 750MW AI inference deployment by 2028

OpenAI and Cerebras confirmed a multi-year partnership to deploy 750MW of Cerebras wafer-scale compute for ultra-low-latency inference on OpenAI's platform. A Master Relationship Agreement effective December 24, 2025 splits capacity into three 250MW segments through the end of 2028.

OpenAI and Cerebras confirm 750MW AI inference deployment by 2028
#OpenAI#Cerebras
Read next
AI

Cerebras claims 5x inference throughput gain via disaggregation

AI

General Compute signs multi-year deal to deploy Cerebras wafer-scale hardware

AI

OpenAI: 80% of enterprise AI problems are deployment, not models

AI

OpenAI previews Ultrafast tier for GPT-5.6 Sol with up to 14x speed