Technologies
Back
Artificial Intelligence & Machine Learning

OpenAI and Cerebras Confirm 750MW AI Inference Deployment Through 2028

Dev.to
Advertisement468 × 90
OpenAI and Cerebras Confirm 750MW AI Inference Deployment Through 2028

OpenAI and Cerebras have announced a multi-year partnership to deploy 750MW of wafer-scale AI compute capacity to support ultra-low-latency inference. The rollout will occur in three 250MW phases, with completion scheduled by the end of 2028. This infrastructure is specifically designed to enhance real-time AI interactions, such as conversational assistants and coding tools, rather than model training. While an SEC filing notes early operational use for an OpenAI Codex Spark model, the companies have not yet disclosed specific product integration plans, pricing, or performance metrics. The agreement represents a significant long-term commitment to scaling AI responsiveness, though the immediate impact on individual customer experiences remains to be seen. Businesses are advised to evaluate their specific technical requirements rather than assuming this infrastructure expansion will automatically resolve all latency or cost challenges.

This is a summary. Read the full article at the original source:

Dev.to
Advertisement468 × 90
Share
Artificial Intelligence & Machine Learning

Related stories

Advertisement970 × 250