Cerebras Runs OpenAI’s GPT-5.6 Sol at 750 Tokens Per Second in New Ultrafast Tier

Cerebras is now running OpenAI’s flagship model at a speed no GPU cloud has publicly matched. On August 13, 2026, the wafer-scale chipmaker announced it powers GPT-5.6 Sol on a new OpenAI service tier called Ultrafast, delivering up to 750 output tokens per second and, by OpenAI’s account, running the model up to 14× faster than Standard processing. Ultrafast launches first in the OpenAI API as a limited preview for a select group of customers, with access expanding as capacity grows. The claim…

Leave a Reply

Your email address will not be published. Required fields are marked *

Back To Top