๐ OpenAI's GPT-5.6 Sol Hits 750 Tokens Per Second in New Cerebras Ultrafast Mode
OpenAI opened a limited API preview of Ultrafast mode for GPT-5.6 Sol on August 13, powered by Cerebras hardware. It delivers up to 750 output tokens per second โ roughly 14x faster than standard โ with no quality degradation on benchmark tasks. No pricing or GA date yet; access is rolling out to select developers first. Could near-real-time inference fundamentally change how people build AI agents?