Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
Preview Ultrafast, a new OpenAI API service tier that runs GPT-5.6 Sol up to 14× faster. Powered by Cerebras, it delivers up to 750 output tokens per second.
What happened
OpenAI announced the launch of the Ultrafast mode for its API, enabling significantly accelerated performance with the GPT-5.6 Sol model.
Why it matters
The rapid processing speeds can enhance developer workflows, allowing for quicker iterations and deployments of AI applications, ultimately improving responsiveness in real-time environments.