Cerebras

Illustration for the OpenAI Ultrafast speed tier story
modelschips

OpenAI previews Ultrafast, running GPT-5.6 Sol at up to 750 tokens per second

On August 13, 2026 OpenAI previewed Ultrafast, a new service tier in its API powered by Cerebras hardware that runs GPT-5.6 Sol at up to 750 output tokens per second, up to 14 times faster than Standard processing. The tier launches first in the API, with this week's release described as an early look. It is the same model at radically lower latency, aimed at agents and latency-sensitive products.