OpenAI has announced a preview of a new API offering called Ultrafast. According to the company, this setup runs the GPT-5.6 Sol model at significantly increased rates, reaching speeds up to 14 times faster than standard execution. The update targets performance improvements for users accessing the model through the API.
The underlying infrastructure for the service tier is powered by Cerebras hardware. Through this integration, the system achieves throughput figures of up to 750 output tokens per second. OpenAI disclosed the preview details to outline the technical capabilities of the new deployment tier.

