OpenAI has introduced a limited preview of Ultrafast, a new API service tier that runs GPT-5.6 Sol up to 14 times faster than Standard processing. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second while retaining Sol’s frontier intelligence.
OpenAI sees applications across incident response, financial research, security, customer support, voice, commerce, and interactive research, where low latency can improve workflows. Early customers include Jane Street, Podium, Basis, and Rogo.
OpenAI is also testing Ultrafast internally for incident response and research. Access is currently limited to selected customers and will expand as capacity increases.




