Models
August 14, 2026

OpenAI previews Ultrafast mode for GPT-5.6 Sol at up to 14X speed

OpenAI has previewed Ultrafast, a new API service tier powered by Cerebras that runs GPT-5.6 Sol up to 14 times faster, reaching 750 output tokens per second.

OpenAI has introduced a limited preview of Ultrafast, a new API service tier that runs GPT-5.6 Sol up to 14 times faster than Standard processing. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second while retaining Sol’s frontier intelligence.

OpenAI sees applications across incident response, financial research, security, customer support, voice, commerce, and interactive research, where low latency can improve workflows. Early customers include Jane Street, Podium, Basis, and Rogo.

OpenAI is also testing Ultrafast internally for incident response and research. Access is currently limited to selected customers and will expand as capacity increases.

#
OpenAI

Read Our Content

See All Blogs
LLM Models

LLM testing of Grok 4.6: A cost-curve event, not a capability

Sarankumar S

August 14, 2026
Read more
LLM Models

Meta Muse Glimmer: Local agentic AI for enterprises

Vishesh Jain

August 13, 2026
Read more