Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
OpenAI is previewing a new Ultrafast API service tier for GPT-5.6 Sol, powered by Cerebras hardware, achieving speeds of up to 750 output tokens per second.
Why it matters: Extremely high token generation speeds enable real-time voice and interactive applications that previously suffered from latency bottlenecks.