01
OpenAI runs its flagship on Cerebras at 14× GPU speed
breakthroughDeveloperComputeFinance
Friday, August 14, 2026
Confidence
Medium · — two primary announcements + Tier-2 corroboration; benchmarks are vendor-run and preview pricing undisclosed
Evidence
OpenAI + Cerebras primary announcements + independent reporting + vendor benchmarks
OpenAI's flagship reasoning model just detached from the GPU stack for a named subset of customers — and the numbers are the story.
- Ultrafast runs GPT-5.6 Sol 14× faster than Standard, up to 750 output tokens per second, on Cerebras.
- Sol on Ultrafast finished Humanity's Last Exam in 11h 11m; Claude Fable 5 needed 78h 27m .
- Launched Aug 13 as a limited preview in the OpenAI API ; pricing not disclosed.
Sources