OpenAI Ultrafast
OpenAI Ultrafast is a new API service tier that runs GPT-5.6 Sol at up to 14 times normal speed, delivering 750 output tokens per second. Powered by Cerebras wafer-scale hardware rather than conventional GPU clusters—a speed no GPU cloud has publicly matched. Designed for enterprise workloads where latency is critical: incident response, customer support, financial analysis, and e-commerce. Currently in limited preview with expanding availability.