OpenAI has announced a major performance breakthrough with the launch of its new 'Ultrafast' mode for the GPT-5.6 Sol model. Designed to accelerate processing throughput up to 14 times the standard speed, the new mode achieves up to 750 output tokens per second, redefining low-latency AI interactions for enterprise systems.
Developed in collaboration with specialized hardware manufacturer Cerebras, Ultrafast eliminates the traditional trade-off between model intelligence and inference speed. The system is particularly targeted at real-time incident response, high-frequency financial market analytics, customer operations, and automated e-commerce workflows where microsecond delays matter.
Currently accessible as an early preview to a select group of enterprise clients, access to the Ultrafast mode will expand across developer APIs and ChatGPT tiers as compute infrastructure scales up globally.
