
OpenAI introduces Ultrafast mode for GPT-5.6 Sol
OpenAI has introduced a new preview mode called Ultrafast, designed to accelerate the processing speed of its GPT-5.6 Sol model. Powered by hardware from Cerebras, the system aims to bring high-throughput generation to enterprise workflows without sacrificing model scale.
Published by Jin · 2 min read · 14 AUG 2026
OpenAI has rolled out a new operating mode called Ultrafast, which is designed to significantly accelerate the pace at which its flagship model, GPT-5.6 Sol, completes tasks. According to the laboratory, the new mode operates at fourteen times the speed of standard processing.
Performance and Metrics
The system is capable of delivering up to 750 output tokens per second. Historically, achieving real-time generation speeds has required deploying smaller or more specialized models. Ultrafast represents an alternative path, maintaining the capabilities of a larger architecture while increasing the volume of output generated per second.
Competitors in the artificial intelligence sector have pursued similar strategies. Anthropic previously introduced accelerated options for its Claude models, though the throughput offered by OpenAI in this release represents a notable shift in generation velocity.
Enterprise Applications
OpenAI suggests that the high-velocity version of GPT-5.6 Sol is suitable for time-sensitive corporate workflows. Notable applications include:
- Incident response operations
- Customer service and support systems
- Financial market analysis
- E-commerce logistics and management
Availability and Infrastructure
Ultrafast is currently powered by a partnership with chipmaker Cerebras. The feature is entering a preview phase and is initially available to a restricted group of enterprise customers. OpenAI stated that broader access will be granted incrementally as processing capacity expands over time.
Source — Original announcement ↗
Worth a read?
Comments · 0