OpenAI has officially launched the new GPT-5.6Sol UltraFast super-fast mode preview for enterprise users. With this mode, the model's performance has seen a revolutionary improvement, reaching up to 14 times the speed of the standard mode, and can generate up to 750 tokens per second.

According to the information, the core technology of this ultra-fast mode is supported by Cerebras. However, in order to ensure the reasonable use of resources, this mode is currently available through an application review system. Enterprise users who need it must submit specific usage scenarios to OpenAI, and whether the application is approved will be comprehensively evaluated by the official based on actual circumstances.

image.png

In terms of application scenarios, OpenAI stated that the ultra-fast mode is not aimed at all regular workloads, but is specifically designed for business scenarios that require high real-time performance. For example: real-time voice interaction, intelligent customer support, business applications, development of intelligent agents (Agents), financial in-depth research, and cutting-edge security research.

As the preview phase progresses, the computing power provided by Cerebras will continue to be optimized. In the future, OpenAI plans to gradually expand the coverage based on feedback from enterprise workloads, allowing more enterprises to smoothly access this ultra-fast AI capability through the API interface.