OpenAI appears to be addressing user demands for a more responsive AI experience with the introduction of a new feature called Ultrafast. This mode promises to significantly enhance the efficiency of its latest model, GPT-5.6 Sol, enabling it to operate at a staggering speed—14 times faster than traditional processing.
According to OpenAI, Ultrafast can generate up to 750 output tokens per second. These tokens are the unique segments of text produced by a language model during interactions with users. The company noted in a recent blog entry that achieving real-time processing speed has typically required users to opt for smaller or specialized models. With Ultrafast, OpenAI aims to deliver more productive outcomes in a shorter amount of time.
Other AI developers, including Anthropic, have also introduced accelerated functionalities for their models. While Anthropic’s Claude offers a fast mode, it does not match the high speeds that OpenAI claims with Ultrafast.
OpenAI envisions that the enhanced capabilities of GPT-5.6 Sol will be advantageous for a variety of corporate applications, particularly in areas like incident response, customer service, financial market analysis, and e-commerce, among others.
Currently, Ultrafast is being rolled out in a preview phase, facilitated by OpenAI’s collaboration with chip manufacturer Cerebras. This preview is initially limited to a small selection of users, but OpenAI has indicated that it plans to broaden access as more resources become available.




