OpenAI is unveiling a groundbreaking update to its latest advanced model, GPT-5.6, allowing it to operate at significantly increased speeds. The newly introduced Ultrafast service tier is claimed to execute GPT-5.6 Sol up to 14 times quicker than standard processing speeds.
This Ultrafast mode is being rolled out initially via the OpenAI API, utilizing Cerebras technology. It can produce as many as 750 output tokens every second, offering a substantial performance boost for applications where speed is crucial, in addition to model sophistication.
GPT-5.6 Sol was launched in June, alongside two additional models: Terra, which balances performance, and Luna, which prioritizes speed. The entire suite became accessible to users in July, integrated across platforms like ChatGPT, Codex, and the API.
OpenAI anticipates that the Ultrafast feature will enhance real-time or near-production applications, including voice recognition, customer support, e-commerce, software development, financial analysis, and security operations.
The company notes that its own developers have utilized this high-speed capability to scrutinize logs and traces during critical events, as well as to streamline research processes that previously took an entire night, now allowing for multiple iterations within a regular workday.
Currently, access to this service tier is restricted to a limited number of customers while OpenAI monitors the impact of the increased speed on actual products and its ability to handle larger demands.
Businesses interested in utilizing the Ultrafast features can join a waitlist by providing information about their workload, speed requirements, anticipated usage, and other pertinent details.




