OpenAI launches Ultrafast Mode, increasing GPT-5.6 Sol speed by 14x

icon币界网
Share
AI summary iconSummary
OpenAI has launched a new token listing mode called Ultrafast, which increases GPT-5.6 Sol processing speed to 14 times the standard rate, with a peak output of 750 tokens per second. The mode is currently in preview for select customers and is designed to support real-time AI requirements in enterprise workflows. OpenAI states that the mode delivers faster responses without relying on smaller models. The update is supported by a partnership with Cerebras. While interest rate news remains a key factor for crypto markets, this development could shift focus toward AI-driven infrastructure.
CoinDesk reports:

OpenAI has released a new model called Ultrafast, designed for faster model response times. The company states that this mode enables GPT 5.6 Sol to operate at 14 times the standard processing speed, with output speeds of up to 750 tokens per second, and it is currently available in preview to a limited number of customers.

This update targets enterprises' demand for real-time AI processing. OpenAI states that previously, achieving responses closer to real-time typically required switching to smaller models or opting for systems optimized for specific tasks. Ultrafast aims to increase the amount of effective work completed per unit of time without solely relying on smaller models.

Improved output speed

According to OpenAI, Ultrafast's key selling point is significantly reducing generation wait times.

  • Maximum output speed of up to 750 tokens per second
  • Processing speed is approximately 14 times faster than standard mode.
  • The current version is still in preview.

TechCrunch noted that competitors like Anthropic had previously launched acceleration modes. For example, Claude’s product also offers a fast mode, but OpenAI’s announced speed metrics are higher.

For enterprise workflows

OpenAI targets Ultrafast for a variety of enterprise use cases, with a focus on incident response, customer service and support, financial market analysis, and e-commerce. These scenarios typically have higher sensitivity to latency, where model response times directly impact the efficiency of human collaboration and the performance of automated processes.

From a product positioning perspective, Ultrafast is not a standalone new model, but rather a high-speed execution method built around GPT 5.6 Sol. For enterprise customers, this means maintaining strong model capabilities while integrating AI more directly into real-time business workflows.

Powered by Cerebras

OpenAI states that Ultrafast is powered by its collaboration with chip company Cerebras. Initial preview access is currently limited to a small group of customers and will gradually expand as computing capacity increases.

This also shows that the competition among large models is extending from parameter and capability comparisons to response speed, deployment efficiency, and underlying computational power coordination. For AI products targeting enterprise markets, speed is becoming a key metric alongside model quality.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.