OpenAI Makes GPT-5.6 Sol 14X Faster With New Ultrafast Mode

New Ultrafast mode pushes GPT-5.6 Sol to 14x standard processing speed (Image: Shutterstock)
New Ultrafast mode pushes GPT-5.6 Sol to 14x standard processing speed (Image: Shutterstock)

OpenAI introduced Ultrafast, a preview service tier that runs GPT-5.6 Sol up to 14 times faster than standard processing and generates up to 750 output tokens per second.

Key Points:

  • Ultrafast runs GPT-5.6 Sol up to 14 times faster than standard processing.
  • Cerebras powers the mode, which can deliver up to 750 output tokens per second.
  • The service is in limited preview, with broader access planned as capacity grows.

OpenAI Ultrafast Preview

OpenAI announced Ultrafast on Aug. 13, launching the speed-focused tier first through its API for a limited group of customers. The company said access will expand as capacity grows.

The service is powered by Cerebras and is designed to preserve GPT-5.6 Sol’s capabilities while cutting the time required to produce responses. Output tokens are pieces of generated text. OpenAI said real-time performance had typically required users to choose a smaller or more specialized model, while Ultrafast targets “more useful work per second.”

OpenAI listed incident response, financial research, customer support, commerce and live research among the workflows it expects to benefit most. Engineers can use the mode to review logs, traces and recent code changes while an outage is unfolding, while support systems can handle multi-step requests during live conversations.

Also Read: Anthropic Pursues A $6B Decart Deal While Preparing To Go Public

GPT-5.6 Sol Impact

Early customers said the speed changes how the model can fit into interactive work, with John Crepezzi of Jane Street calling the increase “impressive” and saying it makes focused work alongside models more practical.

Courtland Lykins, product lead for voice AI at Podium, said Ultrafast has been valuable for the company’s voice stack because “the speed completely changes the call experience” on complex work. OpenAI also said financial researchers can analyze changing market signals and transactions without waiting for slower batch-style processing. That use case depends on rapidly changing inputs.

The preview remains constrained by infrastructure rather than a broad public rollout. OpenAI said it is using feedback from the initial companies to determine where the speed produces the most value before expanding availability.

The Cerebras connection predates Ultrafast. OpenAI announced a partnership with the chipmaker in January to add 750 megawatts of low-latency AI compute to its platform. In February, GPT-5.3-Codex-Spark became the first model tied to that partnership, using Cerebras hardware to exceed 1,000 tokens per second in real-time coding.

Read Next: Virtuals Protocol Trading Volume Explodes 180%, Bulls Eye $0.68

Alexey Bondarev profile photo

Alexey Bondarev

Alexey Bondarev is the Head of Content at Yellow.com, having reported on crypto for the last 10 years. He specializes in in-depth Research and Learn pieces, with a focus on analytical reporting, industry context, and the bigger forces shaping crypto, from the AI era and security technologies to fintech innovation. He believes that everything digital will imminently overcome everything analogue and is working hard to make that come true.

Disclaimer and Risk Warning: The information provided in this article is for educational and informational purposes only and is based on the author's opinion. It does not constitute financial, investment, legal, or tax advice. Cryptocurrency assets are highly volatile and subject to high risk, including the risk of losing all or a substantial amount of your investment. Trading or holding crypto assets may not be suitable for all investors. The views expressed in this article are solely those of the author(s) and do not represent the official policy or position of Yellow, its founders, or its executives. Always conduct your own thorough research (D.Y.O.R.) and consult a licensed financial professional before making any investment decision.
Latest News
Show All News
OpenAI Makes GPT-5.6 Sol 14X Faster With New Ultrafast Mode | Yellow