OpenAI previews GPT-5.6 Sol Ultrafast mode at up to 14× standard speed
OpenAI is previewing Ultrafast, a limited-access API tier for GPT-5.6 Sol. Powered by Cerebras, it is advertised at up to 14× Standard processing speed and up to 750 output tokens per second.
OpenAI is previewing Ultrafast, a new service tier for GPT-5.6 Sol that launches first in the OpenAI API. Access is currently limited to a select group of customers, with the company saying availability will expand to more businesses as capacity grows.
The tier is powered by Cerebras and is advertised at up to 14 times the speed of Standard processing, with a peak of up to 750 output tokens per second. OpenAI positions it for latency-sensitive applications such as real-time voice, customer support, coding, commerce, financial research, and security response.
The 14× and 750-token figures are stated maximums, not guaranteed speeds for every request. The supplied announcement does not provide broad-release timing or complete pricing details.
Source evidence
Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the APIcommunity.openai.com · supporting# Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the API OpenAI is previewing Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, bringing frontier intelligence to products and workflows where every second matters. [Watch the GPT‑5.6 Sol Ultrafast demo]( GPT‑5.6 Sol on Ultrafast mode is launching first in the OpenAI API. It is currently available as a limited preview to a select group of customers, with access expanding as capacity grows. ### `Sign up for Ultrafast access updates` Learn more in the official announcement. Never was the :rocket: emoji more missing from the reactions list! :rocket: [...] ### Related topics | Topic | | Replies | Views | Activity | --- --- | GPT- 5.3-Codex-Spark Research Preview with 1000 Tokens per Second Codex announcement | 7 | 2269 | February 21, 2026 | | GPT-5.4 deep dive: pricing, context limits, and t
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the ...openai.com · supportingAugust 13, 2026 # Previewing Ultrafast mode: GPT‑5.6 Sol at up to 14X the speed A new speed class for frontier intelligence, turning speed into a competitive advantage. Today, we’re sharing an early look at Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing, launching first in the OpenAI API. Powered by Cerebras, Ultrafast generates up to 750 output tokens per second, bringing our most intelligent model to products and workflows where every second matters. [...] ## Powered by Cerebras Ultrafast marks the next step in our partnership with Cerebras to bring ultra-low-latency inference to OpenAI’s platform. Now, with GPT‑5.6 Sol on Ultrafast mode, Cerebras is supporting OpenAI’s most intelligent model, delivering up to 750 output tokens per second, enabling businesses to build more responsive products, make faster decisions, and bring powerful AI directly into their most demanding workflows. ## Availability GPT‑5.6 Sol on Ultrafast mode is
OpenAI previews Ultrafast GPT-5.6 Sol at up to 14x speededtechinnovationhub.com · supportingOpenAI has previewed a new Ultrafast service tier for GPT-5.6 Sol that it says can run the model at up to 14 times the speed of Standard processing, bringing lower-latency access to its frontier model first through the OpenAI API. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, according to OpenAI. The company is initially testing the service with a limited group of customers across coding, commerce, financial research, customer support and other interactive applications. Rather than using a smaller or more specialized model when response speed is critical, developers can use GPT-5.6 Sol at substantially higher inference speeds. [...] #### Cerebras powers the new service tier Ultrafast extends OpenAI's existing partnership with Cerebras, which is providing the infrastructure for the low-latency inference service. OpenAI says Cerebras is supporting GPT-5.6 Sol at speeds of up to 750 output tokens per second. The company describes the service as a new
OpenAI’s GPT-5.6 Sol runs up to 14× faster with Ultrafast mode - Help Net Securityhelpnetsecurity.com · supportingHelp Net Security newsletters: Daily and weekly news, cybersecurity jobs, open source projects, breaking news – subscribe here! Anamarija Pogorelec # OpenAI’s GPT-5.6 Sol runs up to 14× faster with Ultrafast mode OpenAI’s GPT-5.6 Sol on Ultrafast mode is available in limited preview to a select group of customers, launching first through the OpenAI API. The company says the service runs up to 14 times faster than Standard processing and generates up to 750 output tokens per second. Ultrafast is powered by Cerebras as part of the companies’ partnership on ultra-low-latency inference. GPT-5.6 Sol Ultrafast GPT-5.6 Sol Ultrafast ###### GPT-5.6 Sol Ultrafast and standard build a working 3D warehouse simulator from the same text prompt, side by side. (Source: OpenAI) [...] GPT-5.6 Sol Ultrafast and standard build a working 3D warehouse simulator from the same text prompt, side by side. (Source: OpenAI) During the preview, customers are testing Ultrafast across coding, commerce, finan
Ultrafast GPT-5.6 Sol Launched in OpenAI API | OpenAI posted on the topic | LinkedInlinkedin.com · supporting11,620,649 followers A first look at Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. Powered by Cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every second counts. Ultrafast is designed for businesses where faster frontier intelligence creates a measurable advantage, including real-time voice and customer support, commerce, coding and design, financial research, and security response. To view or add a comment, sign in [...] 11,620,649 followers A first look at Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. Powered by Cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every seco
OpenAI on X: "Previewing Ultrafast modex.com · supportingLog inSign up ## Post # OpenAI on X: "Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows." @OpenAI OpenAI @OpenAI Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. 00:00 5:01 PM · Aug 13, 20264.1MViews 799 @OpenAI OpenAI @OpenAI Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. 00:00 5:01 PM · Aug 13, 20264.1MViews 799 @OpenAI OpenAI @OpenAI [...] 799 @OpenAI OpenAI @OpenAI Aug 13 Powered by @cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every second co