OpenAI Previews GPT-5.6 Sol Ultrafast Mode at Up to 14x Speed
OpenAI is previewing an API service tier for GPT-5.6 Sol that is powered by Cerebras and claims output speeds of up to 750 tokens per second.
OpenAI has announced an early preview of Ultrafast, a new service tier for GPT-5.6 Sol. The company says it can run up to 14 times faster than Standard processing and is launching first through the OpenAI API.
Cerebras provides the infrastructure for the low-latency offering, with OpenAI citing a peak generation rate of up to 750 output tokens per second. The tier is intended for products and workflows where response time can affect decisions, user engagement, or operational efficiency.
GPT-5.6 Sol on Ultrafast is currently available only to a select group of customers. OpenAI says access will expand as capacity grows. The speed figures are vendor-stated maximums, not independent performance benchmarks.
Source evidence
Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the APIcommunity.openai.com · supporting# Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the API OpenAI is previewing Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, bringing frontier intelligence to products and workflows where every second matters. [Watch the GPT‑5.6 Sol Ultrafast demo]( GPT‑5.6 Sol on Ultrafast mode is launching first in the OpenAI API. It is currently available as a limited preview to a select group of customers, with access expanding as capacity grows. ### `Sign up for Ultrafast access updates` Learn more in the official announcement. Never was the :rocket: emoji more missing from the reactions list! :rocket: [...] ### Related topics | Topic | | Replies | Views | Activity | --- --- | GPT- 5.3-Codex-Spark Research Preview with 1000 Tokens per Second Codex announcement | 7 | 2269 | February 21, 2026 | | GPT-5.4 deep dive: pricing, context limits, and t
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the ...openai.com · supportingAugust 13, 2026 # Previewing Ultrafast mode: GPT‑5.6 Sol at up to 14X the speed A new speed class for frontier intelligence, turning speed into a competitive advantage. Today, we’re sharing an early look at Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing, launching first in the OpenAI API. Powered by Cerebras, Ultrafast generates up to 750 output tokens per second, bringing our most intelligent model to products and workflows where every second matters. [...] ## Powered by Cerebras Ultrafast marks the next step in our partnership with Cerebras to bring ultra-low-latency inference to OpenAI’s platform. Now, with GPT‑5.6 Sol on Ultrafast mode, Cerebras is supporting OpenAI’s most intelligent model, delivering up to 750 output tokens per second, enabling businesses to build more responsive products, make faster decisions, and bring powerful AI directly into their most demanding workflows. ## Availability GPT‑5.6 Sol on Ultrafast mode is
GPT-5.6 Sol goes 14x faster as OpenAI launches Ultrafast mode powered by Cerebrasthe-decoder.com · supportingThe inference acceleration comes from Cerebras, which signed a ten-billion-dollar partnership with OpenAI earlier this year. The service will initially be available only through the OpenAI API for GPT-5.6 Sol and limited to select customers. OpenAI plans to expand access gradually as capacity grows. Companies that want in can sign up for updates through a form. ## Faster output opens up new use cases [...] Ad OpenAI pitches several other scenarios. In finance, the model could evaluate market signals and flag suspicious transactions while conditions are still shifting. In customer support, complex inquiries could be resolved in real time, even when finding the answer requires multiple steps or systems. In e-commerce, it could answer product questions, check inventory levels, and personalize recommendations before a hesitant buyer abandons their cart. Ad Video: Generating a 3D warehouse simulator, shown on the left in Ultrafast mode [...] Skip to content Sign In Register Subscrib
GPT-5.6 Sol Ultrafast: OpenAI Previews 14x Inferencedigitalapplied.com · supporting## The questions we getevery week. Ultrafast is a new OpenAI service tier, previewed on August 13, 2026, that runs GPT-5.6 Sol — the most capable model in the GPT-5.6 family — on Cerebras hardware at what OpenAI describes as up to 14x the speed of Standard processing, peaking at 750 output tokens per second. It is not a new model: Sol itself was previewed in June 2026 and became broadly available in July. Ultrafast launches first in the OpenAI API as a limited preview for a select group of customers, with access managed through a waitlist that OpenAI says will expand as capacity grows. Related dispatches ## Continue exploringfrontier releases. AI Development #### GPT-5.6 Goes Public: GA Pricing, Ultra Mode and Access [...] vendor-stated ceiling ▼up to Claimed multiple 14x vs Standard processing ▼up to Ultrafast pricing disclosed at preview GA timeline waitlist-gated access GPT-5.6 Sol Ultrafast is OpenAI’s newest service tier, previewed on August 13, 2026: the company’s
OpenAI previews Ultrafast GPT-5.6 Sol at up to 14x speed | ETIH EdTech News — EdTech Innovation Hubedtechinnovationhub.com · supportingOpenAI has previewed a new Ultrafast service tier for GPT-5.6 Sol that it says can run the model at up to 14 times the speed of Standard processing, bringing lower-latency access to its frontier model first through the OpenAI API. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, according to OpenAI. The company is initially testing the service with a limited group of customers across coding, commerce, financial research, customer support and other interactive applications. Rather than using a smaller or more specialized model when response speed is critical, developers can use GPT-5.6 Sol at substantially higher inference speeds. [...] #### Cerebras powers the new service tier Ultrafast extends OpenAI's existing partnership with Cerebras, which is providing the infrastructure for the low-latency inference service. OpenAI says Cerebras is supporting GPT-5.6 Sol at speeds of up to 750 output tokens per second. The company describes the service as a new
OpenAI on X: "Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows." / Xx.com · supporting799 @OpenAI OpenAI @OpenAI Aug 13 Powered by @cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every second counts. Ultrafast is designed for businesses where faster frontier intelligence creates a measurable advantage, including “Previewing Ultrafast mode” in white text over lime-green and sky-blue flowing gradients. Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speedFrom openai.com 36 @OpenAI OpenAI @OpenAI Aug 13 [...] Log inSign up ## Post # OpenAI on X: "Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows." @OpenAI OpenAI @OpenAI Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. 00:00 5: