OpenAI 预览 GPT-5.6 Sol Ultrafast 模式:速度最高提升至标准处理的 14 倍
OpenAI 正在 API 中限量预览 GPT-5.6 Sol 的 Ultrafast 模式。该模式由 Cerebras 提供推理基础设施支持,官方称最高可生成每秒 750 个输出 token,并将逐步扩大客户范围。
OpenAI 正在推出 GPT-5.6 Sol 的 Ultrafast 模式预览版。这一新服务层首先面向 OpenAI API 的少量客户开放,官方称其处理速度最高可达到标准处理的 14 倍;目前访问权限仍将随着可用容量增加而逐步扩大。
Ultrafast 由 Cerebras 提供支持,官方公布的峰值为每秒最多生成 750 个输出 token。该模式面向实时语音、客户支持、编程、商务、金融研究和安全响应等对延迟敏感的场景,让开发者无需改用更小的模型,也能获得更快的 GPT-5.6 Sol 推理响应。
“最高 14 倍”和“每秒 750 个 token”属于官方宣传的峰值指标,实际速度可能因请求、负载和容量而变化。OpenAI 尚未在所提供信息中公布面向所有用户的普遍开放时间或完整定价。
来源证据
Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the APIcommunity.openai.com · supporting# Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the API OpenAI is previewing Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, bringing frontier intelligence to products and workflows where every second matters. [Watch the GPT‑5.6 Sol Ultrafast demo]( GPT‑5.6 Sol on Ultrafast mode is launching first in the OpenAI API. It is currently available as a limited preview to a select group of customers, with access expanding as capacity grows. ### `Sign up for Ultrafast access updates` Learn more in the official announcement. Never was the :rocket: emoji more missing from the reactions list! :rocket: [...] ### Related topics | Topic | | Replies | Views | Activity | --- --- | GPT- 5.3-Codex-Spark Research Preview with 1000 Tokens per Second Codex announcement | 7 | 2269 | February 21, 2026 | | GPT-5.4 deep dive: pricing, context limits, and t
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the ...openai.com · supportingAugust 13, 2026 # Previewing Ultrafast mode: GPT‑5.6 Sol at up to 14X the speed A new speed class for frontier intelligence, turning speed into a competitive advantage. Today, we’re sharing an early look at Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing, launching first in the OpenAI API. Powered by Cerebras, Ultrafast generates up to 750 output tokens per second, bringing our most intelligent model to products and workflows where every second matters. [...] ## Powered by Cerebras Ultrafast marks the next step in our partnership with Cerebras to bring ultra-low-latency inference to OpenAI’s platform. Now, with GPT‑5.6 Sol on Ultrafast mode, Cerebras is supporting OpenAI’s most intelligent model, delivering up to 750 output tokens per second, enabling businesses to build more responsive products, make faster decisions, and bring powerful AI directly into their most demanding workflows. ## Availability GPT‑5.6 Sol on Ultrafast mode is
OpenAI previews Ultrafast GPT-5.6 Sol at up to 14x speededtechinnovationhub.com · supportingOpenAI has previewed a new Ultrafast service tier for GPT-5.6 Sol that it says can run the model at up to 14 times the speed of Standard processing, bringing lower-latency access to its frontier model first through the OpenAI API. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, according to OpenAI. The company is initially testing the service with a limited group of customers across coding, commerce, financial research, customer support and other interactive applications. Rather than using a smaller or more specialized model when response speed is critical, developers can use GPT-5.6 Sol at substantially higher inference speeds. [...] #### Cerebras powers the new service tier Ultrafast extends OpenAI's existing partnership with Cerebras, which is providing the infrastructure for the low-latency inference service. OpenAI says Cerebras is supporting GPT-5.6 Sol at speeds of up to 750 output tokens per second. The company describes the service as a new
OpenAI’s GPT-5.6 Sol runs up to 14× faster with Ultrafast mode - Help Net Securityhelpnetsecurity.com · supportingHelp Net Security newsletters: Daily and weekly news, cybersecurity jobs, open source projects, breaking news – subscribe here! Anamarija Pogorelec # OpenAI’s GPT-5.6 Sol runs up to 14× faster with Ultrafast mode OpenAI’s GPT-5.6 Sol on Ultrafast mode is available in limited preview to a select group of customers, launching first through the OpenAI API. The company says the service runs up to 14 times faster than Standard processing and generates up to 750 output tokens per second. Ultrafast is powered by Cerebras as part of the companies’ partnership on ultra-low-latency inference. GPT-5.6 Sol Ultrafast GPT-5.6 Sol Ultrafast ###### GPT-5.6 Sol Ultrafast and standard build a working 3D warehouse simulator from the same text prompt, side by side. (Source: OpenAI) [...] GPT-5.6 Sol Ultrafast and standard build a working 3D warehouse simulator from the same text prompt, side by side. (Source: OpenAI) During the preview, customers are testing Ultrafast across coding, commerce, finan
Ultrafast GPT-5.6 Sol Launched in OpenAI API | OpenAI posted on the topic | LinkedInlinkedin.com · supporting11,620,649 followers A first look at Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. Powered by Cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every second counts. Ultrafast is designed for businesses where faster frontier intelligence creates a measurable advantage, including real-time voice and customer support, commerce, coding and design, financial research, and security response. To view or add a comment, sign in [...] 11,620,649 followers A first look at Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. Powered by Cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every seco
OpenAI on X: "Previewing Ultrafast modex.com · supportingLog inSign up ## Post # OpenAI on X: "Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows." @OpenAI OpenAI @OpenAI Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. 00:00 5:01 PM · Aug 13, 20264.1MViews 799 @OpenAI OpenAI @OpenAI Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. 00:00 5:01 PM · Aug 13, 20264.1MViews 799 @OpenAI OpenAI @OpenAI [...] 799 @OpenAI OpenAI @OpenAI Aug 13 Powered by @cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every second co