OpenAI 预览 GPT-5.6 Sol Ultrafast 模式:速度最高提升至 14 倍
OpenAI 正在 API 中限量预览由 Cerebras 提供支持的 Ultrafast 服务层,GPT-5.6 Sol 的输出速度最高可达每秒 750 个 token。
OpenAI 发布了 Ultrafast 模式的早期预览。这是一项面向 GPT-5.6 Sol 的新服务层,官方称其处理速度最高可达到 Standard 处理模式的 14 倍,首先通过 OpenAI API 提供。
Ultrafast 由 Cerebras 的基础设施提供支持,官方给出的峰值为每秒生成最多 750 个输出 token。该模式面向需要低延迟响应的产品和业务流程,例如实时决策、交互式应用和复杂客户支持。
目前,GPT-5.6 Sol 的 Ultrafast 模式仅向一小部分客户开放预览。OpenAI 表示,随着可用容量增加,未来将逐步扩大访问范围。上述速度为厂商公布的最高值,不代表所有请求都能达到这一水平。
来源证据
Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the APIcommunity.openai.com · supporting# Ultrafast mode preview: GPT‑5.6 Sol at up to 14X the speed in the API OpenAI is previewing Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, bringing frontier intelligence to products and workflows where every second matters. [Watch the GPT‑5.6 Sol Ultrafast demo]( GPT‑5.6 Sol on Ultrafast mode is launching first in the OpenAI API. It is currently available as a limited preview to a select group of customers, with access expanding as capacity grows. ### `Sign up for Ultrafast access updates` Learn more in the official announcement. Never was the :rocket: emoji more missing from the reactions list! :rocket: [...] ### Related topics | Topic | | Replies | Views | Activity | --- --- | GPT- 5.3-Codex-Spark Research Preview with 1000 Tokens per Second Codex announcement | 7 | 2269 | February 21, 2026 | | GPT-5.4 deep dive: pricing, context limits, and t
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the ...openai.com · supportingAugust 13, 2026 # Previewing Ultrafast mode: GPT‑5.6 Sol at up to 14X the speed A new speed class for frontier intelligence, turning speed into a competitive advantage. Today, we’re sharing an early look at Ultrafast, a new service tier that runs GPT‑5.6 Sol up to 14× faster than Standard processing, launching first in the OpenAI API. Powered by Cerebras, Ultrafast generates up to 750 output tokens per second, bringing our most intelligent model to products and workflows where every second matters. [...] ## Powered by Cerebras Ultrafast marks the next step in our partnership with Cerebras to bring ultra-low-latency inference to OpenAI’s platform. Now, with GPT‑5.6 Sol on Ultrafast mode, Cerebras is supporting OpenAI’s most intelligent model, delivering up to 750 output tokens per second, enabling businesses to build more responsive products, make faster decisions, and bring powerful AI directly into their most demanding workflows. ## Availability GPT‑5.6 Sol on Ultrafast mode is
GPT-5.6 Sol goes 14x faster as OpenAI launches Ultrafast mode powered by Cerebrasthe-decoder.com · supportingThe inference acceleration comes from Cerebras, which signed a ten-billion-dollar partnership with OpenAI earlier this year. The service will initially be available only through the OpenAI API for GPT-5.6 Sol and limited to select customers. OpenAI plans to expand access gradually as capacity grows. Companies that want in can sign up for updates through a form. ## Faster output opens up new use cases [...] Ad OpenAI pitches several other scenarios. In finance, the model could evaluate market signals and flag suspicious transactions while conditions are still shifting. In customer support, complex inquiries could be resolved in real time, even when finding the answer requires multiple steps or systems. In e-commerce, it could answer product questions, check inventory levels, and personalize recommendations before a hesitant buyer abandons their cart. Ad Video: Generating a 3D warehouse simulator, shown on the left in Ultrafast mode [...] Skip to content Sign In Register Subscrib
GPT-5.6 Sol Ultrafast: OpenAI Previews 14x Inferencedigitalapplied.com · supporting## The questions we getevery week. Ultrafast is a new OpenAI service tier, previewed on August 13, 2026, that runs GPT-5.6 Sol — the most capable model in the GPT-5.6 family — on Cerebras hardware at what OpenAI describes as up to 14x the speed of Standard processing, peaking at 750 output tokens per second. It is not a new model: Sol itself was previewed in June 2026 and became broadly available in July. Ultrafast launches first in the OpenAI API as a limited preview for a select group of customers, with access managed through a waitlist that OpenAI says will expand as capacity grows. Related dispatches ## Continue exploringfrontier releases. AI Development #### GPT-5.6 Goes Public: GA Pricing, Ultra Mode and Access [...] vendor-stated ceiling ▼up to Claimed multiple 14x vs Standard processing ▼up to Ultrafast pricing disclosed at preview GA timeline waitlist-gated access GPT-5.6 Sol Ultrafast is OpenAI’s newest service tier, previewed on August 13, 2026: the company’s
OpenAI previews Ultrafast GPT-5.6 Sol at up to 14x speed | ETIH EdTech News — EdTech Innovation Hubedtechinnovationhub.com · supportingOpenAI has previewed a new Ultrafast service tier for GPT-5.6 Sol that it says can run the model at up to 14 times the speed of Standard processing, bringing lower-latency access to its frontier model first through the OpenAI API. Powered by Cerebras, Ultrafast can generate up to 750 output tokens per second, according to OpenAI. The company is initially testing the service with a limited group of customers across coding, commerce, financial research, customer support and other interactive applications. Rather than using a smaller or more specialized model when response speed is critical, developers can use GPT-5.6 Sol at substantially higher inference speeds. [...] #### Cerebras powers the new service tier Ultrafast extends OpenAI's existing partnership with Cerebras, which is providing the infrastructure for the low-latency inference service. OpenAI says Cerebras is supporting GPT-5.6 Sol at speeds of up to 750 output tokens per second. The company describes the service as a new
OpenAI on X: "Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows." / Xx.com · supporting799 @OpenAI OpenAI @OpenAI Aug 13 Powered by @cerebras, Ultrafast generates up to 750 tokens per second, bringing our most intelligent model to products and workflows where every second counts. Ultrafast is designed for businesses where faster frontier intelligence creates a measurable advantage, including “Previewing Ultrafast mode” in white text over lime-green and sky-blue flowing gradients. Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speedFrom openai.com 36 @OpenAI OpenAI @OpenAI Aug 13 [...] Log inSign up ## Post # OpenAI on X: "Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows." @OpenAI OpenAI @OpenAI Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expanded access to more businesses as capacity grows. 00:00 5: