此模式並非新模型,而是在專用硬體上運行現有的 GPT 5.6 Sol,旨在為事件回應、客戶支援、金融分析、程式開發與電子商務等延遲敏感型企業工作流程提供即時 AI 能力。
目前僅限受邀的 API 客戶加入候補名單,定價尚未公布,預計後續將隨 Cerebras 產能擴充而逐步開放。
What is OpenAI's new "Ultrafast" mode for GPT-5.6 Sol, how fast is it compared to standard processing, what token throughput does it achieveOpenAI GPT-5.6 Sol Ultrafast mode: 14× faster inference powered by Cerebras wafer-scale chips (AI-generated visual)
AI 提示詞
Create a landscape editorial hero image for this Studio Global article: What is OpenAI's new "Ultrafast" mode for GPT-5.6 Sol, how fast is it compared to standard processing, what token throughput does it achieve. Article summary: Here is the full breakdown of OpenAI's Ultrafast mode for GPT-5.6 Sol:. Topic tags: general, documentation, general web, user generated. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts with fake numbers, clickbait thumbnails, icons, and tiny thumbnail layouts. Make it useful as an illustrative visual, not as factual evidence.
openai.com
2026 年 8 月 13 日,OpenAI 預覽了全新的 API 服務層「Ultrafast」,讓其最強大的模型 GPT-5.6 Sol 能以極快的推論速度運行 。此服務層由 Cerebras 晶圓級晶片(Cerebras CS-3 系統)驅動,目標是為企業應用消除傳統上模型品質與回應速度之間的取捨 。
速度與 Token 吞吐量
OpenAI 表示,Ultrafast 模式的運行速度比標準處理層快高達 14 倍。該服務層可達到,約為標準 GPT-5.6 Sol 推論吞吐量的 14 倍 。對於延遲敏感的任務,這能將回應時間從數秒縮短至接近即時 。Cerebras 強調,此加速完全沒有犧牲品質,意味著 Sol Ultrafast 能以極快速度提供相同的一流水準表現 。