Red Hat AI 3.5 的主軸不是推出另一個聊天機械人,而是讓企業可在 AI 上線前做安全評估、上線後管治共享 GPU,並持續監察推理服務。 EvalHub 已正式可用;但 AutoRAG 與 Amazon EKS 上的 llm d 分散式推理仍屬 Technology Preview,未必適合承載正式生產工作負載。
What did Red Hat announce with the general availability of Red Hat AI 3.5 on September 11, 2026, and how does the release—including its inteRed Hat AI 3.5 emphasizes safety evaluation, GPU operations and observability for enterprise AI deployments.
AI 提示
Create a landscape editorial hero image for this Studio Global article: What did Red Hat announce with the general availability of Red Hat AI 3.5 on September 11, 2026, and how does the release—including its inte. Article summary: Red Hat announced the general availability of Red Hat AI 3.5 as an enterprise AI platform update focused on making AI a governed, observable, multi-tenant production service across hybrid environments—not merely a collec. Topic tags: general, documentation, general web. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts with fak
openai.com
Red Hat 在 2026 年 9 月 9 日公布 Red Hat AI 3.5,定位並非一項單獨的新模型或聊天功能,而是針對企業 AI「真係要上線營運」而設的更新。到 9 月 11 日的官方重點回顧,Red Hat 再次強調安全、營運控制與效能透明度:目標是把 AI 由零散試行項目,變成可共同使用、可管治的企業服務。4315
今次最明確的生產級功能,是 EvalHub 正式可用(GA)。它是 Red Hat 的模型與 AI agent 評估工具,適用於客戶自行提供或客製化的模型、檢索增強生成(RAG)設定,以及 AI agent。團隊可針對提示注入(prompt injection)、越獄提示(jailbreak)等風險做安全基準測試,並按結果產生面向合規要求的認證文件。13
Red Hat 亦表示,已驗證模型目錄新增超過 20 個模型,包括來自 Google、NVIDIA 及 Alibaba Cloud 的模型。目錄的評估資訊包含 Garak 基準測試結果,以及毒性和可能洩露個人可識別資料(PII)的指標,讓團隊揀模型時有更多可參考的證據。34
共用 GPU:由「邊個搶到就邊個用」變成可管理資源
AI 試行階段通常有專屬基建、用戶亦不多;一旦投入正式運作,工作負載之間會爭資源,不同服務有不同優先次序,成本亦要分攤。Red Hat AI 3.5 新增的控制項包括公平分享排程(fair-share scheduling)、按優先級提供服務、准入控制(admission control),以及按優先級路由請求,目的是管理多人共用 GPU 的環境。48
Red Hat AI 3.5 的主軸不是推出另一個聊天機械人,而是讓企業可在 AI 上線前做安全評估、上線後管治共享 GPU,並持續監察推理服務。
首先要驗證的關鍵點是什麼?
Red Hat AI 3.5 的主軸不是推出另一個聊天機械人,而是讓企業可在 AI 上線前做安全評估、上線後管治共享 GPU,並持續監察推理服務。 EvalHub 已正式可用;但 AutoRAG 與 Amazon EKS 上的 llm d 分散式推理仍屬 Technology Preview,未必適合承載正式生產工作負載。