一龍馬/AI 情報站讀懂消息背後的脈絡
星期四
搜尋

這是廠商單方面引述第三方測試的說法,貼文未公開模型、負載條件與完整報告

中文摘要

AMD 宣稱在 Signal_65 於 TensorWave 代管的 RAG 測試中,Instinct MI355X 的 p99 延遲降低約 42%,並在觸及受測 SLA 前可承受約兩倍並行使用者。這是廠商單方面引述第三方測試的說法,貼文未公開模型、負載條件與完整報告。實際效能與成本效益仍需以原始測試方法與獨立驗證為準。

一龍馬判讀

採購與佈建 RAG 推論的人會拿此數據比較加速器選項,但缺少測試條件就可能誤判 SLA 餘裕。

原文節錄

AMD · @AMD

delivered ~42% lower p99 latency and roughly 2x the concurrent-user headroom…

取得全文 · 不代表內容已獨立查證

查看原文
完整收錄文字與來源

More users shouldn't have to mean slower AI. In @Signal_65 RAG testing hosted via @TensorWave, AMD Instinct MI355X delivered ~42% lower p99 latency and roughly 2x the concurrent-user headroom before breaching the evaluated SLA: https://bit.ly/45hAvl3

收錄日期
2026-10-01
來源
Nitter RSS(公開貼文)
抓取時間
2026/10/01 00:05(台北)