跳到正文
AMD· @AMD · X·· 10 天前AI 评分39
AI 导读

更多用户不应意味着更慢的 AI。 在 @TensorWave 托管的 @Signal_65 RAG 测试中,AMD Instinct MI355X 实现了约 42% 更低的 p99 延迟,且在突破所评估 SLA 之前,并发用户余量约为 2 倍: https://bit.ly/45hAvl3

AI 生成摘要 · 以原文为准

正文

More users shouldn't have to mean slower AI.

In @Signal_65 RAG testing hosted via @TensorWave, AMD Instinct MI355X delivered ~42% lower p99 latency and roughly 2x the concurrent-user headroom before breaching the evaluated SLA: https://bit.ly/45hAvl3

来源:AMD · x.com