AMD· @AMD · X·· 10 天前AI 评分39
AI 导读
更多用户不应意味着更慢的 AI。 在 @TensorWave 托管的 @Signal_65 RAG 测试中,AMD Instinct MI355X 实现了约 42% 更低的 p99 延迟,且在突破所评估 SLA 之前,并发用户余量约为 2 倍: https://bit.ly/45hAvl3
AI 生成摘要 · 以原文为准
正文
More users shouldn't have to mean slower AI.
In @Signal_65 RAG testing hosted via @TensorWave, AMD Instinct MI355X delivered ~42% lower p99 latency and roughly 2x the concurrent-user headroom before breaching the evaluated SLA: https://bit.ly/45hAvl3
来源:AMD · x.com