Intelligence IASep 18, 2026Intelligence IA
Article
Huawei unveiled the OceanStor M900 AI memory storage…
…targeting hyperscale inference with a Lingqu-pooled KV cache up to 64 PB per cluster, ~60μs NPU-to-SSD latency, and ~40 TB/s aggregate bandwidth. The system aims to ease KV cache bottlenecks in large-scale AI serving.
Généré par IA : résumés rédigés par une IA à partir des sources citées. Notre usage de l'IA
Généré par IASource: Pandaily
01
Brief source
Huawei unveiled the OceanStor M900 AI memory storage, targeting hyperscale inference with a Lingqu-pooled KV cache up to 64 PB per cluster, ~60μs NPU-to-SSD latency, and ~40 TB/s aggregate bandwidth. The system aims to ease KV cache bottlenecks in large-scale AI serving.