Skip to content
Intelligence IASep 18, 2026Intelligence IA
Article

Huawei unveiled the OceanStor M900 AI memory storage…

…targeting hyperscale inference with a Lingqu-pooled KV cache up to 64 PB per cluster, ~60μs NPU-to-SSD latency, and ~40 TB/s aggregate bandwidth. The system aims to ease KV cache bottlenecks in large-scale AI serving.

Généré par IA : résumés rédigés par une IA à partir des sources citées. Notre usage de l'IA

Généré par IASource: Pandaily
01

Brief source

Huawei unveiled the OceanStor M900 AI memory storage, targeting hyperscale inference with a Lingqu-pooled KV cache up to 64 PB per cluster, ~60μs NPU-to-SSD latency, and ~40 TB/s aggregate bandwidth. The system aims to ease KV cache bottlenecks in large-scale AI serving.