Skip to content
AI IntelligenceSep 18, 2026AI Intelligence
Article

Huawei unveiled the OceanStor M900 AI memory storage…

…targeting hyperscale inference with a Lingqu-pooled KV cache up to 64 PB per cluster, ~60μs NPU-to-SSD latency, and ~40 TB/s aggregate bandwidth. The system aims to ease KV cache bottlenecks in large-scale AI serving.

AI-generated: summaries written by AI from the linked sources. How we use AI

AI-generatedSource: Pandaily
01

Source Brief

Huawei unveiled the OceanStor M900 AI memory storage, targeting hyperscale inference with a Lingqu-pooled KV cache up to 64 PB per cluster, ~60μs NPU-to-SSD latency, and ~40 TB/s aggregate bandwidth. The system aims to ease KV cache bottlenecks in large-scale AI serving.