Infrastructure Scale
Topic archive • 1 matches
AI-generated: summaries written by AI from the linked sources. How we use AI
2026-09-18
Technology
Huawei unveiled the OceanStor M900 AI memory storage, targeting hyperscale inference with a Lingqu-pooled KV cache up to 64 PB per cluster, ~60μs NPU-to-SSD latency, and ~40 TB/s aggregate bandwidth. The system aims to ease KV cache bottlenecks in large-scale AI serving.
AI Infrastructure • Pandaily
Permalink