Huawei Introduces OceanStor M900 Storage for AI Inference
Huawei introduced OceanStor M900 Context Memory Storage at HUAWEI CONNECT 2026, the company said in a press release. The system gives AI computing clusters a shared memory space with capacity of up to 64 PB per cluster.
OceanStor M900 uses Huawei's UnifiedBus network to manage KV cache across on chip memory, DRAM, and SSDs. Its architecture provides direct NPU to SSD access, with claimed latency of 60 microseconds and aggregate bandwidth of 40 TB per second.
Huawei says the system can double inference cluster token throughput and reduce time to first token by half in typical AI programming workloads. Its storage technology distributes data based on predicted KV cache lifecycles and supports up to 24 drive writes per day.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Silicon Brief or Daily AI Brief.
Also, consider following us on social media:
More from Data Centers
Sep 17 Terragrit Secures Investment From National Grid Partners Sep 17 FirstLight Fiber Signs Hyperscaler for 240 Mile Albany to Boston Route Sep 17 cPanel Releases AI Tools for Website and Application Hosting Sep 17 BrainChip Launches AKD1500 PCIe Card for Edge AI Testing Sep 17 Huawei Cloud Stack Adds Architecture for AI AgentsSilicon Brief
Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.
Whitepaper
Tensordyne Napier: What If One Rack Could Do the Work of Nine?
Tensordyne
This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.
Read moreYou may also like
Huawei Unveils Grid Interactive AI Data Center Solution
Primemas Shows CXL Memory Products for Abaco AI System
MemryX and Lenovo Expand Edge AI Work in Saudi Arabia
ASUS Expands AI Infrastructure From Cloud to Edge
Acer Introduces Veriton RI110 AI Mini Workstation
Daily AI Brief: the AI news that matters, in your inbox.