NSDI 26 - Octopus: Enhancing CXL Memory Pods via Sparse Topology @UsenixOrg
NSDI 26 - Octopus: Enhancing CXL Memory Pods via Sparse Topology  @UsenixOrg
Uploaded June 2026 | Updated September 2026, 3 weeks ago
Octopus: Enhancing CXL Memory Pods via Sparse Topology

Yuhong Zhong, Columbia University; Fiodar Kazhamiaka, Pantea Zardoshti, Shuwei Teng, and Rodrigo Fonseca, Microsoft Azure; Mark D. Hill, University of Wisconsin–Madison; Daniel S. Berger, Microsoft Azure and University of Washington

The Compute Express Link (CXL) interconnect enables compute "pods" that pool memory across servers to reduce cost and improve efficiency. These pods also facilitate pairwise communication whose needs conflict with pooling. Importantly, existing pod designs are small or require indirection through expensive switches. These conventional designs implicitly assume that pods must fully connect all servers to all CXL pooling devices.

This paper breaks with this conventional wisdom by introducing Octopus pods. Octopus directly connects servers to low-port-count CXL pooling devices (e.g., 4 ports) yet scales to large pods without switches by constructing a sparse CXL topology in which each pooling device connects to a carefully chosen subset of servers. Octopus explicitly balances "overlap", where two servers connect to the same pooling device: overlap reduces pooling efficiency but enables low-latency communication. Octopus resolves this tension by grouping servers into "islands" with low-latency intra-island communication and interconnecting islands to favor pooling.

We build a three-server CXL pod prototype and simulate scaled pods with 96 servers under measured device characteristics and physical constraints (1.5m copper cables). On hardware, Octopus RPCs are 3.2× faster than in-rack RDMA and 2.4× faster than CXL switches. In simulation, Octopus achieves net server cost savings of 3–5.4% whereas CXL switches result in a net cost increase.

View the full NSDI '26 program at usenix.org/conference/nsdi26/technical-sessions
NSDI 26 - Octopus: Enhancing CXL Memory Pods via Sparse TopologyNSDI 26 - CCEval: Accurately and Confidently Evaluating Performance Metrics of Congestion...NSDI 26 - HyperEdge: An Edge CDN Infrastructure for Cost Efficient Video StreamingNSDI 26 - OpenOptics: Enabling Open Research and Implementation of Optical Data Center NetworksNSDI 26 - FastServe: Iteration-Level Preemptive Scheduling for Large Language Model InferenceNSDI 26 - Ubers Failover Architecture: Reconciling Reliability and Efficiency in Hyperscale...SREcon26 Americas - Taming the Unpredictable: Reliability in ChaosNSDI 26 - RollPacker: Taming Long-Tail Rollouts for RL Post-Training with Tail BatchingNSDI 26 - KeepON: Supporting Deterministic Traffic on Standard NICsNSDI 26 - FENIX: Enabling In-Network DNN Inference with FPGA-Enhanced Programmable SwitchesComputer Security and Voting, Invited Talk by David Dill at USENIX Security 07PEPR 26 - Toward Provably Private Insights into AI Use
USENIX |

NSDI '26 - Octopus: Enhancing CXL Memory Pods via Sparse Topology

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER