Skip to content
Book Open access

NetLoom: Accelerating the Construction of Large-Scale Network Emulation on a Single Host

Sep 2026 · Proceedings of the 17th ACM SIGOPS Asia-Pacific Workshop on Systems · pp. 103-110 · 0 citations · 36 references

TL;DR

This work revisit VN construction from a kernel-execution perspective, and presents NetLoom, a pipelined network emulation construction framework that decouples VN construction into active execution and blocking phases, co-optimizing them through two complementary techniques: overlapped scheduling to mitigate pipeline starvation, and kernel-level pruning to accelerate execution along the synchronous control path.

Abstract

Container-based network emulation has emerged as a scalable approach for high-fidelity network experiments. However, existing emulators spend a disproportionate amount of time constructing the virtual network (VN) rather than executing experiments, severely degrading efficiency. This bottleneck arises primarily from the substantial overhead of virtual device instantiation and the time-consuming, synchronous kernel execution paths for network configuration. To address this challenge, we revisit VN construction from a kernel-execution perspective. Our analysis shows that the conventional staged workflow strictly adheres to time-consuming kernel synchronization constraints, overlooking the in-kernel opportunities to accelerate the construction. Motivated by these insights, we present NetLoom, a pipelined network emulation construction framework. It decouples VN construction into active execution and blocking phases, co-optimizing them through two complementary techniques: overlapped scheduling to mitigate pipeline starvation, and kernel-level pruning to accelerate execution along the synchronous control path. Experimental results on two representative large-scale topologies show that NetLoom achieves a 4.06–7.03× speedup in setup time across the evaluated scales, compared to the conventional staged workflow on a single host.

Read PDF

Similar papers

Book Open access Aug 2026

Nüwa: A Generative Control Plane for AI Network Simulation

Nüwa is presented, which views routing as a compilation problem, it leverages the hierarchical and symmetric structure common in AI fabrics and compiles a declarative topology description together with routing policies into compact forwarding artifacts that are fast to generate and efficient to look up.

Wenkai Li, Ran Shu, Peng Zhang et al. · 0 citations
Book Open access Aug 2026

Spillway: Orchestrating DPU and Host into a Unified vSwitching Fabric

Spillway introduces a DPU-host hybrid data plane that repurposes idle host CPU resources to process spillover traffic when the DPU becomes the bottleneck, and decouples virtual switching capacity from static DPU hardware limits.

Xiaochong Jiang, Dian Fan, Yilong Lv et al. · 0 citations
Book Open access Aug 2026

PReCCL: Performant and Resilient Collective Communication via Integrated Inband Telemetry and Workload Reallocation

PReCCL is a drop-in NCCL replacement that combines software inband telemetry with cross-VT workload reallocation, and implements in-band monitoring within the CCL, and precisely measures the stall counts of each VT, and piggybacks the telemetry meta-data on existing collective traffic.

Zhiyong Chen, Kaihui Gao, Li Chen et al. · 0 citations
Book Open access Sep 2026

StreamTrace: Fast Trace Analysis for Large-Scale Parallel Applications on a Single Node

Trace-based performance analysis provides essential insights for understanding and optimizing large-scale parallel applications. However, traces from applications running on tens of thousands of processes can easily exceed terabytes, far surpassing the memory capacity of typical computing nodes. Existing approaches eit...

Yu-Yang Jin, Ji-Dong Zhai · 0 citations
Open access Sep 2026

Beyond Hardware Mixers: A Resilient Cloud-Native Architecture for Real-Time Remote Video Production

Professional live video production has traditionally been limited by proprietary hardware mixers and monolithic desktop applications, which impose prohibitive costs and restrict scalable, command-line-based implementation for remote integration (REMI) workflows. To address these limitations, this article presents Vocto...

Martín Herranz Sánchez, Álvaro Llorente, A. del Río et al. · 0 citations
2026

M-CSN: Joint Architecture and Flow Scheduling for Metro-Scale AI Fabric Based on Supernodes

Deploying trillion-parameter large language models across metropolitan environments is required to sustain real-time inference. Urban power constraints, however, prohibit monolithic GPU clusters, forcing the integration of distributed supernodes into a citywide compute fabric. Over 100-km distances, optical propagation...

Liang Guo, Ji-Zhuang Zhao, Wei Quan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.