Skip to content

CAPS: Fine-Tuning CCA Timing

Jul 2026 · arXiv.org · Vol abs/2607.22821 · 0 citations · 19 references
Computer Science

TL;DR

The phase-locked steady state for dumbbell topologies under equal RTT, heterogeneous RTT, and bidirectional traffic is characterized, and the mechanism on a fat-tree under incast, permutation, and all-to-all traffic is validated.

Abstract

Data-center congestion control targets high throughput, fair bandwidth allocation, and low latency. Modern transports couple rate computation and packet scheduling into a single feedback loop, converging to near-optimal rates but leaving standing queues that can scale with the number of flows. We argue that separating the two reveals a simpler design point. Given stable feasible rates, the residual queue problem reduces to a timing problem: if every flow's packets arrive at the bottleneck in the correct slot, the link stays busy and the queue stays empty. Clocked ACK-Paced Synchronization CAPS is a lightweight distributed scheduling layer that achieves this by phase-locking each sender's transmissions to ACK-clocked bottleneck slots, with a per-flow correction that compensates for heterogeneous RTTs. We characterize the phase-locked steady state for dumbbell topologies under equal RTT, heterogeneous RTT, and bidirectional traffic, and validate the mechanism on a fat-tree under incast, permutation, and all-to-all traffic. CAPS reduces worst-case queue occupancy by 5-10x across all tested scenarios without throughput loss.

View source

Similar papers

Book Open access Aug 2026

Synchronizing with the Scheduler: Dual-Loop Congestion Control for 5G Uplink on Commodity Devices

GBR-CC is designed, a dual-loop controller that updates the sender rate on each GBR sample, using GBR for fast adaptation and end-to-end delay trends as a conservative fallback, and improves average throughput over GCC by 50%, while reducing median playout latency by 32–53% and freeze rate by 60%.

Qiang Wu, Yuxin Liu, Tian-Yang Zhang et al. · 0 citations
Preprint Aug 2026

Scaling 5G-TSN Bridges: Operating Regimes, Scheduling, and Time Synchronisation Under Heterogeneous Industrial Traffic

The nascTime framework on OMNeT++/Simu5G is used to evaluate how many TSN endpoints a single 5G NR cell can bridge before per-flow QoS degrades, showing that sub-3 ms TSN deadlines may require radio-configuration changes such as configured grants or higher numerology.

Mohamed A. M. Seliem, U. Roedig, C. Sreenan et al. · 0 citations
Book Open access Aug 2026

CSIG: Congestion Signaling for Datacenter Transports

This work introduces CSIG, a protocol that delivers precise, multi-bit bottleneck congestion signals via a fixed-length Ethernet header, and proposes Fast Ramp-Up, a congestion control primitive that leverages these bottleneck signals to reduce median RPC latency by 20% and unclaimed bandwidth by 60% in production.

Abhiram Ravi, Nandita Dukkipati, Weiwu Pang et al. · 0 citations
#large language models Conference Open access Aug 2026

PSP: Low-Overhead Packet-Level Load Balancing for Stale-State and Bandwidth-Asymmetric Networks

Probabilistic state-proportional (PSP) dispatching, a packet-level load balancing algorithm using a Band-based discrete state representation, provides an effective balance among performance, stability, and overhead for artificial intelligence data centers.

Jia-Qi Liu, Chun-Yang Zhang, Heng Pan et al. · 0 citations
Book Open access Aug 2026

Lynx: Queueing-theoretic Congestion Control Robust to Large Number of Flows

In this paper, we propose a novel congestion control algorithm (CCA) that can maintain a low and nearly constant buffering delay while ensuring high throughput and high throughput fairness even when the number of flows sharing the same bottleneck link increases significantly. Our proposed CCA uses methods formalized in...

Satoshi Utsumi, S. Zabir, Go Hasegawa · 0 citations
Open access 2026

Churn-Aware Spectrum Admission in Low-Latency Mobile Networks

Simulations on synthetic workloads with measurement-verified parameters show that TOA-S substantially reduces reconfiguration churn while maintaining spectrum utilization and latency-compatible execution.

Chi-Jen Wu · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.