M-CSN: Joint Architecture and Flow Scheduling for Metro-Scale AI Fabric Based on Supernodes
Deploying trillion-parameter large language models across metropolitan environments is required to sustain real-time inference. Urban power constraints, however, prohibit monolithic GPU clusters, forcing the integration of distributed supernodes into a citywide compute fabric. Over 100-km distances, optical propagation...