Skip to content

Author

Size Zheng

We have 2 of 12 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

VarioPath: Workload-Aware All-to-All Communication for PCIe GPU Clusters

AlltoAllv communication is a critical primitive in distributed large-model inference, particularly for mixture-of-experts (MoE) models. The growing adoption of PCIe GPU systems for cost-efficient inference makes AlltoAllv performance on these systems increasingly important. Without a dedicated scale-up interconnect (e....

Yao Fei, Jin Fang, Si-Ze Zheng et al. · 0 citations
Book Open access Jul 2026

UniEP: Unified Expert-Parallel MegaKernel MoE for LLM Training

UniEP fuses the MoE communication and computation into MegaKernels, effectively transforming complex architectural tuning into a unified parameter search space for automated adaptability and incorporates a deterministic token ordering mechanism that guarantees numerical consistency with sequential execution, even under...

Size Zheng, Xuegui Zheng, Li-Wen Chang et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.