Skip to content

Author

Yidong Li

We have 6 of 66 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Open access Sep 2026

Caching with Monotonicity and Consistency

We propose monotonic consistent caching (MCC), a cache scheme for applications that demand transactional guarantees. MCC warrants that a transaction-like request always sees a consistent view of the backend database and that observed writes over the cache will not be lost, even if it operates on conventional cache syst...

Shuai An, Yang Cao, Jia Li et al. · 0 citations
Open access Aug 2026

Best of both worlds: contextual feature explanation

Formal feature explanations strictly maintain perfect conformity but are intractable to compute, while heuristic methods are much faster but can lead to problematic explanations due to lack of conformity guarantees. We propose relative keys that have the best of both worlds. Relative keys associate feature explanations...

Shuai An, Yang Cao, Jia Li et al. · 1 citation
Book Open access Sep 2026

Cross-Layer Performance Analysis of Single-GPU Large Language Model Inference

This work presents a cross-layer analysis approach for single-GPU LLM inference that jointly characterizes latency and memory behavior, and systematically characterize representative dense and Mixture-of-Experts models under diverse workloads on a single A100 GPU.

Zong-Xing Zhao, Xia-Qing Li, Ze-Kai Meng et al. · 0 citations
#large language models Book Open access Sep 2026

Cross-Layer Performance Analysis of Single-GPU Large Language Model Inference

The performance of single-GPU LLM inference is characterized by strong cross-layer interactions spanning model architecture, runtime scheduling, operator execution, and GPU microarchitecture. Unfortunately, a unified understanding of single-GPU LLM inference bottlenecks is still lacking due to two limitations: the lack...

Zong-Xing Zhao, Xiaqing Li, Ze-Kai Meng et al. · 0 citations
Open access Sep 2026

Caching with Monotonicity and Consistency

We propose monotonic consistent caching (MCC), a cache scheme for applications that demand transactional guarantees. MCC warrants that a transaction-like request always sees a consistent view of the backend database and that observed writes over the cache will not be lost, even if it operates on conventional cache syst...

Shuai An, Yang Cao, Jia Li et al. · 0 citations

CSFL: Communication-Efficient Semi-Asynchronous Federated Learning Method in Resource-Constrained Edge Computing

Federated learning (FL) is a distributed machine learning (ML) paradigm that has been widely used to train ML models on massive amounts of data in edge computing (EC) environments. However, FL faces significant challenges from device heterogeneity, edge dynamics, and limited communication resources. To address these ch...

Junyi Deng, Jiahua Liu, Yanheng Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.