Skip to content

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

K-Bench: A Benchmark for LLM Unlearning in Agentic Deployments

K-Bench is introduced, a benchmark that scores LLM unlearning under agentic deployment and certifies forgetting by reading the model's final answer, where a model that refuses to answer already counts as having forgotten.

Guang-Sheng Yu, Yan-Na Jiang, Qin Wang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Diversity Combining for Multi-Path LLM Reasoning

Multi-path reasoning methods such as self-consistency (SC) sample $K$ reasoning paths and choose the most frequent answer. However, their gains quickly plateau as $K$ increases, and existing methods do not predict when this saturation will occur. We formalize multi-path LLM reasoning as a diversity combining problem fr...

Guang-Sheng Yu, Litianyi Zhang, Qin Wang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

K-Bench: A Benchmark for LLM Unlearning in Agentic Deployments

Unlearning benchmarks such as TOFU and MUSE certify forgetting by reading the model's final answer, where a model that refuses to answer already counts as having forgotten. We show that this model-level certificate does not transfer once the model is deployed as an agent. We introduce K-Bench, a benchmark that scores L...

Guangsheng Yu, Yanna Jiang, Qin Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.