K-Bench: A Benchmark for LLM Unlearning in Agentic Deployments
K-Bench is introduced, a benchmark that scores LLM unlearning under agentic deployment and certifies forgetting by reading the model's final answer, where a model that refuses to answer already counts as having forgotten.