Preprint
Aug 2026
MaliciousSkillBench: A Comprehensive Benchmark for Malicious Agent Skill Detection
The results show that reliable malicious-Skill detection requires both broader cross-source benchmark coverage and evaluation that jointly measures attack detection and benign over-flagging.
Yue Wang, Yi Liu, Gelei Deng et al.
· 0 citations