Preprint
Jul 2026
Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents
This paper proposes SafeKeep, an inference-time safeguard that decouples safety judgment from tool execution: it assesses requests using flattened textual tool specifications while retaining the original schema-formatted specifications for execution.
Minghui Pan, Jiayuxuan Yang, Yuanyuan Yuan et al.
· 0 citations