Preprint
Aug 2026
Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems
It is asked whether committees of language-model agents deliberating on a shared workspace can be gamed by shortcuts, cues a benchmark rewards but a clinician would ignore, and what games a committee is social plausibility.
S. Ordóñez, Agastya Munnangi, Aldo Marzullo et al.
· 0 citations