How Perturbations Propagate: A Multi-Level Analysis of Robustness in Large Language Models
The results show that robustness claims based on a single behavioral or representational metric can be misleading, and motivate multi-level evaluation of how perturbations alter language-model computation.
Dun-Li Chan, Emily Liu, Niyathi Allu et al.
· 0 citations