Cross-Domain Faithfulness Evaluation of SHAP and Attention-Based Explanations in Transformer NLP Models
The results demonstrate that superior predictive performance does not necessarily correspond to higher explanation faithfulness or stronger cross-domain stability, and highlight the importance of jointly evaluating predictive performance, explanation faithfulness, and explanation robustness when developing trustworthy transformer-based NLP systems.