Book
Open access
Aug 2026
Interpretability in the Era of Large Language Models: Mechanistic Methodology, Empirical Practices, and Applications
This tutorial provides a comprehensive, end-to-end view of LLM interpretability, transitioning from microscopic neural analysis to macroscopic application and deployment, and explores how these interpretability paradigms scale and inspire the design of frontier architectures, agentic systems, and thinking models.
Wei Zhang, Zhengfu He, Lucia Zhang et al.
· Proceedings of the 32nd ACM... · 0 citations