Preprint
Aug 2026
How Much Does a Reasoning Summary Reveal? An Observability Ladder for Large Language Models
An observability ladder is introduced that holds each completed run fixed and varies only what a reader inspects to judge whether the answer is correct: the response, a self-summary the model writes from the trace, the trace itself, and internal signals, each with and without the prompt.
A. Algaba, Francesca Carlon, Lynn Delcon et al.
· 0 citations