Skip to content

Author

S. Pandere

We have 2 of 3 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Rewired or Gated? How Instruction Tuning Shapes Knowledge-Conflict Circuits in LLMs

In language models, the choice between believing the prompt and believing the weights is made by a handful of identifiable attention heads. Instruction tuning changes how models behave under conflict, but whether it rewires the underlying circuit or merely gates/reweights already present components, remains unknown. We...

S. Pandere, Gautam Ranka, Ritika Varshney et al. · 0 citations
#machine learning Preprint Sep 2026

Topographic Training Concentrates Causal Circuits Without Improving Neuron Monosemanticity

Mechanistic interpretability of vision transformers seeks to decompose model computation into human-readable units, but learned representations entangle many concepts in each neuron. Feature superposition is widely treated as the central obstacle to this decomposition, yet most mitigations (sparse autoencoders, diction...

Gautam Ranka, S. Pandere, Aiden Dsouza · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.