Skip to content

Emergent Latent-State Computation under Stochastic Volatility

Jul 2026 · arXiv.org · Vol abs/2607.25459 · 0 citations · 38 references
Computer Science Economics

TL;DR

Stochastic volatility models provide a useful benchmark for mechanistic interpretability under noisy latent dynamics and partial observability, and output-head replacement shows that part of the degradation under noisy MSE training arises from readout misalignment rather than representation failure.

Abstract

Mechanistic interpretability has largely focused on language models and deterministic toy tasks. Much less is known about how sequence models internally represent latent stochastic dynamics under noisy, partially observed observations. We study this question in a controlled multivariate stochastic volatility setting, where models observe only returns while the ground-truth latent volatility state is known to the researcher. This setting provides a useful benchmark for mechanistic interpretability under partial observability: the latent state is hidden from the model but directly available for evaluation. Across architectures, losses, and output heads, we find evidence for a two-stage computation. Hidden representations encode substantial information about the next latent volatility state, and the output head maps this representation to squared return forecasts. Furthermore, in Transformers, latent-state decodability emerges at identifiable architectural stages whose location depends on the volatility period. In long-cycle regimes, this computation simplifies into an explicit latent-state filter consisting of a learned linear projection followed by $\ell^2$ normalization. Output-head replacement further shows that part of the degradation under noisy MSE training arises from readout misalignment rather than representation failure. These results suggest that stochastic volatility models provide a useful benchmark for mechanistic interpretability under noisy latent dynamics and partial observability.

View source

Similar papers

Preprint Aug 2026

Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Models

A calibrated stochastic world model can reveal how uncertain a future is without revealing why it branches. The same conditional future law can arise because an observation aliases physical states or because dynamics remain random after the declared full state is fixed. We prove that ordinary transitions cannot identif...

Yi-Xin Dong · 0 citations
Jul 2026

Susceptible Reservoir Architectures for Regime-Conditional Volatility Forecasting

Susceptible Architectures (SUSA), a reservoir-design principle for volatility forecasting, and its two concrete implementations, based on complex-valued open-chain and periodic reservoirs and regime-conditioned experts to interpret reservoir features across calm, onset, recovery, and persistent-stress states are introd...

Aliaksei Kaliutau · 0 citations
Preprint Aug 2026

UNVaMP: Neural Knowledge Tracing with Variational Regularization of Latent Knowledge Dynamics

We introduce the Unified Neural Variational Measurement of Proficiency (UNVaMP) architecture, a knowledge tracing method that integrates observed student-item interactions with internal memory to produce evolving latent representations of student knowledge. These representations support accurate predictions of future r...

Carson Cook, Ahmed Zerouali, Anthony Schmidt et al. · 0 citations
Preprint Jul 2026

The Fundamental Structure of Risk: From Characteristics to Covariance

Estimating the covariance structure of financial assets typically relies on historical returns, making risk models dependent on noisy and asset-specific time series. We propose the Characteristic-Driven Dynamic Factor Model (CD-DFM), a non-linear latent factor model that instead constructs a representation of the asset...

Alexandre Alouadi, Charles-Albert Lehalle · 0 citations
Jul 2026

Information Bottleneck Learning for Faithful Time Series Forecasting Explanations

As forecasts increasingly drive decisions in fields such as energy, transportation, and healthcare, understanding the historical data behind these predictions has become as crucial as the predictions themselves. Although existing interpretable-by-design forecasters reveal their internal structures, they offer no guaran...

Xu Zheng, Wei Cheng, Zhuomin Chen et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.