Skip to content
Preprint

Deep Learning of Robust Market Making under Regime-Switching Order Flow

Sep 2026 · 0 citations · 39 references
Economics

TL;DR

A deep reinforcement-learning market maker (RLMM) is developed - a Rainbow-style distributional DQN (C51) which is calibrated and tested in a zero-intelligence limit order book and finds that, in the stationary setting, RLMM outperforms GLFT across the entire observed risk-return frontier.

Abstract

Classical market-making strategies based on stochastic control, such as the Avellaneda-Stoikov and the Gu\'{e}ant-Lehalle-Fernandez-Tapia (GLFT) extension, provide closed-form quoting rules, but rest on assumptions that break down at realistic microstructure timescales. One of them is that order flow is stationary, while empirical evidence points to the existence of regimes, possibly associated with algorithmic execution of metaorders. In this case, existing methods provide negative PnL. In this paper, we develop a deep reinforcement-learning market maker (RLMM) - a Rainbow-style distributional DQN (C51) which is calibrated and tested in a zero-intelligence limit order book. We find that, in the stationary setting, RLMM outperforms GLFT across the entire observed risk-return frontier. The RLMM is more robust to flow asymmetry than GLFT, but, like any stationarily trained strategy, it still suffers large drawdowns from inventory saturation under persistent directional imbalance. Augmenting the state of RLMM with two auxiliary signals - a Bayesian online change-point filter over the directional flow bias and a queue-adjusted quote-exposure imbalance -restores profitability. A final scenario-bandit step that reweights low-return regime scenarios further improves performance under random-persistence and correlated-direction stress.

View source

Similar papers

#machine learning Preprint Sep 2026

Robust Market Making with Hawkes Order Flow and Price Impact via Adversarial Reinforcement Learning

Market-making strategies in real limit order book markets face substantial model uncertainty and regime-shift risk. Existing adversarial reinforcement learning approaches improve robustness by formulating the Avellaneda--Stoikov market-making problem as a zero-sum game between a market maker and an environmental advers...

Hao-Hao Yang, Zheng Xu · 0 citations
Book Open access Aug 2026

Reinforcement Learning with Scenario-Context Rollout in Portfolio Management

When economic structures and market dynamics shift, classic portfolio rebalancing algorithms often suffer from unstable and degraded performance. To improve the return and robustness of portfolio management, we explore reinforcement learning (RL) and propose Scenario-Context Rollout (SCR), a macroeconomics-guided feedb...

Vanya Priscillia Bendatu, Yao Lu · 0 citations
Sep 2026

Volatility-Consistent Reward Shaping for Deep Reinforcement Learning Market Makers

Deep reinforcement learning (DRL) is increasingly utilized for optimal execution and market making. However, standard DRL formulations typically rely on static inventory penalties to control risk. In this paper, we observe that applying a static inventory penalty may induce a volatility-inconsistent implicit risk prefe...

Chun-Ming Zeng · 0 citations
Conference Aug 2026

Uncertainty-Aware Forecast-Conditioned Reinforcement Learning for Multi-Asset Algorithmic Trading

Financial markets are challenging to navigate due to changing regimes, high volatility, and unpredictable investor behavior, often leading to model misspecification in classical stationary frameworks like Moving Average (MA) and Autoregressive (AR) models. To address this, we propose an uncertainty-aware framework that...

A. Verma, Arti M. K., Surjeet Kumar · 0 citations
Book Open access Aug 2026

Beyond Black Boxes: An Energy-Based Unified Framework for Interpretable Stock Selection

Quantitative stock selection from large-scale market data is critical for achieving excess returns in financial markets. A persistent challenge is tail risk, which triggers non-stationary regime shifts that invalidate learned patterns and distort portfolio logic. Despite advances in deep learning, existing approaches c...

Junren Xiao, Ruiyao Miao, Zixuan Yuan et al. · 0 citations
Open access 2026

Dynamic-Flooding Transformer Ensembles for Reinforcement-Learning-Based Equity Market Timing

Accurate directional forecasting of equity time series is challenging due to small-sample bottlenecks, regime-dependent non-stationarity, and dominant idiosyncratic noise. This paper presents a coordinated quantitative framework addressing these difficulties. First, causal Transformer classifiers operate over 78 featur...

Tung-Liang Chen, Jyh-Shing Roger Jang · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.