Skip to content
Open access

Dynamic Customer Equity Optimization: A Reinforcement Learning Framework for Sequential Marketing Decision-Making in Volatile Markets

Jul 2026 · International Journal of Business and Management Sciences · Vol 7, pp. 351-370 · 0 citations

TL;DR

It is argued that dynamic customer equity optimization is a paradigm shift; instead of reactive, campaign-based marketing, dynamic customer equity optimization is proactive, relationship-oriented value co-creation.

Abstract

In a world of unparalleled market volatility and fragmented customer journeys, the old customer equity management models based on fixed segmentation and post-hoc analytics have not been sufficient to capture the dynamic development of customer-firm relationships. This paper presents an elaborate reinforcement learning (RL) model of dynamic customer equity optimization, which views marketing decisions as adaptive interventions that are sequential in non-stationary environment. To construct a practically implementable and theoretically based architecture of real-time marketing decision-making, we combine recent developments in the deep reinforcement learning, causal inference, and customer lifetime value (CLV) modeling. The framework combines: (1) multi-response state models that maintain Markov properties whilst learn online customer value signals; (2) conservative Q-learning to ensure reliable policy learning on offline data; (3) factor sensitive reward designs that include time varying customer engagement dynamics; and (4) multi-objective optimization that balances acquisition, retention and profitability goals. Empirical results on a variety of industry applications show that RL-based methods obtain significant improvements over constant baselines, and reported improvements in targeting efficiency of 27% (Qini coefficient), ROI gains of 18-58 and CLV impact gains of 45-85 (Wang and Chen, 2025). We cover theoretical background, issues in implementation and research directions in the future by arguing that dynamic customer equity optimization is a paradigm shift; instead of reactive, campaign-based marketing, dynamic customer equity optimization is proactive, relationship-oriented value co-creation. The paper ends by highlighting research gaps that are crucial to fill and outlining an agenda to further develop the combination of reinforcement learning and customer equity theory.

Read PDF

Similar papers

Open access Sep 2026

A Market-Aware Dynamic Trading Strategy Based on Unsupervised Clustering and Reinforcement Learning

The complexity and dynamics of the stock market make traditional absolute price prediction models based on supervised learning face great risks and limitations in practical business applications. This report proposes an innovative Hybrid Deep Reinforcement Learning trading framework to shift the forecast target from st...

Yuan Zhuang · 0 citations
Preprint Aug 2026

Concentrated Liquidity Provision: a Reinforcement Learning Perspective

Automated market makers (AMMs) are a cornerstone of decentralised finance (DeFi). Constant product markets with concentrated liquidity, such as UniswapV3, are now a well-established design. In these markets, liquidity providers (LPs) face a sequential decision problem: they must decide when to rebalance their positions...

Georgios Chionas, Charalampos Kleitsikas, Stefanos Leonardos et al. · 0 citations
Conference Aug 2026

Uncertainty-Aware Forecast-Conditioned Reinforcement Learning for Multi-Asset Algorithmic Trading

Financial markets are challenging to navigate due to changing regimes, high volatility, and unpredictable investor behavior, often leading to model misspecification in classical stationary frameworks like Moving Average (MA) and Autoregressive (AR) models. To address this, we propose an uncertainty-aware framework that...

A. Verma, Arti M. K., Surjeet Kumar · 0 citations
Review Open access Aug 2026

Forecast-driven discount optimization in FMCG distribution: an ensemble learning review

Trade discounts are among the fastest-acting profit levers available to a distributor of fast-moving consumer goods (FMCG). A distributor typically receives a limited discount budget from the manufacturer and must decide how to spread it across the matrix of stock-keeping units (SKUs) and customers so that company prof...

N. Rabimov, A. Akhatov, M. Khamidov · 0 citations
Book Open access Aug 2026

Reinforcement Learning with Scenario-Context Rollout in Portfolio Management

When economic structures and market dynamics shift, classic portfolio rebalancing algorithms often suffer from unstable and degraded performance. To improve the return and robustness of portfolio management, we explore reinforcement learning (RL) and propose Scenario-Context Rollout (SCR), a macroeconomics-guided feedb...

Vanya Priscillia Bendatu, Yao Lu · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.