Volatility-Consistent Reward Shaping for Deep Reinforcement Learning Market Makers
Deep reinforcement learning (DRL) is increasingly utilized for optimal execution and market making. However, standard DRL formulations typically rely on static inventory penalties to control risk. In this paper, we observe that applying a static inventory penalty may induce a volatility-inconsistent implicit risk prefe...