Policy Optimization and Statistical Inference for Online Contextual Matrix Games
This work introduces online contextual matrix games and proposes OnGameLearn, an online learning algorithm that efficiently balances exploration and exploitation across both player actions and contexts and effectively navigates the intertwined challenges of strategic and contextual decision-making.
Liner Xiang, Yixin Wang, Hengrui Cai
· 0 citations