A comparison between Thompson sampling and greedy algorithm in portfolio selection
The findings suggest that in standard passive investment settings, the additional complexity of exploration-based reinforcement learning is not justified, and simpler estimation-based approaches deliver comparable performance.
Sasikarn Choobun
· 0 citations