Parallel Heuristic Search as Inference for Actor-Critic Reinforcement Learning Models (Extended Abstract)
Actor-critic models are a class of model-free deep reinforcement learning (RL) algorithms that have demonstrated effectiveness across various robot learning tasks. While considerable research has focused on improving training stability and data sampling efficiency, most deployment strategies have remained relatively si...