Learning Together and Against Each Other: How Multiple Reinforcement-Learning Agents Coordinate, Compete and Where the Field is Heading
The obstacles that distinguish MARL from its singleagent counterpart are discussed, namely a moving-target learning problem, the difficulty of dividing a shared reward among team members, limited local views, and growth of the joint action space, and why training with global information but acting on local information...