Open access
Aug 2026
A dynamic recommendation strategy for Chinese language teaching resources driven by reinforcement learning
A new Decentralized Distributed Proximal using Dueling Deep Q Network (D2P-D2QN) is presented, which combines the accuracy of the D2QN estimation with the robustness of proximal policy optimization in a multi-agent setting that is distributed.
Ming Li
· Discover Artificial Intellig... · 0 citations