| Tesseract: Tensorised Actors for Multi-Agent Reinforcement Learning | May 31, 2021 | Learning TheoryMulti-agent Reinforcement Learning | —Unverified | 0 |
| Meta Learning the Step Size in Policy Gradient Methods | May 20, 2021 | Meta-LearningMeta Reinforcement Learning | —Unverified | 0 |
| Controlling an Inverted Pendulum with Policy Gradient Methods-A Tutorial | May 17, 2021 | OpenAI GymPolicy Gradient Methods | —Unverified | 0 |
| On the Linear convergence of Natural Policy Gradient Algorithm | May 4, 2021 | Policy Gradient Methodsreinforcement-learning | —Unverified | 0 |
| Semi-On-Policy Training for Sample Efficient Multi-Agent Policy Gradients | Apr 27, 2021 | Multi-agent Reinforcement LearningPolicy Gradient Methods | —Unverified | 0 |
| Model-free Policy Learning with Reward Gradients | Mar 9, 2021 | Continuous Controlmodel | CodeCode Available | 1 |
| Softmax Policy Gradient Methods Can Take Exponential Time to Converge | Feb 22, 2021 | Policy Gradient Methods | —Unverified | 0 |
| Factored Policy Gradients: Leveraging Structure for Efficient Learning in MOMDPs | Feb 20, 2021 | Policy Gradient Methods | —Unverified | 0 |
| Strategic bidding in freight transport using deep reinforcement learning | Feb 18, 2021 | Deep Reinforcement LearningFairness | —Unverified | 0 |
| Provably Efficient Policy Optimization for Two-Player Zero-Sum Markov Games | Feb 17, 2021 | Policy Gradient MethodsVocal Bursts Valence Prediction | —Unverified | 0 |