| DEER: A Delay-Resilient Framework for Reinforcement Learning with Variable Delays | Jun 5, 2024 | MuJoCoReinforcement Learning (RL) | —Unverified | 0 |
| Value Improved Actor Critic Algorithms | Jun 3, 2024 | MuJoCo | —Unverified | 0 |
| Trust the Model Where It Trusts Itself -- Model-Based Actor-Critic with Uncertainty-Aware Rollout Adaption | May 29, 2024 | modelModel-based Reinforcement Learning | CodeCode Available | 0 |
| A Pontryagin Perspective on Reinforcement Learning | May 28, 2024 | MuJoCoreinforcement-learning | —Unverified | 0 |
| Imitating from auxiliary imperfect demonstrations via Adversarial Density Weighted Regression | May 28, 2024 | Imitation LearningMuJoCo | CodeCode Available | 0 |
| Symmetric Reinforcement Learning Loss for Robust Learning on Diverse Tasks and Model Scales | May 27, 2024 | Atari GamesMuJoCo | CodeCode Available | 0 |
| Adaptive Q-Network: On-the-fly Target Selection for Deep Reinforcement Learning | May 25, 2024 | Atari GamesAutoML | —Unverified | 0 |
| Variational Delayed Policy Optimization | May 23, 2024 | MuJoCoReinforcement Learning (RL) | CodeCode Available | 0 |
| Learning rigid-body simulators over implicit shapes for large-scale scenes and vision | May 22, 2024 | MuJoCo | —Unverified | 0 |
| Pure Planning to Pure Policies and In Between with a Recursive Tree Planner | May 21, 2024 | MuJoCo | —Unverified | 0 |