| Co-Activation Graph Analysis of Safety-Verified and Explainable Deep Reinforcement Learning Policies | Jan 6, 2025 | Decision MakingDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Learning Coverage Paths in Unknown Environments with Deep Reinforcement Learning | Jun 29, 2023 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| Training a Resilient Q-Network against Observational Interference | Feb 18, 2021 | Causal InferenceDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| CDT: Cascading Decision Trees for Explainable Reinforcement Learning | Nov 15, 2020 | Deep Reinforcement LearningExplainable Models | CodeCode Available | 1 | 5 |
| Enhancing Battery Storage Energy Arbitrage with Deep Reinforcement Learning and Time-Series Forecasting | Oct 25, 2024 | Deep Reinforcement LearningTime Series | CodeCode Available | 1 | 5 |
| Enhancing Cooperative Multi-Agent Reinforcement Learning with State Modelling and Adversarial Exploration | May 8, 2025 | Deep Reinforcement LearningMulti-agent Reinforcement Learning | CodeCode Available | 1 | 5 |
| Character Controllers Using Motion VAEs | Mar 26, 2021 | Continuous ControlDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| A fast balance optimization approach for charging enhancement of lithium-ion battery packs through deep reinforcement learning | Apr 24, 2024 | Deep Reinforcement Learningenergy management | CodeCode Available | 1 | 5 |
| An End-to-end Deep Reinforcement Learning Approach for the Long-term Short-term Planning on the Frenet Space | Nov 26, 2020 | Decision MakingDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Chip Placement with Deep Reinforcement Learning | Apr 22, 2020 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| EpidemiOptim: A Toolbox for the Optimization of Control Policies in Epidemiological Models | Oct 9, 2020 | Deep Reinforcement LearningEpidemiology | CodeCode Available | 1 | 5 |
| ERL-Re^2: Efficient Evolutionary Reinforcement Learning with Shared State Representation and Individual Policy Representation | Oct 26, 2022 | continuous-controlContinuous Control | CodeCode Available | 1 | 5 |
| Affordance Learning from Play for Sample-Efficient Policy Learning | Mar 1, 2022 | Deep Reinforcement LearningMotion Planning | CodeCode Available | 1 | 5 |
| Cleanba: A Reproducible and Efficient Distributed Reinforcement Learning Platform | Sep 29, 2023 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| An Equivalence between Loss Functions and Non-Uniform Sampling in Experience Replay | Jul 12, 2020 | Deep Reinforcement LearningMuJoCo | CodeCode Available | 1 | 5 |
| Execution-based Code Generation using Deep Reinforcement Learning | Jan 31, 2023 | Code CompletionCode Generation | CodeCode Available | 1 | 5 |
| Co-designing Intelligent Control of Building HVACs and Microgrids | Jul 18, 2021 | Deep Reinforcement LearningReinforcement Learning (RL) | CodeCode Available | 1 | 5 |
| Deep reinforcement learning for large-scale epidemic control | Mar 30, 2020 | Computational EfficiencyDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Exploration and Anti-Exploration with Distributional Random Network Distillation | Jan 18, 2024 | D4RLDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Age-Based Scheduling for Mobile Edge Computing: A Deep Reinforcement Learning Approach | Dec 1, 2023 | Deep Reinforcement LearningEdge-computing | CodeCode Available | 1 | 5 |
| Collective eXplainable AI: Explaining Cooperative Strategies and Agent Contribution in Multiagent Reinforcement Learning with Shapley Values | Oct 4, 2021 | Decision MakingDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Collaborative Target Search with a Visual Drone Swarm: An Adaptive Curriculum Embedded Multistage Reinforcement Learning Approach | Apr 26, 2022 | Deep Reinforcement LearningManagement | CodeCode Available | 1 | 5 |
| ColO-RAN: Developing Machine Learning-based xApps for Open RAN Closed-loop Control on Programmable Experimental Platforms | Dec 17, 2021 | Deep Reinforcement LearningScheduling | CodeCode Available | 1 | 5 |
| Exploring Deep Reinforcement Learning-Assisted Federated Learning for Online Resource Allocation in Privacy-Persevering EdgeIoT | Feb 15, 2022 | Deep Reinforcement LearningEdge-computing | CodeCode Available | 1 | 5 |
| RL-I2IT: Image-to-Image Translation with Deep Reinforcement Learning | Sep 24, 2023 | Auxiliary LearningDecision Making | CodeCode Available | 1 | 5 |
| An experimental evaluation of Deep Reinforcement Learning algorithms for HVAC control | Jan 11, 2024 | Deep Reinforcement LearningIncremental Learning | CodeCode Available | 1 | 5 |
| Combining Reinforcement Learning and Constraint Programming for Combinatorial Optimization | Jun 2, 2020 | Combinatorial OptimizationDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Combining Deep Reinforcement Learning and Search for Imperfect-Information Games | Jul 27, 2020 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| Deep Reinforcement Learning for Joint Spectrum and Power Allocation in Cellular Networks | Dec 19, 2020 | Deep Reinforcement LearningManagement | CodeCode Available | 1 | 5 |
| Combining Semantic Guidance and Deep Reinforcement Learning For Generating Human Level Paintings | Nov 25, 2020 | Deep Reinforcement LearningModel-based Reinforcement Learning | CodeCode Available | 1 | 5 |
| Deep Reinforcement Learning for List-wise Recommendations | Dec 30, 2017 | Deep Reinforcement LearningRecommendation Systems | CodeCode Available | 1 | 5 |
| Agent with Warm Start and Adaptive Dynamic Termination for Plane Localization in 3D Ultrasound | Mar 26, 2021 | Deep Reinforcement LearningReinforcement Learning (RL) | CodeCode Available | 1 | 5 |
| An Efficient Asynchronous Method for Integrating Evolutionary and Gradient-based Policy Search | Dec 10, 2020 | continuous-controlContinuous Control | CodeCode Available | 1 | 5 |
| Computational Performance of Deep Reinforcement Learning to find Nash Equilibria | Apr 26, 2021 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| A Deep Reinforcement Learning Framework for the Financial Portfolio Management Problem | Jun 30, 2017 | Deep Reinforcement LearningManagement | CodeCode Available | 1 | 5 |
| Deep Reinforcement Learning for Entity Alignment | Mar 7, 2022 | Decision MakingDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| ContainerGym: A Real-World Reinforcement Learning Benchmark for Resource Allocation | Jul 6, 2023 | Decision MakingDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Connecting Deep-Reinforcement-Learning-based Obstacle Avoidance with Conventional Global Planners using Waypoint Generators | Apr 8, 2021 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| Contention Window Optimization in IEEE 802.11ax Networks with Deep Reinforcement Learning | Mar 3, 2020 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| For SALE: State-Action Representation Learning for Deep Reinforcement Learning | Jun 4, 2023 | continuous-controlContinuous Control | CodeCode Available | 1 | 5 |
| A Closer Look at Invalid Action Masking in Policy Gradient Algorithms | Jun 25, 2020 | Deep Reinforcement LearningReal-Time Strategy Games | CodeCode Available | 1 | 5 |
| GANav: Efficient Terrain Segmentation for Robot Navigation in Unstructured Outdoor Environments | Mar 7, 2021 | Deep Reinforcement LearningRobot Navigation | CodeCode Available | 1 | 5 |
| 2-Level Reinforcement Learning for Ships on Inland Waterways: Path Planning and Following | Jul 25, 2023 | Deep Reinforcement Learningreinforcement-learning | CodeCode Available | 1 | 5 |
| Continuous control with deep reinforcement learning | Sep 9, 2015 | Action Detectioncontinuous-control | CodeCode Available | 1 | 5 |
| Acme: A Research Framework for Distributed Reinforcement Learning | Jun 1, 2020 | Deep Reinforcement LearningDQN Replay Dataset | CodeCode Available | 1 | 5 |
| Continuous-Time Fitted Value Iteration for Robust Policies | Oct 5, 2021 | continuous-controlContinuous Control | CodeCode Available | 1 | 5 |
| Continuous Deep Q-Learning with Model-based Acceleration | Mar 2, 2016 | continuous-controlContinuous Control | CodeCode Available | 1 | 5 |
| Continuous Coordination As a Realistic Scenario for Lifelong Learning | Mar 4, 2021 | Continual LearningDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| Deep Reinforcement Learning for Conservation Decisions | Jun 15, 2021 | BIG-bench Machine LearningDeep Reinforcement Learning | CodeCode Available | 1 | 5 |
| An Application of Deep Reinforcement Learning to Algorithmic Trading | Apr 7, 2020 | Algorithmic TradingDeep Reinforcement Learning | CodeCode Available | 1 | 5 |