SOTAVerified

Offline RL

Papers

Showing 541550 of 755 papers

TitleStatusHype
DeepThermal: Combustion Optimization for Thermal Power Generating Units Using Offline Reinforcement Learning0
Delphic Offline Reinforcement Learning under Nonidentifiable Hidden Confounding0
Deploying Offline Reinforcement Learning with Human Feedback0
Design from Policies: Conservative Test-Time Adaptation for Offline Policy Optimization0
Development and Validation of Heparin Dosing Policies Using an Offline Reinforcement Learning Algorithm0
Dialogue Evaluation with Offline Reinforcement Learning0
DIAR: Diffusion-model-guided Implicit Q-learning with Adaptive Revaluation0
DiffPoGAN: Diffusion Policies with Generative Adversarial Networks for Offline Reinforcement Learning0
DiffStitch: Boosting Offline Reinforcement Learning with Diffusion-based Trajectory Stitching0
Diffused Task-Agnostic Milestone Planner0
Show:102550
← PrevPage 55 of 76Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1KFCAverage Reward81.8Unverified
2ADMPOAverage Reward81Unverified
3Decision Transformer (DT)Average Reward73.5Unverified
#ModelMetricClaimedVerifiedStatus
1ParPID4RL Normalized Score151.4Unverified