SOTAVerified

Value prediction

Papers

Showing 51–60 of 83 papers

TitleStatusHype
AutoDIME: Automatic Design of Interesting Multi-Agent Environments—0
Why Should I Trust You, Bellman? The Bellman Error is a Poor Replacement for Value Error—0
CoRGi: Content-Rich Graph Neural Networks with Attention—0
X-model: Improving Data Efficiency in Deep Learning with A Minimax Model—0
Uncertainty-Based Offline Reinforcement Learning with Diversified Q-EnsembleCode1
Why Should I Trust You, Bellman? Evaluating the Bellman Objective with Off-Policy Data—0
Understanding and Leveraging Overparameterization in Recursive Value Estimation—0
On the Estimation Bias in Double Q-LearningCode0
Generative Self-training for Cross-domain Unsupervised Tagged-to-Cine MRI Synthesis—0
RCURRENCY: Live Digital Asset Trading Using a Recurrent Neural Network-based Forecasting System—0
Show:102550
← PrevPage 6 of 9Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1DFSudMRR73.6—Unverified