SOTAVerified

Atari Games

The Atari 2600 Games task (and dataset) involves training an agent to achieve high game scores.

( Image credit: Playing Atari with Deep Reinforcement Learning )

Papers

Showing 376–400 of 625 papers

TitleStatusHype
Combating Reinforcement Learning's Sisyphean Curse with Intrinsic Fear—0
Combining Off and On-Policy Training in Model-Based Reinforcement Learning—0
Combining policy gradient and Q-learning—0
Compression and Localization in Reinforcement Learning for ATARI Games—0
Compute- and Memory-Efficient Reinforcement Learning with Latent Experience Replay—0
Constraining Action Sequences with Formal Languages for Deep Reinforcement Learning—0
Contingency-Aware Exploration in Reinforcement Learning—0
Continuous-time Value Function Approximation in Reproducing Kernel Hilbert Spaces—0
Control in Stochastic Environment with Delays: A Model-based Reinforcement Learning Approach—0
Convex Regularization in Monte-Carlo Tree Search—0
CrowdPlay: Crowdsourcing human demonstration data for offline learning in Atari games—0
Curiosity in Hindsight: Intrinsic Exploration in Stochastic Environments—0
Data Efficient Training for Reinforcement Learning with Adaptive Behavior Policy Sharing—0
Deep Apprenticeship Learning for Playing Games—0
Deep AutoRegressive Networks—0
Deep Bayesian Reward Learning from Preferences—0
Deep Conservative Policy Iteration—0
Deep Learning for Real-Time Atari Game Play Using Offline Monte-Carlo Tree Search Planning—0
Deep Learning for Reward Design to Improve Monte Carlo Tree Search in ATARI Games—0
Deep Learning of Intrinsically Motivated Options in the Arcade Learning Environment—0
Deep Q-Learning with Low Switching Cost—0
Deep Q-Network for AI Soccer—0
Deep Reinforcement Learning Boosted by External Knowledge—0
Deep Reinforcement Learning for Doom using Unsupervised Auxiliary Tasks—0
Deep Reinforcement Learning for NLP—0
Show:102550
← PrevPage 16 of 25Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GDI-H3Score864—Unverified
2GDI-H3(200M frames)Score864—Unverified
3GDI-I3(200M frames)Score864—Unverified
4GDI-I3Score864—Unverified
5Bootstrapped DQNScore855—Unverified
6FQFScore854.2—Unverified
7R2D2Score837.7—Unverified
8Ape-XScore800.9—Unverified
9Agent57Score790.4—Unverified
10IMPALA (deep)Score787.34—Unverified
#ModelMetricClaimedVerifiedStatus
1GDI-H3(200M frames)Score34—Unverified
2TRPO-hashScore34—Unverified
3IQNScore34—Unverified
4NoisyNet-DuelingScore34—Unverified
5QR-DQN-1Score34—Unverified
6Go-ExploreScore34—Unverified
7GDI-I3Score34—Unverified
8GDI-H3Score34—Unverified
9Bootstrapped DQNScore33.9—Unverified
10ASL DDQNScore33.9—Unverified
#ModelMetricClaimedVerifiedStatus
1Agent57Score580,328.14—Unverified
2QR-DQN-1Score572,510—Unverified
3R2D2Score408,850—Unverified
4IMPALA (deep)Score351,200.12—Unverified
5Ape-XScore302,391.3—Unverified
6A2C + SILScore104,975.6—Unverified
7MuZero (Res2 Adam)Score94,906.25—Unverified
8DreamerV2Score94,688—Unverified
9MuZeroScore72,276—Unverified
10DNAScore52,398—Unverified
#ModelMetricClaimedVerifiedStatus
1GDI-H3Score1,000,000—Unverified
2GDI-H3(200M frames)Score1,000,000—Unverified
3Agent57Score999,997.63—Unverified
4R2D2Score999,996.7—Unverified
5MuZeroScore999,976.52—Unverified
6MuZero (Res2 Adam)Score999,659.18—Unverified
7GDI-I3Score943,910—Unverified
8Ape-XScore392,952.3—Unverified
9C51 noopScore266,434—Unverified
10Duel noopScore50,254.2—Unverified