SOTAVerified

Atari Games

The Atari 2600 Games task (and dataset) involves training an agent to achieve high game scores.

( Image credit: Playing Atari with Deep Reinforcement Learning )

Papers

Showing 1–25 of 625 papers

TitleStatusHype
CALE: Continuous Arcade Learning EnvironmentCode7
MineDojo: Building Open-Ended Embodied Agents with Internet-Scale KnowledgeCode4
Distributed Prioritized Experience ReplayCode3
Rainbow: Combining Improvements in Deep Reinforcement LearningCode3
Streaming Deep Reinforcement Learning Finally WorksCode3
Intelligent Go-Explore: Standing on the Shoulders of Giant Foundation ModelsCode2
Harfang3D Dog-Fight Sandbox: A Reinforcement Learning Research Platform for the Customized Control Tasks of Fighter AircraftsCode2
Benchmarking Deep Reinforcement Learning for Continuous ControlCode2
Mastering Atari Games with Limited DataCode2
Conformal Symplectic Optimization for Stable Reinforcement LearningCode2
Accelerated Methods for Deep Reinforcement LearningCode2
Mastering Atari, Go, Chess and Shogi by Planning with a Learned ModelCode2
Deep Recurrent Q-Learning for Partially Observable MDPsCode1
Deep Hierarchical Planning from PixelsCode1
Deep Reinforcement Learning with Double Q-learningCode1
CURL: Contrastive Unsupervised Representations for Reinforcement LearningCode1
Contrastive Variational Reinforcement Learning for Complex ObservationsCode1
Decision Transformer: Reinforcement Learning via Sequence ModelingCode1
Developing an OpenAI Gym-compatible framework and simulation environment for testing Deep Reinforcement Learning agents solving the Ambulance Location ProblemCode1
Behavior From the Void: Unsupervised Active Pre-TrainingCode1
Reincarnating Reinforcement Learning: Reusing Prior Computation to Accelerate ProgressCode1
Atari-5: Distilling the Arcade Learning Environment down to Five GamesCode1
AdaRL: What, Where, and How to Adapt in Transfer Reinforcement LearningCode1
A Distributional Perspective on Reinforcement LearningCode1
Asynchronous Methods for Deep Reinforcement LearningCode1
Show:102550
← PrevPage 1 of 25Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GDI-H3(200M frames)Score864—Unverified
2GDI-I3(200M frames)Score864—Unverified
3GDI-H3Score864—Unverified
4GDI-I3Score864—Unverified
5Bootstrapped DQNScore855—Unverified
6FQFScore854.2—Unverified
7R2D2Score837.7—Unverified
8Ape-XScore800.9—Unverified
9Agent57Score790.4—Unverified
10IMPALA (deep)Score787.34—Unverified
#ModelMetricClaimedVerifiedStatus
1TRPO-hashScore34—Unverified
2IQNScore34—Unverified
3GDI-H3(200M frames)Score34—Unverified
4GDI-I3Score34—Unverified
5Go-ExploreScore34—Unverified
6QR-DQN-1Score34—Unverified
7NoisyNet-DuelingScore34—Unverified
8GDI-H3Score34—Unverified
9C51 noopScore33.9—Unverified
10ASL DDQNScore33.9—Unverified
#ModelMetricClaimedVerifiedStatus
1Agent57Score580,328.14—Unverified
2QR-DQN-1Score572,510—Unverified
3R2D2Score408,850—Unverified
4IMPALA (deep)Score351,200.12—Unverified
5Ape-XScore302,391.3—Unverified
6A2C + SILScore104,975.6—Unverified
7MuZero (Res2 Adam)Score94,906.25—Unverified
8DreamerV2Score94,688—Unverified
9MuZeroScore72,276—Unverified
10DNAScore52,398—Unverified
#ModelMetricClaimedVerifiedStatus
1GDI-H3Score1,000,000—Unverified
2GDI-H3(200M frames)Score1,000,000—Unverified
3Agent57Score999,997.63—Unverified
4R2D2Score999,996.7—Unverified
5MuZeroScore999,976.52—Unverified
6MuZero (Res2 Adam)Score999,659.18—Unverified
7GDI-I3Score943,910—Unverified
8Ape-XScore392,952.3—Unverified
9C51 noopScore266,434—Unverified
10Duel noopScore50,254.2—Unverified