paper
-with-
me
Papers
Browse State-of-the-Art
Datasets
Methods
AI Agents
Trends
Digest
🌙
Atari Games
벤치마크
Atari Games on Atari 2600 Surround
30개 결과 ·
⬇ CSV
·
JSON
Score
-7.8
-3.35
1.1
5.55
10
2015-12
2026-09
Persistent AL — 0.72 (2015-12-15)
Persistent AL — 0.72 (2015-12-15)
NoisyNet-Dueling — 10.0 (2017-06-30)
NoisyNet-Dueling — 10.0 (2017-06-30)
QR-DQN-1 — 8.2 (2017-10-27)
QR-DQN-1 — 8.2 (2017-10-27)
IMPALA (deep) — 7.56 (2018-02-05)
IMPALA (deep) — 7.56 (2018-02-05)
Ape-X — 7.1 (2018-03-02)
Ape-X — 7.1 (2018-03-02)
IQN — 9.4 (2018-06-14)
IQN — 9.4 (2018-06-14)
R2D2 — 9.9 (2019-05-01)
R2D2 — 9.9 (2019-05-01)
MuZero — 9.99 (2019-11-19)
MuZero — 9.99 (2019-11-19)
Agent57 — 9.5 (2020-03-30)
Agent57 — 9.5 (2020-03-30)
MuZero (Res2 Adam) — 9.9 (2021-04-13)
MuZero (Res2 Adam) — 9.9 (2021-04-13)
GDI-I3 — -7.8 (2021-06-11)
GDI-I3 — -7.8 (2021-06-11)
GDI-H3 — 2.606 (2022-06-07)
GDI-I3 — -7.8 (2022-06-07)
GDI-H3 — 2.606 (2022-06-07)
GDI-I3 — -7.8 (2022-06-07)
DNA — 5.3 (2022-06-20)
DNA — 5.3 (2022-06-20)
ASL DDQN — 2.5 (2023-05-07)
ASL DDQN — 2.5 (2023-05-07)
Persistent AL — 0.72 (2015-12-15)
NoisyNet-Dueling — 10.0 (2017-06-30)
2015-12-15 — Persistent AL: Score 0.72
2017-06-30 — NoisyNet-Dueling: Score 10.0
Rank
Model
Score
Paper
Code
Year
1
NoisyNet-Dueling
10
Noisy Networks for Exploration
opendilab/DI-engine
·
Curt-Park/rainbow-is-all-you-need
·
chainer/chainerrl
·
+12
2017
2
MuZero
9.99
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model
werner-duvaud/muzero-general
·
opendilab/LightZero
·
koulanurag/muzero-pytorch
·
+15
2019
3
R2D2
9.9
Recurrent Experience Replay in Distributed Reinforcement Learning
opendilab/DI-engine
·
michaelnny/deep_rl_zoo
·
garymm/earl
2019
3
MuZero (Res2 Adam)
9.9
Online and Offline Reinforcement Learning by Planning with a Learned Model
DHDev0/Muzero-unplugged
·
enpasos/muzero
2021
5
Agent57
9.5
Agent57: Outperforming the Atari Human Benchmark
michaelnny/deep_rl_zoo
·
pocokhc/agent57
·
yuta0821/agent57_pytorch
·
+2
2020
6
IQN
9.4
Implicit Quantile Networks for Distributional Reinforcement Learning
opendilab/DI-engine
·
chainer/chainerrl
·
Kchu/DeepRL_CK
·
+16
2018
7
QR-DQN-1
8.2
Distributional Reinforcement Learning with Quantile Regression
DLR-RM/stable-baselines3
·
facebookresearch/Horizon
·
facebookresearch/ReAgent
·
+14
2017
8
IMPALA (deep)
7.56
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures
ray-project/ray
·
opendilab/DI-engine
·
deepmind/haiku
·
+21
2018
9
Ape-X
7.1
Distributed Prioritized Experience Replay
ray-project/ray
·
vwxyzjn/cleanrl
·
opendilab/DI-engine
·
+12
2018
10
DNA
5.3
DNA: Proximal Policy Optimization with a Dual Network Architecture
maitchison/PPO
2022
11
GDI-H3
2.606
Generalized Data Distribution Iteration
2022
12
ASL DDQN
2.5
Train a Real-world Local Path Planner in One Hour via Partially Decoupled Reinforcement Learning and Vectorized Diversity
xinjinghao/color
2023
13
Persistent AL
0.72
Increasing the Action Gap: New Operators for Reinforcement Learning
janhuenermann/neurojs
·
chainer/chainerrl
2015
14
GDI-I3
-7.8
GDI: Rethinking What Makes Reinforcement Learning Different From Supervised Learning
2021
14
GDI-I3
-7.8
Generalized Data Distribution Iteration
2022
16
NoisyNet-Dueling
10
Noisy Networks for Exploration
opendilab/DI-engine
·
Curt-Park/rainbow-is-all-you-need
·
chainer/chainerrl
·
+12
2017
17
MuZero
9.99
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model
werner-duvaud/muzero-general
·
opendilab/LightZero
·
koulanurag/muzero-pytorch
·
+15
2019
18
R2D2
9.9
Recurrent Experience Replay in Distributed Reinforcement Learning
opendilab/DI-engine
·
michaelnny/deep_rl_zoo
·
garymm/earl
2019
18
MuZero (Res2 Adam)
9.9
Online and Offline Reinforcement Learning by Planning with a Learned Model
DHDev0/Muzero-unplugged
·
enpasos/muzero
2021
20
Agent57
9.5
Agent57: Outperforming the Atari Human Benchmark
michaelnny/deep_rl_zoo
·
pocokhc/agent57
·
yuta0821/agent57_pytorch
·
+2
2020
1–20 / 30
다음 →
페이지당
10
20
50
100