paper-with-me

Atari Games 벤치마크

Atari Games on Atari 2600 Surround

30개 결과 · ⬇ CSV · JSON

Score

-7.8 -3.35 1.1 5.55 10 2015-12 2026-09 Persistent AL — 0.72 (2015-12-15) Persistent AL — 0.72 (2015-12-15) NoisyNet-Dueling — 10.0 (2017-06-30) NoisyNet-Dueling — 10.0 (2017-06-30) QR-DQN-1 — 8.2 (2017-10-27) QR-DQN-1 — 8.2 (2017-10-27) IMPALA (deep) — 7.56 (2018-02-05) IMPALA (deep) — 7.56 (2018-02-05) Ape-X — 7.1 (2018-03-02) Ape-X — 7.1 (2018-03-02) IQN — 9.4 (2018-06-14) IQN — 9.4 (2018-06-14) R2D2 — 9.9 (2019-05-01) R2D2 — 9.9 (2019-05-01) MuZero — 9.99 (2019-11-19) MuZero — 9.99 (2019-11-19) Agent57 — 9.5 (2020-03-30) Agent57 — 9.5 (2020-03-30) MuZero (Res2 Adam) — 9.9 (2021-04-13) MuZero (Res2 Adam) — 9.9 (2021-04-13) GDI-I3 — -7.8 (2021-06-11) GDI-I3 — -7.8 (2021-06-11) GDI-H3 — 2.606 (2022-06-07) GDI-I3 — -7.8 (2022-06-07) GDI-H3 — 2.606 (2022-06-07) GDI-I3 — -7.8 (2022-06-07) DNA — 5.3 (2022-06-20) DNA — 5.3 (2022-06-20) ASL DDQN — 2.5 (2023-05-07) ASL DDQN — 2.5 (2023-05-07) Persistent AL — 0.72 (2015-12-15) NoisyNet-Dueling — 10.0 (2017-06-30)
RankModel Score PaperCodeYear
1 NoisyNet-Dueling 10 Noisy Networks for Exploration opendilab/DI-engine · Curt-Park/rainbow-is-all-you-need · chainer/chainerrl · +12 2017
2 MuZero 9.99 Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model werner-duvaud/muzero-general · opendilab/LightZero · koulanurag/muzero-pytorch · +15 2019
3 R2D2 9.9 Recurrent Experience Replay in Distributed Reinforcement Learning opendilab/DI-engine · michaelnny/deep_rl_zoo · garymm/earl 2019
3 MuZero (Res2 Adam) 9.9 Online and Offline Reinforcement Learning by Planning with a Learned Model DHDev0/Muzero-unplugged · enpasos/muzero 2021
5 Agent57 9.5 Agent57: Outperforming the Atari Human Benchmark michaelnny/deep_rl_zoo · pocokhc/agent57 · yuta0821/agent57_pytorch · +2 2020
6 IQN 9.4 Implicit Quantile Networks for Distributional Reinforcement Learning opendilab/DI-engine · chainer/chainerrl · Kchu/DeepRL_CK · +16 2018
7 QR-DQN-1 8.2 Distributional Reinforcement Learning with Quantile Regression DLR-RM/stable-baselines3 · facebookresearch/Horizon · facebookresearch/ReAgent · +14 2017
8 IMPALA (deep) 7.56 IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures ray-project/ray · opendilab/DI-engine · deepmind/haiku · +21 2018
9 Ape-X 7.1 Distributed Prioritized Experience Replay ray-project/ray · vwxyzjn/cleanrl · opendilab/DI-engine · +12 2018
10 DNA 5.3 DNA: Proximal Policy Optimization with a Dual Network Architecture maitchison/PPO 2022
11 GDI-H3 2.606 Generalized Data Distribution Iteration 2022
12 ASL DDQN 2.5 Train a Real-world Local Path Planner in One Hour via Partially Decoupled Reinforcement Learning and Vectorized Diversity xinjinghao/color 2023
13 Persistent AL 0.72 Increasing the Action Gap: New Operators for Reinforcement Learning janhuenermann/neurojs · chainer/chainerrl 2015
14 GDI-I3 -7.8 GDI: Rethinking What Makes Reinforcement Learning Different From Supervised Learning 2021
14 GDI-I3 -7.8 Generalized Data Distribution Iteration 2022
16 NoisyNet-Dueling 10 Noisy Networks for Exploration opendilab/DI-engine · Curt-Park/rainbow-is-all-you-need · chainer/chainerrl · +12 2017
17 MuZero 9.99 Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model werner-duvaud/muzero-general · opendilab/LightZero · koulanurag/muzero-pytorch · +15 2019
18 R2D2 9.9 Recurrent Experience Replay in Distributed Reinforcement Learning opendilab/DI-engine · michaelnny/deep_rl_zoo · garymm/earl 2019
18 MuZero (Res2 Adam) 9.9 Online and Offline Reinforcement Learning by Planning with a Learned Model DHDev0/Muzero-unplugged · enpasos/muzero 2021
20 Agent57 9.5 Agent57: Outperforming the Atari Human Benchmark michaelnny/deep_rl_zoo · pocokhc/agent57 · yuta0821/agent57_pytorch · +2 2020
1–20 / 30 다음 → 페이지당 10 20 50 100