MuJoCo Games 벤치마크
MuJoCo Games on Walker2d
| Rank | Model | Mean | Paper | Code | Year |
|---|---|---|---|---|---|
| 1 | IQ-Learn | 5134 | IQ-Learn: Inverse soft-Q Learning for Imitation | Div99/IQ-Learn · robfiras/ls-iq · google-deepmind/csil · +2 | 2021 |
| 2 | POP3D | 3966.01 | Policy Optimization With Penalized Point Probability Distance: An Alternative To Proximal Policy Optimization | cxxgtxy/POP3D · paperwithcode/pop3d | 2018 |