paper-with-me

MuJoCo Games 벤치마크

MuJoCo Games on Walker2d

2개 결과 · ⬇ CSV · JSON

RankModel Mean PaperCodeYear
1 IQ-Learn 5134 IQ-Learn: Inverse soft-Q Learning for Imitation Div99/IQ-Learn · robfiras/ls-iq · google-deepmind/csil · +2 2021
2 POP3D 3966.01 Policy Optimization With Penalized Point Probability Distance: An Alternative To Proximal Policy Optimization cxxgtxy/POP3D · paperwithcode/pop3d 2018
1–2 / 2 페이지당 10 20 50 100