paper-with-me

홈 › Papers

OpenAI Gym

2016-06-05 · Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, Wojciech Zaremba

OpenAI Gym is a toolkit for reinforcement learning research. It includes a growing collection of benchmark problems that expose a common interface, and a website where people can share their results and compare the performance of algorithms. This whitepaper discusses the components of OpenAI Gym and the design decisions that went into the software.

📄 PDF Abstract BibTeX arXiv:1606.01540

Code (45)

09jvilla/CS234_gym tf
1110589721/openia-gym tf
1nf0rmagician/gym-bimanualPickAndPlace tf
AnejSvete/gym tf
Christopheraburns/openai-gym tf
DartEnv/dart-env tf
EndingCredits/gym tf
KarlXing/gym tf
LucasSilve/openAIgym tf
M46N3/learn2drive tf
MindCode-4/code-2/tree/main/openai mindspore
Rowing0914/gym_modified tf
RyanRizzo96/gym-FetchCustom tf
StanfordVL/Gym tf
Steve--Hunter/gym tf
YanglanWang/classic_control tf
a-ozeki/openAI_gym tf
abhiksingla/gym tf
aysaha/gym tf
cbellinger27/adaptive_optics_gym pytorch
davidsonic/self_brewed_gym tf
developer-smartwebzone/-reinforcement-learning-algorithms. tf
felipeescallon/openai-gym tf
franroldans/custom_gym tf
gtrll/dartenv tf
jangirrishabh/gym tf
jturner65/Getup-DartEnv tf
ligy2016/gym tf
n0o8o0n1um/Gym1 tf
nmsquared/CS7641-Assignment-4 tf
nohboogy/gym tf
openai/gym tf
ozcell/gym_wmgds_ma tf
polixir/neorl2 pytorch
qihongl/demo-advantage-actor-critic pytorch
sahandrez/gym_forestfire pytorch
shinian123/gym tf
shriram2112/Reinforcement_learning_gym tf
sibeshkar/gym tf
skyepodium/DeepLearning-Everywhere tf
tjcdev/mlpgym tf
tsinghua-rll/gym-gridworld
varuncs2011/rl tf
vvanirudh/imitation-learning-gym tf
zubair-irshad/imitation_learning tf

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

An Empirical Study of OpenAI API Discussions on Stack Overflow

2025-05-07 · Xiang Chen, Jibin Wang, Chaoyang Gao, Xiaolin Ju 외

The rapid advancement of large language models (LLMs), represented by OpenAI's GPT series, has significantly impacted various domains such as natural language processing, software development, education, healthcare, fina…

Prompt Engineering

ORRB -- OpenAI Remote Rendering Backend

2019-06-26 · Maciek Chociej, Peter Welinder, Lilian Weng

We present the OpenAI Remote Rendering Backend (ORRB), a system that allows fast and customizable rendering of robotics environments. It is based on the Unity3d game engine and interfaces with the MuJoCo physics simulati…

MuJoCo

Dota 2 with Large Scale Deep Reinforcement Learning

2019-12-13 · Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung 외

On April 13th, 2019, OpenAI Five became the first AI system to defeat the world champions at an esports game. The game of Dota 2 presents novel challenges for AI systems such as long time horizons, imperfect information,…

Deep Reinforcement LearningDota 2reinforcement-learningReinforcement Learning+1

The ADAIO System at the BEA-2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues

2023-06-08 · Adaeze Adigwe, Zheng Yuan

This paper presents the ADAIO team's system entry in the Building Educational Applications (BEA) 2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues. The task aims to assess the performance of st…

Few-Shot LearningResponse Generation

Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit

2024-12-09 · Joshua Freeman, Chloe Rippe, Edoardo Debenedetti, Maksym Andriushchenko

Copyright infringement in frontier LLMs has received much attention recently due to the New York Times v. OpenAI lawsuit, filed in December 2023. The New York Times claims that GPT-4 has infringed its copyrights by repro…

ArticlesMemorization