paper-with-me

홈 › Papers

How Crucial is Transformer in Decision Transformer?

2022-11-26 · Max Siebenborn, Boris Belousov, Junning Huang, Jan Peters

Decision Transformer (DT) is a recently proposed architecture for Reinforcement Learning that frames the decision-making process as an auto-regressive sequence modeling problem and uses a Transformer model to predict the next action in a sequence of states, actions, and rewards. In this paper, we analyze how crucial the Transformer model is in the complete DT architecture on continuous control tasks. Namely, we replace the Transformer by an LSTM model while keeping the other parts unchanged to obtain what we call a Decision LSTM model. We compare it to DT on continuous control tasks, including pendulum swing-up and stabilization, in simulation and on physical hardware. Our experiments show that DT struggles with continuous control problems, such as inverted pendulum and Furuta pendulum stabilization. On the other hand, the proposed Decision LSTM is able to achieve expert-level performance on these tasks, in addition to learning a swing-up controller on the real system. These results suggest that the strength of the Decision Transformer for continuous control tasks may lie in the overall sequential modeling architecture and not in the Transformer per se.

📄 PDF Abstract BibTeX arXiv:2211.14655

Code (1)

max7born/decision-lstm 공식 구현 pytorch

Tasks

continuous-controlContinuous ControlDecision Making

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Tanh Activation 설명 없음
Sigmoid Activation 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Adam 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…

Similar Papers 제목 키워드 기반

Solving Multi-Goal Robotic Tasks with Decision Transformer

2024-10-08 · Paul Gajewski, Dominik Żurek, Marcin Pietroń, Kamil Faber

Artificial intelligence plays a crucial role in robotics, with reinforcement learning (RL) emerging as one of the most promising approaches for robot control. However, several key challenges hinder its broader applicatio…

Multi-Goal Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer

2025-05-14 · Minh Hoang Nguyen, Linh Le Pham Van, Thommen George Karimpanal, Sunil Gupta 외

Decision Transformers (DT) play a crucial role in modern reinforcement learning, leveraging offline datasets to achieve impressive results across various domains. However, DT requires high-quality, comprehensive data to …

counterfactualCounterfactual ReasoningD4RLDecision Making+2

Robust Load Prediction of Power Network Clusters Based on Cloud-Model-Improved Transformer

2024-07-30 · Cheng Jiang, Gang Lu, Xue Ma, Di wu

Load data from power network clusters indicates economic development in each area, crucial for predicting regional trends and guiding power enterprise decisions. The Transformer model, a leading method for load predictio…

Unlocking Bias Detection: Leveraging Transformer-Based Models for Content Analysis

2023-09-30 · Shaina Raza, Oluwanifemi Bamgbose, Veronica Chatrath, Shardul Ghuge 외

Bias detection in text is crucial for combating the spread of negative stereotypes, misinformation, and biased decision-making. Traditional language models frequently face challenges in generalizing beyond their training…

Bias DetectionDecision MakingMisinformationSentence

Decision Transformer: Reinforcement Learning via Sequence Modeling

2021-06-02 · NeurIPS 2021 12 · Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee 외

We introduce a framework that abstracts Reinforcement Learning (RL) as a sequence modeling problem. This allows us to draw upon the simplicity and scalability of the Transformer architecture, and associated advances in l…

Atari GamesD4RLLanguage ModelingLanguage Modelling+5