paper-with-me

홈 › Papers

Is Bang-Bang Control All You Need? Solving Continuous Control with Bernoulli Policies

2021-11-03 · NeurIPS 2021 12 · Tim Seyde, Igor Gilitschenski, Wilko Schwarting, Bartolomeo Stellato, Martin Riedmiller, Markus Wulfmeier, Daniela Rus

Reinforcement learning (RL) for continuous control typically employs distributions whose support covers the entire action space. In this work, we investigate the colloquially known phenomenon that trained agents often prefer actions at the boundaries of that space. We draw theoretical connections to the emergence of bang-bang behavior in optimal control, and provide extensive empirical evaluation across a variety of recent RL algorithms. We replace the normal Gaussian by a Bernoulli distribution that solely considers the extremes along each action dimension - a bang-bang controller. Surprisingly, this achieves state-of-the-art performance on several continuous control benchmarks - in contrast to robotic hardware, where energy and maintenance cost affect controller choices. Since exploration, learning,and the final solution are entangled in RL, we provide additional imitation learning experiments to reduce the impact of exploration on our analysis. Finally, we show that our observations generalize to environments that aim to model real-world challenges and evaluate factors to mitigate the emergence of bang-bang solutions. Our findings emphasize challenges for benchmarking continuous control algorithms, particularly in light of potential real-world applications.

📄 PDF Abstract BibTeX arXiv:2111.02552

Code (0)

등록된 구현이 없습니다.

Tasks

AllBenchmarkingcontinuous-controlContinuous ControlImitation LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Solving Continuous Control via Q-learning

2022-10-22 · Tim Seyde, Peter Werner, Wilko Schwarting, Igor Gilitschenski 외

While there has been substantial success for solving continuous control with actor-critic methods, simpler critic-only methods such as Q-learning find limited application in the associated high-dimensional action spaces.…

continuous-controlContinuous ControlMulti-agent Reinforcement LearningQ-Learning

Growing Q-Networks: Solving Continuous Control Tasks with Adaptive Control Resolution

2024-04-05 · Tim Seyde, Peter Werner, Wilko Schwarting, Markus Wulfmeier 외

Recent reinforcement learning approaches have shown surprisingly strong capabilities of bang-bang policies for solving continuous control benchmarks. The underlying coarse action space discretizations often yield favoura…

continuous-controlContinuous ControlQ-Learning

BanglaSTEM: A Parallel Corpus for Technical Domain Bangla-English Translation

2025-11-05 · Kazi Reyazul Hasan, Mubasshira Musarrat, A. B. M. Alim Al Islam, Muhammad Abdullah Adnan arxiv

Large language models work well for technical problem solving in English but perform poorly when the same questions are asked in Bangla. A simple solution would be to translate Bangla questions into English first and the…

Vashantor: A Large-scale Multilingual Benchmark Dataset for Automated Translation of Bangla Regional Dialects to Bangla Language

2023-11-18 · Fatema Tuj Johora Faria, Mukaffi Bin Moin, Ahmed Al Wase, Mehidi Ahmmed 외

The Bangla linguistic variety is a fascinating mix of regional dialects that adds to the cultural diversity of the Bangla-speaking community. Despite extensive study into translating Bangla to English, English to Bangla,…

Machine TranslationTranslation

MultiBanAbs: A Comprehensive Multi-Domain Bangla Abstractive Text Summarization Dataset

2025-11-24 · Md. Tanzim Ferdous, Naeem Ahsan Chowdhury, Prithwiraj Bhattacharjee arxiv

This study developed a new Bangla abstractive summarization dataset to generate concise summaries of Bangla articles from diverse sources. Most existing studies in this field have concentrated on news articles, where jou…

Abstractive Text SummarizationTransfer Learning