paper-with-me

Papers

ReCoDe: Reinforcement Learning-based Dynamic Constraint Design for Multi-Agent Coordination

2025-07-25 · Michael Amir, Guang Yang, Zhan Gao, Keisuke Okumura, Heedo Woo, Amanda Prorok arxiv

Constraint-based optimization is a cornerstone of robotics, enabling the design of controllers that reliably encode task and safety requirements such as collision avoidance or formation adherence. However, handcrafted constraints can fail in multi-agent settings that demand complex coordination. We introduce ReCoDe--Reinforcement-based Constraint Design--a decentralized, hybrid framework that merges the reliability of optimization-based controllers with the adaptability of multi-agent reinforcement learning. Rather than discarding expert controllers, ReCoDe improves them by learning additional, dynamic constraints that capture subtler behaviors, for example, by constraining agent movements to prevent congestion in cluttered scenarios. Through local communication, agents collectively constrain their allowed actions to coordinate more effectively under changing conditions. In this work, we focus on applications of ReCoDe to multi-agent navigation tasks requiring intricate, context-based movements and consensus, where we show that it outperforms purely handcrafted controllers, other hybrid approaches, and standard MARL baselines. We give empirical (real robot) and theoretical evidence that retaining a user-defined controller, even when it is imperfect, is more efficient than learning from scratch, especially because ReCoDe can dynamically change the degree to which it relies on this controller.

📄 PDF Abstract BibTeX arXiv:2507.19151

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement LearningCollision Avoidance

Similar Papers 제목 키워드 기반

Doubly-Dynamic ISAC Precoding for Vehicular Networks: A Constrained Deep Reinforcement Learning (CDRL) Approach

2024-05-23 · Zonghui Yang, Shijian Gao, Xiang Cheng

Integrated sensing and communication (ISAC) technology is essential for supporting vehicular networks. However, the communication channel in this scenario exhibits time variations, and the potential targets may move rapi…

Deep Reinforcement LearningIntegrated sensing and communicationISAC

Asymptotic Analysis of One-bit Quantized Box-Constrained Precoding in Large-Scale Multi-User Systems

2025-02-05 · Xiuxiu Ma, Abla Kammoun, Mohamed-Slim Alouini, Tareq Y. Al-Naffouri

This paper addresses the design of multi-antenna precoding strategies, considering hardware limitations such as low-resolution digital-to-analog converters (DACs), which necessitate the quantization of transmitted signal…

Quantization

ReCode: Updating Code API Knowledge with Reinforcement Learning

2025-06-25 · Haoze Wu, Yunzhi Yao, Wenhao Yu, Huajun Chen 외

Large Language Models (LLMs) exhibit remarkable code generation capabilities but falter when adapting to frequent updates in external library APIs. This critical limitation, stemming from reliance on outdated API knowled…

Code Generationreinforcement-learningReinforcement Learning

ESPRIT-Oriented Precoder Design for mmWave Channel Estimation

2023-01-04 · Musa Furkan Keskin, Alessio Fascista, Fan Jiang, Angelo Coluccia 외

We consider the problem of ESPRIT-oriented precoder design for beamspace angle-of-departure (AoD) estimation in downlink mmWave multiple-input single-output communications. Standard precoders (i.e., directional/sum beams…

PrecoderNet: Hybrid Beamforming for Millimeter Wave Systems with Deep Reinforcement Learning

2019-07-31 · Qisheng Wang, Keming Feng, Xiao Li, Shi Jin

In this letter, we investigate the hybrid beamforming for millimeter wave massive multiple-input multiple-output (MIMO) system based on deep reinforcement learning (DRL). Imperfect channel state information (CSI) is assu…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)