paper-with-me

홈 › Papers

Physics Enhanced Residual Policy Learning (PERPL) for safety cruising in mixed traffic platooning under actuator and communication delay

2024-09-23 · Keke Long, Haotian Shi, Yang Zhou, Xiaopeng Li

Linear control models have gained extensive application in vehicle control due to their simplicity, ease of use, and support for stability analysis. However, these models lack adaptability to the changing environment and multi-objective settings. Reinforcement learning (RL) models, on the other hand, offer adaptability but suffer from a lack of interpretability and generalization capabilities. This paper aims to develop a family of RL-based controllers enhanced by physics-informed policies, leveraging the advantages of both physics-based models (data-efficient and interpretable) and RL methods (flexible to multiple objectives and fast computing). We propose the Physics-Enhanced Residual Policy Learning (PERPL) framework, where the physics component provides model interpretability and stability. The learning-based Residual Policy adjusts the physics-based policy to adapt to the changing environment, thereby refining the decisions of the physics model. We apply our proposed model to decentralized control to mixed traffic platoon of Connected and Automated Vehicles (CAVs) and Human-driven Vehicles (HVs) using a constant time gap (CTG) strategy for cruising and incorporating actuator and communication delays. Experimental results demonstrate that our method achieves smaller headway errors and better oscillation dampening than linear models and RL alone in scenarios with artificially extreme conditions and real preceding vehicle trajectories. At the macroscopic level, overall traffic oscillations are also reduced as the penetration rate of CAVs employing the PERPL scheme increases.

📄 PDF Abstract BibTeX arXiv:2409.15595

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Physics-Regulated Deep Reinforcement Learning: Invariant Embeddings

2023-05-26 · Hongpeng Cao, Yanbing Mao, Lui Sha, Marco Caccamo

This paper proposes the Phy-DRL: a physics-regulated deep reinforcement learning (DRL) framework for safety-critical autonomous systems. The Phy-DRL has three distinguished invariant-embedding designs: i) residual action…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning

Physical Deep Reinforcement Learning Towards Safety Guarantee

2023-03-29 · Hongpeng Cao, Yanbing Mao, Lui Sha, Marco Caccamo

Deep reinforcement learning (DRL) has achieved tremendous success in many complex decision-making tasks of autonomous systems with high-dimensional state and/or action spaces. However, the safety and stability still rema…

Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement Learning

HetGPS: Scalable Graph Multi-Agent Reinforcement Learning with Physics-Anchored Adaptive Safety for EV Charging

2026-08-01 · Xiangwei Wang, Nanduni Nimalsiri, Yu Xia, Peng Wang 외 arxiv

Safety interventions for large populations of network-coupled agents must protect shared constraints without unnecessarily overriding task-oriented policy decisions. We present HetGPS, a hybrid graph-control framework sy…

Multi-agent Reinforcement Learning

Trustworthy Human-AI Collaboration: Reinforcement Learning with Human Feedback and Physics Knowledge for Safe Autonomous Driving

2024-09-01 · Zilin Huang, Zihao Sheng, Sikai Chen

In the field of autonomous driving, developing safe and trustworthy autonomous driving policies remains a significant challenge. Recently, Reinforcement Learning with Human Feedback (RLHF) has attracted substantial atten…

Autonomous DrivingPhilosophyreinforcement-learningReinforcement Learning

A Physics Enhanced Residual Learning (PERL) Framework for Vehicle Trajectory Prediction

2023-09-26 · Keke Long, Zihao Sheng, Haotian Shi, Xiaopeng Li 외

In vehicle trajectory prediction, physics models and data-driven models are two predominant methodologies. However, each approach presents its own set of challenges: physics models fall short in predictability, while dat…

PredictionTrajectory Prediction