paper-with-me

홈 › Papers

Online No-regret Model-Based Meta RL for Personalized Navigation

2022-04-05 · Yuda Song, Ye Yuan, Wen Sun, Kris Kitani

The interaction between a vehicle navigation system and the driver of the vehicle can be formulated as a model-based reinforcement learning problem, where the navigation systems (agent) must quickly adapt to the characteristics of the driver (environmental dynamics) to provide the best sequence of turn-by-turn driving instructions. Most modern day navigation systems (e.g, Google maps, Waze, Garmin) are not designed to personalize their low-level interactions for individual users across a wide range of driving styles (e.g., vehicle type, reaction time, level of expertise). Towards the development of personalized navigation systems that adapt to a variety of driving styles, we propose an online no-regret model-based RL method that quickly conforms to the dynamics of the current user. As the user interacts with it, the navigation system quickly builds a user-specific model, from which navigation commands are optimized using model predictive control. By personalizing the policy in this way, our method is able to give well-timed driving instructions that match the user's dynamics. Our theoretical analysis shows that our method is a no-regret algorithm and we provide the convergence rate in the agnostic setting. Our empirical analysis with 60+ hours of real-world user data using a driving simulator shows that our method can reduce the number of collisions by more than 60%.

📄 PDF Abstract BibTeX arXiv:2204.01925

Code (0)

등록된 구현이 없습니다.

Tasks

Model-based Reinforcement LearningModel Predictive Control

Similar Papers 제목 키워드 기반

Meta-Learning Online Control for Linear Dynamical Systems

2022-08-18 · Deepan Muthirayan, Dileep Kalathil, Pramod P. Khargonekar

In this paper, we consider the problem of finding a meta-learning online control algorithm that can learn across the tasks when faced with a sequence of $N$ (similar) control tasks. Each task involves controlling a linea…

Meta-Learning

Meta-LinEXP3: Online-within-Online Learning for Adversarial Linear Contextual Bandits

2026-09-09 · Hao Li, Jie Xu, Zheng Xie arxiv

Meta-learning has emerged as an effective paradigm for transferring knowledge across sequential bandit tasks. While substantial progress has been made for stochastic bandits and non-contextual adversarial bandits, meta-l…

Accelerating Distributed Online Meta-Learning via Multi-Agent Collaboration under Limited Communication

2020-12-15 · Sen Lin, Mehmet Dedeoglu, Junshan Zhang

Online meta-learning is emerging as an enabling technique for achieving edge intelligence in the IoT ecosystem. Nevertheless, to learn a good meta-model for within-task fast adaptation, a single agent alone has to learn …

Meta-Learning

A Unified Framework for Analyzing Meta-algorithms in Online Convex Optimization

2024-02-13 · Mohammad Pedramfar, Vaneet Aggarwal

In this paper, we analyze the problem of online convex optimization in different settings, including different feedback types (full-information/semi-bandit/bandit/etc) in either stochastic or non-stochastic setting and d…

No-regret Non-convex Online Meta-Learning

2019-10-22 · Zhenxun Zhuang, Yunlong Wang, Kezi Yu, Songtao Lu

The online meta-learning framework is designed for the continual lifelong learning setting. It bridges two fields: meta-learning which tries to extract prior knowledge from past tasks for fast learning of future tasks, a…

Lifelong learningMeta-Learning