paper-with-me

홈 › Papers

DIAL: Distribution-Informed Adaptive Learning of Multi-Task Constraints for Safety-Critical Systems

2025-01-30 · Se-Wook Yoo, Seung-Woo Seo

Safe reinforcement learning has traditionally relied on predefined constraint functions to ensure safety in complex real-world tasks, such as autonomous driving. However, defining these functions accurately for varied tasks is a persistent challenge. Recent research highlights the potential of leveraging pre-acquired task-agnostic knowledge to enhance both safety and sample efficiency in related tasks. Building on this insight, we propose a novel method to learn shared constraint distributions across multiple tasks. Our approach identifies the shared constraints through imitation learning and then adapts to new tasks by adjusting risk levels within these learned distributions. This adaptability addresses variations in risk sensitivity stemming from expert-specific biases, ensuring consistent adherence to general safety principles even with imperfect demonstrations. Our method can be applied to control and navigation domains, including multi-task and meta-task scenarios, accommodating constraints such as maintaining safe distances or adhering to speed limits. Experimental results validate the efficacy of our approach, demonstrating superior safety performance and success rates compared to baselines, all without requiring task-specific constraint definitions. These findings underscore the versatility and practicality of our method across a wide range of real-world tasks.

📄 PDF Abstract BibTeX arXiv:2501.18086

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingImitation LearningSafe Reinforcement Learning

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Adaptively trained Physics-informed Radial Basis Function Neural Networks for Solving Multi-asset Option Pricing Problems

2026-01-19 · Yan Ma, Yumeng Ren, Elisabeth Larsson arxiv

The present study investigates the numerical solution of Black-Scholes partial differential equation (PDE) for option valuation with multiple underlying assets. We develop a physics-informed (PI) machine learning algorit…

LAVA: Latent Action Spaces via Variational Auto-encoding for Dialogue Policy Optimization

2020-11-18 · COLING 2020 8 · Nurul Lubis, Christian Geishauser, Michael Heck, Hsien-Chin Lin 외

Reinforcement learning (RL) can enable task-oriented dialogue systems to steer the conversation towards successful task completion. In an end-to-end setting, a response can be constructed in a word-level sequential decis…

Decision MakingReinforcement Learning (RL)Sequential Decision MakingTask-Oriented Dialogue Systems

Kernel-Adaptive PI-ELMs for Forward and Inverse Problems in PDEs with Sharp Gradients

2025-07-14 · Vikas Dwivedi, Balaji Srinivasan, Monica Sigovan, Bruno Sixou arxiv

Physics-informed machine learning frameworks such as Physics-Informed Neural Networks (PINNs) and Physics-Informed Extreme Learning Machines (PI-ELMs) have shown great promise for solving partial differential equations (…

Dialogue Act Patterns in GenAI-Mediated L2 Oral Practice: A Sequential Analysis of Learner-Chatbot Interactions

2026-04-07 · Liqun He, Shijun, Chen, Mutlu Cukurova 외 arxiv

While generative AI (GenAI) voice chatbots offer scalable opportunities for second language (L2) oral practice, the interactional processes related to learners' gains remain underexplored. This study investigates dialogu…

DialogGraph-LLM: Graph-Informed LLMs for End-to-End Audio Dialogue Intent Recognition

2025-11-14 · HongYu Liu, Junxin Li, Changxi Guo, Hao Chen 외 arxiv

Recognizing speaker intent in long audio dialogues among speakers has a wide range of applications, but is a non-trivial AI task due to complex inter-dependencies in speaker utterances and scarce annotated data. To addre…

Intent Recognition