paper-with-me

홈 › Papers

Differentiable Environment-Trajectory Co-Optimization for Safe Multi-Agent Navigation

2026-04-08 · Zhan Gao, Gabriele Fadini, Stelian Coros, Amanda Prorok arxiv

The environment plays a critical role in multi-agent navigation by imposing spatial constraints, rules, and limitations that agents must navigate around. Traditional approaches treat the environment as fixed, without exploring its impact on agents' performance. This work considers environment configurations as decision variables, alongside agent actions, to jointly achieve safe navigation. We formulate a bi-level problem, where the lower-level sub-problem optimizes agent trajectories that minimize navigation cost and the upper-level sub-problem optimizes environment configurations that maximize navigation safety. We develop a differentiable optimization method that iteratively solves the lower-level sub-problem with interior point methods and the upper-level sub-problem with gradient ascent. A key challenge lies in analytically coupling these two levels. We address this by leveraging KKT conditions and the Implicit Function Theorem to compute gradients of agent trajectories w.r.t. environment parameters, enabling differentiation throughout the bi-level structure. Moreover, we propose a novel metric that quantifies navigation safety as a criterion for the upper-level environment optimization, and prove its validity through measure theory. Our experiments validate the effectiveness of the proposed framework in a variety of safety-critical navigation scenarios, inspired from warehouse logistics to urban transportation. The results demonstrate that optimized environments provide navigation guidance, improving both agents' safety and efficiency.

📄 PDF Abstract BibTeX arXiv:2604.06972

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DiffCo: Auto-Differentiable Proxy Collision Detection with Multi-class Labels for Safety-Aware Trajectory Optimization

2021-02-15 · Yuheng Zhi, Nikhil Das, Michael Yip

The objective of trajectory optimization algorithms is to achieve an optimal collision-free path between a start and goal state. In real-world scenarios where environments can be complex and non-homogeneous, a robot need…

pdSTL: Probabilistic Differentiable Signal Temporal Logic for Stochastic Systems

2026-06-17 · Bennett Dogbey, Hemanth Manjunatha arxiv

Autonomous robots operating in uncertain environments must satisfy complex temporal and safety specifications despite stochastic dynamics and sensing noise. While Signal Temporal Logic (STL) offers robustness measures fo…

LeTO: Learning Constrained Visuomotor Policy with Differentiable Trajectory Optimization

2024-01-30 · Zhengtong Xu, Yu She

This paper introduces LeTO, a method for learning constrained visuomotor policy with differentiable trajectory optimization. Our approach integrates a differentiable optimization layer into the neural network. By formula…

Imitation Learning

CSSDF-Net: Safe Motion Planning Based on Neural Implicit Representations of Configuration Space Distance Field

2026-03-19 · Haohua Chen, Yixuan Zhou, Yifan Zhou, Hesheng Wang arxiv

High-dimensional manipulator operation in unstructured environments requires a differentiable, scene-agnostic distance query mechanism to guide safe motion generation. Existing geometric collision checkers are typically …

Zero-shot GeneralizationCollision AvoidanceMotion Planning

Safe Reinforcement Learning Using Black-Box Reachability Analysis

2022-04-15 · Mahmoud Selim, Amr Alanwar, Shreyas Kousik, Grace Gao 외

Reinforcement learning (RL) is capable of sophisticated motion planning and control for robots in uncertain environments. However, state-of-the-art deep RL approaches typically lack safety guarantees, especially when the…

Motion Planningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1