paper-with-me

홈 › Papers

Distilling Collaborative Dynamics into Latent Space for Implicit Coordination in Decentralized Multi-Agent Manipulation

2026-06-22 · Chanyoung Park, Minsung Yoon, Andrew Jeong, Sung-eui Yoon arxiv

Multi-arm manipulation demands precise spatiotemporal coordination, yet many centralized approaches scale poorly as team size increases. To address this, we propose CLS-DP, a decentralized multi-agent framework that enables implicit coordination under partial observability without shared global views, explicit state information, or inter-agent communication. Under the centralized training and decentralized execution (CTDE) paradigm, CLS-DP distills privileged multi-agent dynamics into a latent space. At deployment, each agent infers a collaborative latent from its local RGB observation and a shared task instruction; it then conditions the diffusion denoising process on this latent. This design enables implicit coordination with a per-agent cost independent of team size. Across six RoboFactory benchmark tasks spanning two to four agents, CLS-DP achieves a 38% mean success rate, outperforming the best centralized baseline (20%) and a decentralized ablation without the collaborative latent (9%). It also maintains superior parameter efficiency across all agent configurations. Attribution maps show that an agent conditioned on the collaborative latent places high attribution on the joints and grippers of both itself and its teammates throughout execution. This suggests that the learned latent efficiently encodes collaborative dynamics from local observation, which facilitates implicit coordination in realistic settings characterized by partial observability.

📄 PDF Abstract BibTeX arXiv:2606.22982

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Learning to Watermark in the Latent Space of Generative Models

2026-01-22 · Sylvestre-Alvise Rebuffi, Tuan Tran, Valeriu Lacatusu, Pierre Fernandez 외 arxiv

Existing approaches for watermarking AI-generated images often rely on post-hoc methods applied in pixel space, introducing computational overhead and potential visual artifacts. In this work, we explore latent space wat…

Expressive dynamics models with nonlinear injective readouts enable reliable recovery of latent features from neural activity

2023-09-12 · Christopher Versteeg, Andrew R. Sedler, Jonathan D. McCart, Chethan Pandarinath

The advent of large-scale neural recordings has enabled new methods to discover the computational mechanisms of neural circuits by understanding the rules that govern how their state evolves over time. While these \texti…

Wasserstein Auto-encoded MDPs: Formal Verification of Efficiently Distilled RL Policies with Many-sided Guarantees

2023-03-22 · Florent Delgrange, Ann Nowé, Guillermo A. Pérez

Although deep reinforcement learning (DRL) has many success stories, the large-scale deployment of policies learned through these advanced techniques in safety-critical scenarios is hindered by their lack of formal guara…

Deep Reinforcement Learning

The Gaussian Process Prior VAE for Interpretable Latent Dynamics from Pixels

2019-10-16 · pproximateinference AABI Symposium 2019 12 · Michael Arthur Leopold Pearce

We consider the problem of unsupervised learning of a low dimensional, interpretable, latent state of a video containing a moving object. The problem of distilling dynamics from pixels has been extensively considered thr…

State Space Models

System-1.5 Reasoning: Traversal in Language and Latent Spaces with Dynamic Shortcuts

2025-05-25 · Xiaoqiang Wang, Suyuchen Wang, Yun Zhu, Bang Liu

Chain-of-thought (CoT) reasoning enables large language models (LLMs) to move beyond fast System-1 responses and engage in deliberative System-2 reasoning. However, this comes at the cost of significant inefficiency due …

GSM8K