paper-with-me

홈 › Papers

Latent Actions from Factorized Transition Effects under Agent Ambiguity

2026-06-29 · Heejeong Nam, Chandradithya S Jonnalagadda, Harshit Aggarwal, Eric Xu, Randall Balestriero arxiv

Latent Action Models (LAMs) learn action-like proxies from observation transitions. However, in multi-object or distractor-rich scenes, these visual effects mix agent motion with distractors, camera dynamics, and background changes, making the underlying action source ambiguous without supervision. Structuring this mixture as reusable transition effects provides an intermediate representation from which action-like latents can be more robustly formed. We introduce Observed Transition Factorization (OTF), which decomposes each transition into a sparse set of observed transition primitives. Using these primitives as the transition interface, we propose OTF-LAM, which abstracts motion primitives into action-like latents within the standard inverse-forward dynamics framework, and OTF-LAM-Dino, a decoder-free variant that predicts future states in a frozen DINOv2 representation space. Empirically, OTF primitives transfer zeroshot across controlled carrier and morphology shifts, showing reusability. Furthermore, downstream policy learning results match or outperform baselines under complex transition ambiguity.

📄 PDF Abstract BibTeX arXiv:2606.30544

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MALA: Cross-Domain Dialogue Generation with Action Learning

2019-12-18 · Xinting Huang, Jianzhong Qi, Yu Sun, Rui Zhang

Response generation for task-oriented dialogues involves two basic components: dialogue planning and surface realization. These two components, however, have a discrepancy in their objectives, i.e., task completion and l…

Dialogue GenerationResponse GenerationSemantic SimilaritySemantic Textual Similarity+1

ConTraIRL: Factorized Contrastive Abstractions for Transferable IRL

2026-06-02 · Yikang Gui, Bikramjit Banerjee, Prashant Doshi arxiv

Reward transfer in Inverse Reinforcement Learning (IRL) is unreliable when policies must generalize to unseen combinations of environment dynamics and task goals. We propose Factorized Contrastive Abstractions for Transf…

Reinforcement LearningContinuous Control

Characterizing A Database of Sequential Behaviors with Latent Dirichlet Hidden Markov Models

2013-05-24 · Yin Song, Longbing Cao, Xuhui Fan, Wei Cao 외

This paper proposes a generative model, the latent Dirichlet hidden Markov models (LDHMM), for characterizing a database of sequential behaviors (sequences). LDHMMs posit that each sequence is generated by an underlying …

General Classification

Factored Latent Action World Models

2026-02-18 · Zizhao Wang, Chang Shi, Jiaheng Hu, Kevin Rohling 외 arxiv

Learning latent actions from action-free video has emerged as a powerful paradigm for scaling up controllable world model learning. Latent actions provide a natural interface for users to iteratively generate and manipul…

Video Generation

SPECTRA: Sparse Entity-centric Transitions

2019-09-25 · Rim Assouel, Yoshua Bengio

Learning an agent that interacts with objects is ubiquituous in many RL tasks. In most of them the agent's actions have sparse effects : only a small subset of objects in the visual scene will be affected by the action …