paper-with-me

Papers

ICPRL: Acquiring Physical Intuition from Interactive Control

2026-03-01 · Xinrun Xu, Pi Bu, Ye Wang, Börje F. Karlsson, Ziming Wang, Tengtao Song, Qi Zhu, Jun Song, Shuo Zhang, Zhiming Ding, Bo Zheng arxiv

VLMs excel at static perception but falter in interactive reasoning in dynamic physical environments, which demands planning and adaptation to dynamic outcomes. Existing physical reasoning methods often depend on abstract symbolic inputs or lack the ability to learn and adapt from direct, pixel-based visual interaction in novel scenarios. We introduce ICPRL (In-Context Physical Reinforcement Learning), a framework inspired by In-Context Reinforcement Learning (ICRL) that empowers VLMs to acquire physical intuition and adapt their policies in-context. Our approach trains a vision-grounded policy model via multi-turn Group Relative Policy Optimization (GRPO) over diverse multi-episode interaction histories. This enables the agent to adapt strategies by conditioning on past trial-and-error sequences, without requiring any weight updates. This adaptive policy works in concert with a separately trained world model that provides explicit physical reasoning by predicting the results of potential actions. At inference, the policy proposes candidate actions, while the world model predicts outcomes to guide a root-node PUCT search to select the most promising action. Evaluated on the diverse physics-based puzzle-solving tasks in the DeepPHY benchmark, ICPRL demonstrates significant improvements across both its I. policy-only, and II. world-model-augmented stages. Notably, these gains are retained in unseen physical environments, demonstrating that our framework facilitates genuine in-context acquisition of the environment's physical dynamics from interactive experience.

📄 PDF Abstract BibTeX arXiv:2603.13295

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningPhysical Intuition

Similar Papers 제목 키워드 기반

INTENTION: Inferring Tendencies of Humanoid Robot Motion Through Interactive Intuition and Grounded VLM

2025-08-06 · Jin Wang, Weijie Wang, Boyuan Deng, Heng Zhang 외 arxiv

Traditional control and planning for robotic manipulation heavily rely on precise physical models and predefined action sequences. While effective in structured environments, such approaches often fail in real-world scen…

Acquiring Grounded Representations of Words with Situated Interactive Instruction

2025-02-28 · Shiwali Mohan, Aaron H. Mininger, James R. Kirk, John E. Laird

We present an approach for acquiring grounded representations of words from mixed-initiative, situated interactions with a human instructor. The work focuses on the acquisition of diverse types of knowledge including per…

Learning in Context, Guided by Choice: A Reward-Free Paradigm for Reinforcement Learning with Transformers

2026-02-09 · Juncheng Dong, Bowen He, Moyang Guo, Ethan X. Fang 외 arxiv

In-context reinforcement learning (ICRL) leverages the in-context learning capabilities of transformer models (TMs) to efficiently generalize to unseen sequential decision-making tasks without parameter updates. However,…

Reinforcement LearningContinuous Control

3D Reconstruction of Crime Scenes and Design Considerations for an Interactive Investigation Tool

2015-12-10 · Erkan Bostanci

Crime Scene Investigation (CSI) is a carefully planned systematic process with the purpose of acquiring physical evidences to shed light upon the physical reality of the crime and eventually detect the identity of the cr…

3D Reconstruction

Acquiring Target Stacking Skills by Goal-Parameterized Deep Reinforcement Learning

2017-11-01 · ICLR 2018 1 · Wenbin Li, Jeannette Bohg, Mario Fritz

Understanding physical phenomena is a key component of human intelligence and enables physical interaction with previously unseen environments. In this paper, we study how an artificial agent can autonomously acquire thi…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)