paper-with-me

홈 › Papers

FAWAM: Force-Aware World Action Models for Closed-Loop Contact-Rich Manipulation

2026-06-07 · Haotian He, Zeyu Yan, Qipeng Liu, Ning Guo, Wenzhao Lian arxiv

Force signals provide critical interaction cues for contact-rich robotic manipulation. However, existing methods mostly use force as an additional observation modality, without fully exploiting its role in modeling future interaction dynamics or guiding execution-time feedback correction. In this paper, we propose FAWAM, a force-aware world action model that incorporates force information at three levels: perception, prediction, and closed-loop execution. FAWAM first encodes historical 6-axis force/torque signals to modulate action generation, then jointly predicts future actions and end-effector wrenches to explicitly model contact evolution. It further introduces a residual correction module that uses the predicted wrench trajectory as an execution-time reference to refine actions online based on real-time force feedback. Real-world experiments across multiple contact-rich tasks show that FAWAM improves the average success rate by 36.25% over vision-only baselines and 21.25% over existing force-aware baselines, demonstrating the effectiveness of our force-aware framework for robust contact-rich manipulation.

📄 PDF Abstract BibTeX arXiv:2606.08555

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ForceVLA2: Unleashing Hybrid Force-Position Control with Force Awareness for Contact-Rich Manipulation

2026-03-16 · Yang Li, Zhaxizhuoma, Hongru Jiang, Junjie Xia 외 arxiv

Embodied intelligence for contact-rich manipulation has predominantly relied on position control, while explicit awareness and regulation of interaction forces remain under-explored, limiting stability, precision, and ro…

HaWMPO: Hallucination-Aware World Model-based Policy Optimization for Generalist Robot Policy

2026-09-09 · Zengjue Chen, Peidong Liu, Jiawei Li, Qi Wang arxiv

Generalist robot policies have demonstrated strong generalization across robotic manipulation tasks, yet their success rates remain limited in com- plex long-horizon scenarios. Recent methods improve Visual-Language-Acti…

Reinforcement Learning

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation

2026-05-15 · Baining Zhao, Jiacheng Xu, Weicheng Feng, Xin Zhang 외 arxiv

Aerial vision-language navigation (VLN) requires agents to follow natural-language instructions through closed-loop perception and action in 3D environments. We argue that aerial VLN can be formulated as a prediction-dri…

Vision-Language NavigationReinforcement Learning

PulseCX: Breaking the Closed-World Assumption in Real-Time CX

2026-06-19 · Rajat Agarwal, Suvidha Tripathi, Shubham Sharma arxiv

Conversational AI agents in Customer Experience (CX) typically suffer from a Closed-World Constraint, ignoring high-velocity external shifts like viral trends or outages. Ad-hoc web search attempts to bridge this gap but…

Occlusion-Aware Search for Object Retrieval in Clutter

2020-11-06 · Wissam Bejjani, Wisdom C. Agboh, Mehmet R. Dogar, Matteo Leonetti

We address the manipulation task of retrieving a target object from a cluttered shelf. When the target object is hidden, the robot must search through the clutter for retrieving it. Solving this task requires reasoning o…

ObjectRetrieval