paper-with-me

Papers

Towards Interpretable Visuo-Tactile Predictive Models for Soft Robot Interactions

2024-07-16 · Enrico Donato, Thomas George Thuruthel, Egidio Falotico

Autonomous systems face the intricate challenge of navigating unpredictable environments and interacting with external objects. The successful integration of robotic agents into real-world situations hinges on their perception capabilities, which involve amalgamating world models and predictive skills. Effective perception models build upon the fusion of various sensory modalities to probe the surroundings. Deep learning applied to raw sensory modalities offers a viable option. However, learning-based perceptive representations become difficult to interpret. This challenge is particularly pronounced in soft robots, where the compliance of structures and materials makes prediction even harder. Our work addresses this complexity by harnessing a generative model to construct a multi-modal perception model for soft robots and to leverage proprioceptive and visual information to anticipate and interpret contact interactions with external objects. A suite of tools to interpret the perception model is furnished, shedding light on the fusion and prediction processes across multiple sensory inputs after the learning phase. We will delve into the outlooks of the perception model and its implications for control purposes.

📄 PDF Abstract BibTeX arXiv:2407.12197

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation

2026-03-19 · Yuhang Zheng, Songen Gu, Weize Li, Yupeng Zheng 외 arxiv

Contact-rich manipulation tasks, such as wiping and assembly, require accurate perception of contact forces, friction changes, and state transitions that cannot be reliably inferred from vision alone. Despite growing int…

MuPNet: Multi-modal Predictive Coding Network for Place Recognition by Unsupervised Learning of Joint Visuo-Tactile Latent Representations

2019-09-16 · Oliver Struckmeier, Kshitij Tiwari, Shirin Dora, Martin J. Pearson 외

Extracting and binding salient information from different sensory modalities to determine common features in the environment is a significant challenge in robotics. Here we present MuPNet (Multi-modal Predictive Coding N…

RoboPack: Learning Tactile-Informed Dynamics Models for Dense Packing

2024-07-01 · Bo Ai, Stephen Tian, Haochen Shi, YiXuan Wang 외

Tactile feedback is critical for understanding the dynamics of both rigid and deformable objects in many manipulation tasks, such as non-prehensile manipulation and dense packing. We introduce an approach that combines v…

Graph Neural NetworkModel Predictive Control

TacSE3: Equivariant SE(3) Motion Estimation from Low-Texture Visuotactile Images for In-Gripper Tracking and Compensation

2026-05-18 · Zhongyuan Liao, Junzhe Wang, Qingyang Liu, Zhenmin Huang 외 arxiv

Robotic in-hand manipulation requires reliable object-motion tracking under frequent visual occlusion, yet low-texture visuotactile images provide few stable correspondences for conventional image- or geometry-matching m…

UniVTAC: A Unified Simulation Platform for Visuo-Tactile Manipulation Data Generation, Learning, and Benchmarking

2026-02-10 · Baijun Chen, Weijie Wan, Tianxing Chen, Xianda Guo 외 arxiv

Robotic manipulation has seen rapid progress with vision-language-action (VLA) policies. However, visuo-tactile perception is critical for contact-rich manipulation, as tasks such as insertion are difficult to complete r…