paper-with-me

홈 › Papers

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

2026-07-06 · Hongyu Li, Wanjia Fu, Xiaoyan Cong, Zekun Li, Binghao Huang, Hanxiao Jiang, Xintong He, Yiqing Liang, Rao Fu, Tao Lu, Srinath Sridhar, Kevin A. Smith, George Konidaris, Yunzhu Li arxiv

Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling deformable objects presents a particularly difficult case due to their high-dimensional state spaces and complex material properties. While current world models approach this through two distinct paradigms: learning the dynamics over the 2D pixel space or more explicit 3D geometric space. A systematic understanding of their relative strengths and limitations remains elusive due to the lack of diverse, large-scale real-world data. To address this, we present Deform360, a large-scale visuotactile dataset featuring 198 daily-life objects, 1,980 interaction sequences, and over 215 hours of observations from 41 surround-view cameras and bimanual tactile grippers to capture both global motion and contact-induced local deformations. Leveraging a novel markerless visuotactile 3D tracking pipeline to extract dense geometry and motion, we systematically evaluate current state-of-the-art world models, comparing 2D video models against 3D particle models. Finally, we provide a preliminary demonstration indicating the real-world applicability of our dataset by performing robot planning tasks on deformable objects. Our analysis reveals key insights into the trade-offs between structural priors and scalability, providing a solid benchmark for future research in generalizable deformable object-centric world modeling. Project website: https://deform360.lhy.xyz

📄 PDF Abstract BibTeX arXiv:2607.05390

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Monocular Depth Estimation for Soft Visuotactile Sensors

2021-01-05 · Rares Ambrus, Vitor Guizilini, Naveen Kuppuswamy, Andrew Beaulieu 외

Fluid-filled soft visuotactile sensors such as the Soft-bubbles alleviate key challenges for robust manipulation, as they enable reliable grasps along with the ability to obtain high-resolution sensory feedback on contac…

Depth EstimationMonocular Depth Estimation

ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers

2025-05-26 · Fotios Lygerakis, Ozan Özdenizci, Elmar Rückert

Tactile sensing provides local essential information that is complementary to visual perception, such as texture, compliance, and force. Despite recent advances in visuotactile representation learning, challenges remain …

cross-modal alignmentPositionRepresentation LearningRobotic Grasping+3

Learning Visuotactile Skills with Two Multifingered Hands

2024-04-25 · Toru Lin, Yu Zhang, Qiyang Li, Haozhi Qi 외

Aiming to replicate human-like dexterity, perceptual experiences, and motion patterns, we explore learning from human demonstrations using a bimanual system with multifingered hands and visuotactile data. Two significant…

ViHOPE: Visuotactile In-Hand Object 6D Pose Estimation with Shape Completion

2023-09-11 · Hongyu Li, Snehal Dikhale, Soshi Iba, Nawid Jamali

In this letter, we introduce ViHOPE, a novel framework for estimating the 6D pose of an in-hand object using visuotactile perception. Our key insight is that the accuracy of the 6D object pose estimate can be improved by…

6D Pose EstimationGenerative Adversarial NetworkObjectPose Estimation

MoiréTac: A Dual-Mode Visuotactile Sensor for Multidimensional Perception Using Moiré Pattern Amplification

2025-09-16 · Kit-Wa Sou, Junhao Gong, Shoujie Li, Chuqiao Lyu 외 arxiv

Visuotactile sensors typically employ sparse marker arrays that limit spatial resolution and lack clear analytical force-to-image relationships. To solve this problem, we present \textbf{MoiréTac}, a dual-mode sensor tha…