paper-with-me

Papers

TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training

2026-07-03 · Shengbang Liu, Yueru Jia, Yuyang Yan, Jiaming Liu, Xinran Zhang, Qiuxuan Feng, Yandong Guo, Shiji Zhou, Boxin Shi, Shanghang Zhang arxiv

Vision-Language-Action (VLA) models have shown promising generalization in robotic manipulation, but they still struggle with contact-rich tasks, where minor contact perturbations can cause unrecoverable failures that are hard to detect from vision alone. Since these failures are localized rather than task-level semantic errors, tactile-aware corrective post-training offers an efficient way to improve recovery. However, scaling such supervision through human intervention is costly. Recent works have explored world models to synthesize imagined rollouts for policy improvement, but vision-only world models may produce visually plausible yet contact-inconsistent trajectories. We therefore introduce TACO, a tactile-aware world-model-driven framework for scalable VLA post-training in contact-rich manipulation. Given real robot rollouts, TACO follows a Recognize-Imagine-Label loop with a tactile-aware world model: a unified progress-action model recognizes failure-adjacent states using progress estimates, a visuo-tactile generation model imagines local correction segments, and the progress-action model labels them with executable corrective actions. To incorporate tactile corrective supervision into VLA post-training, TACO combines knowledge-insulated tactile adaptation with advantage-conditioned training, enabling the policy to learn from imagined corrections without degrading pretrained visual-language priors. These components enable TACO to convert real-world failures into imagined visuo-tactile corrections for iterative VLA post-training. Experiments on real-world contact-rich manipulation tasks show that TACO achieves 44% absolute success rate improvement over the base policy and 32% over the policy without knowledge-insulated tactile adaptation.

📄 PDF Abstract BibTeX arXiv:2607.02840

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Visual-Tactile Sensing for In-Hand Object Reconstruction

2023-03-25 · CVPR 2023 1 · Wenqiang Xu, Zhenjun Yu, Han Xue, Ruolin Ye 외

Tactile sensing is one of the modalities humans rely on heavily to perceive the world. Working with vision, this modality refines local geometry structure, measures deformation at the contact area, and indicates the hand…

ObjectObject Reconstruction

TaCo: A Benchmark for Lossless and Lossy Codecs of Heterogeneous Tactile Data

2026-02-10 · Zhengxue Cheng, Yan Zhao, Keyu Wang, Hengdi Zhang 외 arxiv

Tactile sensing is crucial for embodied intelligence, providing fine-grained perception and control in complex environments. However, efficient tactile data compression, which is essential for real-time robotic applicati…

Robotic Grasping

TacO: Benchmarking Tactile Sensors for Object Manipulation

2026-05-21 · Anya Zorin, Zilin Si, Myungsun Park, Junsung Park 외 arxiv

Vision-based learning from demonstrations has achieved remarkable success in enabling robots to perform manipulation tasks and high-level semantic reasoning, yet it remains insufficient for complex, contact-rich manipula…

Robot Manipulation

Visuotactile Affordances for Cloth Manipulation with Local Control

2022-12-09 · Neha Sunil, Shaoxiong Wang, Yu She, Edward Adelson 외

Cloth in the real world is often crumpled, self-occluded, or folded in on itself such that key regions, such as corners, are not directly graspable, making manipulation difficult. We propose a system that leverages visua…

Edge ClassificationPose Estimation

Metacognition for Unknown Situations and Environments (MUSE)

2024-11-20 · Rodolfo Valiente, Praveen K. Pilly

Metacognition--the awareness and regulation of one's cognitive processes--is central to human adaptability in unknown situations. In contrast, current autonomous agents often struggle in novel environments due to their l…