paper-with-me

홈 › Papers

Can Explicit Physical Feasibility Benefit VLA Learning? An Empirical Study

2026-04-20 · Yubai Wei, Chen Wu, Hashem Haghbayan arxiv

Vision-Language-Action (VLA) models map multimodal inputs directly to robot actions and are typically trained through large-scale imitation learning. While this paradigm has shown strong performance, prevailing VLA training procedures do not explicitly supervise hard physical constraints such as obstacle avoidance or kinematic feasibility. As a result, the geometric structure underlying physically feasible behavior must be inferred only implicitly from demonstrations. In this paper, we study whether introducing explicit feasibility supervision can provide effective structured guidance for VLA policies. We formulate a simple geometry-grounded feasibility objective and integrate it into the training stage of a diffusion-based VLA policy. To evaluate this idea systematically, we use obstacle-aware manipulation as a controlled probe of geometry-dependent physical feasibility. Empirical results show that augmenting VLA training with feasibility supervision improves both physical reliability and overall task performance, while also enhancing learning efficiency in the low-data regime. These findings indicate that explicit feasibility signals can effectively complement imitation-based VLA learning, highlighting their potential for developing more reliable VLA policies.

📄 PDF Abstract BibTeX arXiv:2604.17896

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GPLD3D: Latent Diffusion of 3D Shape Generative Models by Enforcing Geometric and Physical Priors

2024-01-01 · CVPR 2024 1 · Yuan Dong, Qi Zuo, Xiaodong Gu, Weihao Yuan 외

State-of-the-art man-made shape generative models usually adopt established generative models under a suitable implicit shape representation. A common theme is to perform distribution alignment which does not explici…

Denoising

ScenePilot: Controllable Boundary-Driven Critical Scenario Generation for Autonomous Driving

2026-05-20 · Qiyu Ruan, Yuxuan Wang, He Li, Zhenning Li 외 arxiv

Safety-critical scenarios are central to evaluating autonomous driving systems, yet their rarity in naturalistic logs makes simulation-based stress testing indispensable. Most scenario generation methods treat surroundin…

Reinforcement LearningAutonomous Driving

PhysReflect-VLA: Physical Feasibility and Self-Reflective Regulation for Reliable Vision-Language-Action Policies

2026-06-25 · Jiayu Yang, Tao Yang, Weijun Li, Xiang Chang 외 arxiv

Long-horizon robotic manipulation is highly sensitive to physically infeasible transitions, contact-induced disturbances, and the lack of effective self-correction during execution. Although Vision-Language-Action (VLA) …

CWM: Contrastive World Models for Action Feasibility Learning in Embodied Agent Pipelines

2026-02-25 · Chayan Banerjee arxiv

A reliable action feasibility scorer is a critical bottleneck in embodied agent pipelines: before any planning or reasoning occurs, the agent must identify which candidate actions are physically executable in the current…

SIPTraj: Map-Free End-to-End Trajectory Prediction via Physics-Guided Scene Interaction

2026-08-01 · Feifei Liu, Zejun Wei, Haozhe Wang, Yazhi Ye 외 arxiv

Trajectory prediction of surrounding agents is a prerequisite for safe planning and decision making in autonomous driving. Without high-definition (HD) maps, sensor-derived bird's-eye-view (BEV) features provide no expli…

Trajectory PredictionAutonomous DrivingDecision Making