paper-with-me

Papers

Learning What to Say to Your VLA: Mostly Harmless Vision Language Action Model Steering

2026-06-10 · Hyun Joe Jeong, Gokul Swamy, Andrea Bajcsy arxiv

Vision-Language-Action (VLA) models provide a natural language interface to robot control, but the mapping from language to behavior is often brittle and unintuitive: semantically similar instructions can induce drastically different behaviors, while some capabilities may not be elicitable through prompting alone. As a result, both human instructions and zero-shot language models can fail to reliably steer VLAs toward successful task execution. In this work, we propose a framework that interactively searches for language sequences that improve closed-loop VLA task performance, distills these sequences into a test-time language feedback policy (LFP), and learns an improvement head that predicts when language steering will improve performance. We conformalize this improvement head to prevent harmful steering interventions, where the LFP decreases task performance relative to the original instruction on out-of-distribution scenarios. Crucially, our approach operates on arbitrary frozen pre-trained VLAs, requiring neither access to the original training distribution nor fine-tuning of the underlying model. On seen environments, our conformalized LFP improves base VLA performance by 24.7% in simulation and 65.0% in hardware. On visual and semantic perturbations, our conformalized LFP has strong harmlessness guarantees, and produces recovery behaviors not observed with open-loop prompting.

📄 PDF Abstract BibTeX arXiv:2606.12299

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Your CLIP has 164 dimensions of noise: Exploring the embeddings covariance eigenspectrum of contrastively pretrained vision-language transformers

2026-05-14 · Jakub Grzywaczewski, Dawid Płudowski, Przemysław Biecek arxiv

Contrastively pre-trained Vision-Language Models (VLMs) serve as powerful feature extractors. Yet, their shared latent spaces are prone to structural anomalies and act as repositories for non-semantic, multi-modal noise.…

Controllable Preference Optimization: Toward Controllable Multi-Objective Alignment

2024-02-29 · Yiju Guo, Ganqu Cui, Lifan Yuan, Ning Ding 외

Alignment in artificial intelligence pursues the consistency between model responses and human preferences as well as values. In practice, the multifaceted nature of human preferences inadvertently introduces what is kno…

Navigate

Predicting Personalized Academic and Career Roads: First Steps Toward a Multi-Uses Recommender System

2020-01-03 · Alexandre Nadjem, Juan-Manuel Torres-Moreno, Marc El-Bèze, Guillaume Marrel 외

Nobody knows what one's do in the future and everyone will have had a different answer to the question : how do you see yourself in five years after your current job/diploma? In this paper we introduce concepts, large ca…

Recommendation Systems

What Your Posts Reveal: A Benchmark and Agentic Framework for User-Level Privacy Leakage on Social Media

2026-06-05 · Zifan Peng, Yini Huang, Aiwen Lu, Qiming Ye 외 arxiv

Public social media posts can reveal private information through weak cues scattered across text, images, or metadata. Such leakage is often cumulative and cross-post: cues that appear harmless in isolation may jointly e…

MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?

2024-06-22 · Xirui Li, Hengguang Zhou, Ruochen Wang, Tianyi Zhou 외

Humans are prone to cognitive distortions -- biased thinking patterns that lead to exaggerated responses to specific stimuli, albeit in very different contexts. This paper demonstrates that advanced Multimodal Large Lang…

Language ModelingLanguage Modelling