paper-with-me

홈 › Papers

Scaling by Diversified Experience for Vision-Language-Action Models

2026-06-08 · Leiyu Wang, Zhaofengnian Wang, Xueqi Li, Luoyi Fan, Cewu Lu, Nanyang Ye arxiv

Vision-Language-Action models face significant challenges in real-world deployment due to the entanglement of high-level reasoning with low-level control, and the instability of policy optimization. In this paper, we introduce SyVLA, a robust VLA model trained with diversified experiences. We propose an Intention Decoupling algorithm to isolate control-relevant features from reasoning contexts and a similar-sample guided RL pipeline to stabilize policy updates and mitigate distribution shift. Extensive experiments on real-world robotic tasks and multi-modal benchmarks demonstrate that SyVLA achieves superior task success rates and stronger out-of-distribution generalization compared to existing methods, while effectively preserving core vision-language capabilities. Codes and Datasets is released on \href{https://sy-vla.github.io/}{project page}.

📄 PDF Abstract BibTeX arXiv:2606.09009

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Diversified Scaling Inference in Time Series Foundation Models

2026-01-24 · Ruijin Hua, Zichuan Liu, Kun Zhang, Yiyuan Yang arxiv

The advancement of Time Series Foundation Models (TSFMs) has been driven primarily by large-scale pre-training, but inference-time compute potential remains largely untapped. This work systematically investigates two que…

On Diversified Preferences of Large Language Model Alignment

2023-12-12 · Dun Zeng, Yong Dai, Pengyu Cheng, Longyue Wang 외

Aligning large language models (LLMs) with human preferences has been recognized as the key to improving LLMs' interaction quality. However, in this pluralistic world, human preferences can be diversified due to annotato…

Language ModelingLanguage ModellingLarge Language Model

Experience Scaling: Post-Deployment Evolution For Large Language Models

2025-09-23 · Xingkun Yin, Kaibin Huang, Dong In Kim, Hongyang Du arxiv

Scaling model size, training data, and compute power have driven advances in large language models (LLMs), but these approaches are reaching saturation as human-generated text is exhausted and further gains diminish. We …

ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory

2025-09-29 · Siru Ouyang, Jun Yan, I-Hung Hsu, Yanfei Chen 외 arxiv

With the growing adoption of large language model agents in persistent real-world roles, they naturally encounter continuous streams of tasks. A key limitation, however, is their failure to learn from the accumulated int…

Bandit Learning for Diversified Interactive Recommendation

2019-07-01 · Yong Liu, Yingtai Xiao, Qiong Wu, Chunyan Miao 외

Interactive recommender systems that enable the interactions between users and the recommender system have attracted increasing research attentions. Previous methods mainly focus on optimizing recommendation accuracy. Ho…

Bayesian InferenceDiversityInteractive RecommendationRecommendation Systems+1