paper-with-me

Papers

ChainVLA: Chaining Vision-Language-Action Queries through a Unified Execution State for Long-Horizon Manipulation

2026-08-03 · Yuzhi Huang, Weijue Bu, Ziyi Xiong, Jie Wu, Fanding Huang, Jingyan Jiang, Zhi Wang arxiv

Humans perform long-horizon manipulation by retaining knowledge of what earlier actions have established while continuously adapting the motion underway. By contrast, action-chunked vision-language-action (VLA) policies repeatedly replan from the current input at each query. Existing methods preserve either long-term task evidence through memory or short-term motion through action reuse and ensembling, leaving the cross-query handoff incomplete. We introduce ChainVLA, a 1.2B-parameter VLA policy that chains successive queries through a joint and revisable execution state. Progress Context combines a recurrent Working State with sparse event memory to carry observation-derived task progress, while Motion Tail feeds the preceding prediction's unexecuted continuation into state construction and action generation. Together, the two components condition a decoder that regenerates each action horizon under the latest observation, allowing the carried state to guide the next prediction without fixing it. ChainVLA reaches 62.8% average success on RMBench and 98.8% across four LIBERO suites, while removing Motion Tail or Progress Context reduces RMBench success to 11.2% and 3.0%, respectively. These asymmetric ablations are consistent with motion continuity helping preserve the observation stream from which task progress is inferred.

📄 PDF Abstract BibTeX arXiv:2608.02326

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

An Agent-based Model of the Cognitive Mechanisms Underlying the Origins of Creative Cultural Evolution

2013-10-14 · Liane Gabora, Maryam Saberi

Human culture is uniquely cumulative and open-ended. Using a computational model of cultural evolution in which neural network based agents evolve ideas for actions through invention and imitation, we tested the hypothes…

Cultural Vocal Bursts Intensity PredictionDiversity

RE-GAINS & EnChAnT: Intelligent Tool Manipulation Systems For Enhanced Query Responses

2024-01-28 · Sahil Girhepuje, Siva Sankar Sajeev, Purvam Jain, Arya Sikder 외

Large Language Models (LLMs) currently struggle with tool invocation and chaining, as they often hallucinate or miss essential steps in a sequence. We propose RE-GAINS and EnChAnT, two novel frameworks that empower LLMs …

An Agentic Flow for Finite State Machine Extraction using Prompt Chaining

2025-07-15 · Fares Wael, Youssef Maklad, Ali Hamdi, Wael Elsersy arxiv

Finite-State Machines (FSMs) are critical for modeling the operational logic of network protocols, enabling verification, analysis, and vulnerability discovery. However, existing FSM extraction techniques face limitation…

Computing FO-Rewritings in EL in Practice: from Atomic to Conjunctive Queries

2018-04-18 · Peter Hansen, Carsten Lutz

A prominent approach to implementing ontology-mediated queries (OMQs) is to rewrite into a first-order query, which is then executed using a conventional SQL database system. We consider the case where the ontology is fo…

Bi-Chainer: Automated Large Language Models Reasoning with Bidirectional Chaining

2024-06-05 · Shuqi Liu, Bowei He, Linqi Song

Large Language Models (LLMs) have shown human-like reasoning abilities but still face challenges in solving complex logical problems. Existing unidirectional chaining methods, such as forward chaining and backward chaini…

Logical Reasoning