paper-with-me

홈 › Papers

RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies

2026-03-04 · Yinpei Dai, Hongze Fu, Jayjun Lee, Yuejiang Liu, Haoran Zhang, Jianing Yang, Chelsea Finn, Nima Fazeli, Joyce Chai arxiv

Memory is critical for long-horizon and history-dependent robotic manipulation. Such tasks often involve counting repeated actions or manipulating objects that become temporarily occluded. Recent vision-language-action (VLA) models have begun to incorporate memory mechanisms; however, their evaluations remain confined to narrow, non-standardized settings. This limits systematic understanding, comparison, and progress measurement. To address these challenges, we introduce RoboMME: a large-scale standardized benchmark for evaluating and advancing VLA models in long-horizon, history-dependent scenarios. Our benchmark comprises 16 manipulation tasks constructed under a carefully designed taxonomy that evaluates temporal, spatial, object, and procedural memory. We further develop a suite of 14 memory-augmented VLA variants built on the π0.5 backbone to systematically explore different memory representations across multiple integration strategies. Experimental results show that the effectiveness of memory representations is highly task-dependent, with each design offering distinct advantages and limitations across different tasks. Videos and code can be found at our website https://robomme.github.io.

📄 PDF Abstract BibTeX arXiv:2603.04639

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RoboMME-Interference: Benchmarking Robot Memory Under Interference

2026-06-21 · Soumil Rathi arxiv

Robots deployed in realistic settings will accumulate experience across many sessions and tasks over their deployment. The robot's tasks may often require it to remember information from multiple sessions ago, making lon…

Robot Policy Evaluation for Sim-to-Real Transfer: A Benchmarking Perspective

2025-08-14 · Xuning Yang, Clemens Eppner, Jonathan Tremblay, Dieter Fox 외 arxiv

Current vision-based robotics simulation benchmarks have significantly advanced robotic manipulation research. However, robotics is fundamentally a real-world problem, and evaluation for real-world applications has lagge…

Vesta: A Generalist Embodied Reasoning Model

2026-06-18 · Johan Bjorck, Zhiqi Li, Yunze Man, Jing Wang 외 arxiv

Robots operating in open-world environments must seamlessly integrate localization, spatial reasoning, navigation, and long-horizon planning. While specialist models excel at individual tasks, deploying a multi-model sta…

Spatial Reasoning

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

2026-04-10 · Xuning Yang, Rishit Dagli, Alex Zook, Hugo Hadfield 외 arxiv

The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid performance saturation and a lack of true generalization testing. Exis…

Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation

2024-10-10 · Qingwen Bu, Hongyang Li, Li Chen, Jisong Cai 외

The increasing demand for versatile robotic systems to operate in diverse and dynamic environments has emphasized the importance of a generalist policy, which leverages a large cross-embodiment data corpus to facilitate …

Robot ManipulationVision-Language-Action