paper-with-me

Papers

MemoAct: Atkinson-Shiffrin-Inspired Memory-Augmented Visuomotor Policy for Robotic Manipulation

2026-03-19 · Liufan Tan, Jiale Li, Gangshan Jing arxiv

Memory-augmented robotic policies are essential in handling memory-dependent tasks. However, existing approaches typically rely on simple observation window extensions, struggling to simultaneously achieve precise task state tracking and robust long-horizon retention. To overcome these challenges, inspired by the Atkinson-Shiffrin memory model, we propose MemoAct, a hierarchical memory-based policy that leverages distinct memory tiers to tackle specific bottlenecks. Specifically, lossless short-term memory ensures precise task state tracking, while compressed long-term memory enables robust long-horizon retention. To enrich the evaluation landscape, we construct MemoryRTBench based on RoboTwin 2.0, specifically tailored to assess policy capabilities in task state tracking and long-horizon retention. Extensive experiments across simulated and real-world scenarios demonstrate that MemoAct achieves superior performance compared to both existing Markovian baselines and history-aware policies. The project page is \href{https://tlf-tlf.github.io/MemoActPage/}{available}.

📄 PDF Abstract BibTeX arXiv:2603.18494

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

XMem: Long-Term Video Object Segmentation with an Atkinson-Shiffrin Memory Model

2022-07-14 · Ho Kei Cheng, Alexander G. Schwing

We present XMem, a video object segmentation architecture for long videos with unified feature memory stores inspired by the Atkinson-Shiffrin memory model. Prior work on video object segmentation typically only uses one…

2D Human Pose Estimation2D Object Detection3D Absolute Human Pose EstimationSegmentation+4

LightMem: Lightweight and Efficient Memory-Augmented Generation

2025-10-21 · Jizhan Fang, Xinle Deng, Haoming Xu, Ziyan Jiang 외 arxiv

Despite their remarkable capabilities, Large Language Models (LLMs) struggle to effectively leverage historical interaction information in dynamic and complex environments. Memory systems enable LLMs to move beyond state…

MUlti-Store Tracker (MUSTer): A Cognitive Psychology Inspired Approach to Object Tracking

2015-06-01 · CVPR 2015 6 · Zhibin Hong, Zhe Chen, Chaohui Wang, Xue Mei 외

Variations in the appearance of a tracked object, such as changes in geometry/photometry, camera viewpoint, illumination, or partial occlusion, pose a major challenge to object tracking. Here, we adopt cognitive psycholo…

ObjectObject Tracking

MovieChat: From Dense Token to Sparse Memory for Long Video Understanding

2023-07-31 · CVPR 2024 1 · Enxin Song, Wenhao Chai, Guanhong Wang, Yucheng Zhang 외

Recently, integrating video foundation models and large language models to build a video understanding system can overcome the limitations of specific pre-defined vision tasks. Yet, existing systems can only handle video…

Multiple-choiceQuestion AnsweringVideo-based Generative Performance BenchmarkingVideo-based Generative Performance Benchmarking (Consistency)+12

Black-box Unsupervised Domain Adaptation with Bi-directional Atkinson-Shiffrin Memory

2023-08-25 · ICCV 2023 1 · Jingyi Zhang, Jiaxing Huang, Xueying Jiang, Shijian Lu

Black-box unsupervised domain adaptation (UDA) learns with source predictions of target data without accessing either source data or source models during training, and it has clear superiority in data privacy and flexibi…

Domain Adaptationimage-classificationImage ClassificationMemorization+4