paper-with-me

홈 › Papers

Semantic Decomposition and Recognition of Long and Complex Manipulation Action Sequences

2016-10-18 · Eren Erdal Aksoy, Adil Orhan, Florentin Woergoetter

Understanding continuous human actions is a non-trivial but important problem in computer vision. Although there exists a large corpus of work in the recognition of action sequences, most approaches suffer from problems relating to vast variations in motions, action combinations, and scene contexts. In this paper, we introduce a novel method for semantic segmentation and recognition of long and complex manipulation action tasks, such as "preparing a breakfast" or "making a sandwich". We represent manipulations with our recently introduced "Semantic Event Chain" (SEC) concept, which captures the underlying spatiotemporal structure of an action invariant to motion, velocity, and scene context. Solely based on the spatiotemporal interactions between manipulated objects and hands in the extracted SEC, the framework automatically parses individual manipulation streams performed either sequentially or concurrently. Using event chains, our method further extracts basic primitive elements of each parsed manipulation. Without requiring any prior object knowledge, the proposed framework can also extract object-like scene entities that exhibit the same role in semantically similar manipulations. We conduct extensive experiments on various recent datasets to validate the robustness of the framework.

📄 PDF Abstract BibTeX arXiv:1610.05693

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic Segmentation

Similar Papers 제목 키워드 기반

Gentle Manipulation Policy Learning via Demonstrations from VLM Planned Atomic Skills

2025-11-08 · Jiayu Zhou, Qiwei Wu, Jian Li, Zhe Chen 외 arxiv

Autonomous execution of long-horizon, contact-rich manipulation tasks traditionally requires extensive real-world data and expert engineering, posing significant cost and scalability challenges. This paper proposes a nov…

Reinforcement LearningKnowledge Distillation

ODYSSEY: Open-World Quadrupeds Exploration and Manipulation for Long-Horizon Tasks

2025-08-11 · Kaijun Wang, Liqin Lu, Mingyu Liu, Jianuo Jiang 외 arxiv

Language-guided long-horizon mobile manipulation has long been a grand challenge in embodied semantic reasoning, generalizable manipulation, and adaptive locomotion. Three fundamental limitations hinder progress: First, …

Spatial Reasoning

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation

2026-08-04 · Andrea Protopapa, Davide Buoso, Francesca Pistilli, Georgia Chalvatzaki 외 arxiv

Learning long-horizon manipulation skills with reinforcement learning remains challenging due to the complexity of reward design, the limited guidance of sparse rewards, and the high cost of manual subtask annotation. Vi…

Reinforcement LearningGraph Neural Network

Lifelong Language-Conditioned Robotic Manipulation Learning

2026-03-05 · Xudong Wang, Zebin Han, Zhiyu Liu, Gan Li 외 arxiv

Traditional language-conditioned manipulation agent sequential adaptation to new manipulation skills leads to catastrophic forgetting of old skills, limiting dynamic scene practical deployment. In this paper, we propose …

OmniContact: Chaining Meta-Skills via Contact Flow for Generalizable Humanoid Loco-Manipulation

2026-06-24 · Runyi Yu, Xiaoyi Lin, Ji Ma, Yinhuai Wang 외 arxiv

Learning long-horizon humanoid loco-manipulation poses a dual challenge: it requires not only the robust execution of meta-skills but also their seamless, closed-loop chaining equipped with autonomous recovery. Existing …