paper-with-me

홈 › Papers

HiMaCon: Discovering Hierarchical Manipulation Concepts from Unlabeled Multi-Modal Data

2025-10-13 · Ruizhe Liu, Pei Zhou, Qian Luo, Li Sun, Jun Cen, Yibing Song, Yanchao Yang arxiv

Effective generalization in robotic manipulation requires representations that capture invariant patterns of interaction across environments and tasks. We present a self-supervised framework for learning hierarchical manipulation concepts that encode these invariant patterns through cross-modal sensory correlations and multi-level temporal abstractions without requiring human annotation. Our approach combines a cross-modal correlation network that identifies persistent patterns across sensory modalities with a multi-horizon predictor that organizes representations hierarchically across temporal scales. Manipulation concepts learned through this dual structure enable policies to focus on transferable relational patterns while maintaining awareness of both immediate actions and longer-term goals. Empirical evaluation across simulated benchmarks and real-world deployments demonstrates significant performance improvements with our concept-enhanced policies. Analysis reveals that the learned concepts resemble human-interpretable manipulation primitives despite receiving no semantic supervision. This work advances both the understanding of representation learning for manipulation and provides a practical approach to enhancing robotic performance in complex scenarios.

📄 PDF Abstract BibTeX arXiv:2510.11321

Code (0)

등록된 구현이 없습니다.

Tasks

Representation Learning

Similar Papers 제목 키워드 기반

HINT: Hierarchical Neuron Concept Explainer

2022-03-27 · CVPR 2022 1 · Andong Wang, Wei-Ning Lee, Xiaojuan Qi

To interpret deep networks, one main approach is to associate neurons with human-understandable concepts. However, existing methods often ignore the inherent relationships of different concepts (e.g., dog and cat both be…

Object LocalizationWeakly-Supervised Object Localization

Learning Hierarchical Semantic Image Manipulation through Structured Representations

2018-08-22 · NeurIPS 2018 12 · Seunghoon Hong, Xinchen Yan, Thomas Huang, Honglak Lee

Understanding, reasoning, and manipulating semantic concepts of images have been a fundamental research problem for decades. Previous work mainly focused on direct manipulation on natural image manifold through color str…

Image GenerationImage ManipulationObject

A Pseudo-Label Method for Coarse-to-Fine Multi-Label Learning with Limited Supervision

2019-03-25 · ICLR Workshop LLD 2019 · Cheng-Yu Hsieh, Miao Xu, Gang Niu, Hsuan-Tien Lin 외

The goal of multi-label learning (MLL) is to associate a given instance with its relevant labels from a set of concepts. Previous works of MLL mainly focused on the setting where the concept set is assumed to be fixed, w…

Meta-LearningMulti-Label LearningPseudo Label

SCAN: Learning Hierarchical Compositional Visual Concepts

2017-07-11 · ICLR 2018 1 · Irina Higgins, Nicolas Sonnerat, Loic Matthey, Arka Pal 외

The seemingly infinite diversity of the natural world arises from a relatively small set of coherent rules, such as the laws of physics or chemistry. We conjecture that these rules give rise to regularities that can be d…

Understanding Distributed Representations of Concepts in Deep Neural Networks without Supervision

2023-12-28 · Wonjoon Chang, Dahee Kwon, Jaesik Choi

Understanding intermediate representations of the concepts learned by deep learning classifiers is indispensable for interpreting general model behaviors. Existing approaches to reveal learned concepts often rely on huma…

Deep Learning