paper-with-me

Papers

IAAO: Interactive Affordance Learning for Articulated Objects in 3D Environments

2025-01-01 · CVPR 2025 1 · Can Zhang, Gim Hee Lee

This work presents IAAO, a novel framework that builds an explicit 3D model for intelligent agents to gain understanding of articulated objects in their environment through interaction. Unlike prior methods that rely on task-specific networks and assumptions about movable parts, our IAAO leverages large foundation models to estimate interactive affordances and part articulations in three stages. We first build hierarchical features and label fields for each object state using 3D Gaussian Splatting (3DGS) by distilling mask features and view-consistent labels from multi-view images. We then perform object- and part-level queries on the 3D Gaussian primitives to identify static and articulated elements, estimating global transformations and local articulation parameters along with affordances. Finally, scenes from different states are merged and refined based on the estimated transformations, enabling robust affordance-based interaction and manipulation of objects. Experimental results demonstrate the effectiveness of our method.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

3DGS

Similar Papers 제목 키워드 기반

Ditto in the House: Building Articulation Models of Indoor Scenes through Interactive Perception

2023-02-02 · Cheng-Chun Hsu, Zhenyu Jiang, Yuke Zhu

Virtualizing the physical world into virtual models has been a critical technique for robot navigation and planning in the real world. To foster manipulation with articulated objects in everyday life, this work explores …

Robot Navigation

AdaAfford: Learning to Adapt Manipulation Affordance for 3D Articulated Objects via Few-shot Interactions

2021-12-01 · Yian Wang, Ruihai Wu, Kaichun Mo, Jiaqi Ke 외

Perceiving and interacting with 3D articulated objects, such as cabinets, doors, and faucets, pose particular challenges for future home-assistant robots performing daily tasks in human environments. Besides parsing the …

Friction

ManipGPT: Is Affordance Segmentation by Large Vision Models Enough for Articulated Object Manipulation?

2024-12-13 · Taewhan Kim, Hojin Bae, Zeming Li, Xiaoqi Li 외

Visual actionable affordance has emerged as a transformative approach in robotics, focusing on perceiving interaction areas prior to manipulation. Traditional methods rely on pixel sampling to identify successful interac…

Robot Manipulation

ArtiBench and ArtiBrain: Benchmarking Generalizable Vision-Language Articulated Object Manipulation

2025-11-25 · Yuhan Wu, Tiantian Wei, Shuo Wang, ZhiChao Wang 외 arxiv

Interactive articulated manipulation requires long-horizon, multi-step interactions with appliances while maintaining physical consistency. Existing vision-language and diffusion-based policies struggle to generalize acr…

ArtiWorld: LLM-Driven Articulation of 3D Objects in Scenes

2025-11-17 · Yixuan Yang, Luyang Xie, Zhen Luo, Zixiang Zhao 외 arxiv

Building interactive simulators and scalable robot-learning environments requires a large number of articulated assets. However, most existing 3D assets in simulation are rigid, and manually converting them into articula…