paper-with-me

홈 › Papers

Articulate-Anything: Automatic Modeling of Articulated Objects via a Vision-Language Foundation Model

2024-10-03 · Long Le, Jason Xie, William Liang, Hung-Ju Wang, Yue Yang, Yecheng Jason Ma, Kyle Vedder, Arjun Krishna, Dinesh Jayaraman, Eric Eaton

Interactive 3D simulated objects are crucial in AR/VR, animations, and robotics, driving immersive experiences and advanced automation. However, creating these articulated objects requires extensive human effort and expertise, limiting their broader applications. To overcome this challenge, we present Articulate-Anything, a system that automates the articulation of diverse, complex objects from many input modalities, including text, images, and videos. Articulate-Anything leverages vision-language models (VLMs) to generate code that can be compiled into an interactable digital twin for use in standard 3D simulators. Our system exploits existing 3D asset datasets via a mesh retrieval mechanism, along with an actor-critic system that iteratively proposes, evaluates, and refines solutions for articulating the objects, self-correcting errors to achieve a robust outcome. Qualitative evaluations demonstrate Articulate-Anything's capability to articulate complex and even ambiguous object affordances by leveraging rich grounded inputs. In extensive quantitative experiments on the standard PartNet-Mobility dataset, Articulate-Anything substantially outperforms prior work, increasing the success rate from 8.7-11.6% to 75% and setting a new bar for state-of-the-art performance. We further showcase the utility of our system by generating 3D assets from in-the-wild video inputs, which are then used to train robotic policies for fine-grained manipulation tasks in simulation that go beyond basic pick and place. These policies are then transferred to a real robotic system.

📄 PDF Abstract BibTeX arXiv:2410.13882

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

URDF-Anything: Constructing Articulated Objects with 3D Multimodal Language Model

2025-11-02 · Zhe Li, Xiang Bai, Jieyu Zhang, Zhuangzhe Wu 외 arxiv

Constructing accurate digital twins of articulated objects is essential for robotic simulation training and embodied AI world model building, yet historically requires painstaking manual modeling or multi-stage pipelines…

Parameter Prediction

Survey on Modeling of Human-made Articulated Objects

2024-03-22 · Jiayi Liu, Manolis Savva, Ali Mahdavi-Amiri

3D modeling of articulated objects is a research problem within computer vision, graphics, and robotics. Its objective is to understand the shape and motion of the articulated components, represent the geometry and mobil…

ObjectSurvey

GAMMA: Generalizable Articulation Modeling and Manipulation for Articulated Objects

2023-09-28 · Qiaojun Yu, JunBo Wang, Wenhai Liu, Ce Hao 외

Articulated objects like cabinets and doors are widespread in daily life. However, directly manipulating 3D articulated objects is challenging because they have diverse geometrical shapes, semantic categories, and kineti…

Manner Of Articulation DetectionRobot ManipulationTrajectory Planning

V-MAO: Generative Modeling for Multi-Arm Manipulation of Articulated Objects

2021-11-07 · Xingyu Liu, Kris M. Kitani

Manipulating articulated objects requires multiple robot arms in general. It is challenging to enable multiple robot arms to collaboratively complete manipulation tasks on articulated objects. In this paper, we present $…

MuJoCoObject

Articulate AnyMesh: Open-Vocabulary 3D Articulated Objects Modeling

2025-02-04 · Xiaowen Qiu, Jincheng Yang, Yian Wang, Zhehuan Chen 외

3D articulated objects modeling has long been a challenging problem, since it requires to capture both accurate surface geometries and semantically meaningful and spatially precise structures, parts, and joints. Existing…

ObjectVisual Prompting