paper-with-me

Papers

Instruct-Particulate: Scaling Feed-Forward 3D Object Articulation with Kinematic Control

2026-06-12 · Ruining Li, Yuxin Yao, Matt Zhou, Chuanxia Zheng, Christian Rupprecht, Joan Lasenby, Shangzhe Wu, Andrea Vedaldi arxiv

Reconstructing articulated 3D objects is important for animation, gaming, and robotic simulations. Recent neural networks can estimate the articulated structure of 3D objects, but their generalization remains limited by the scarcity of annotated data for this task. To address this gap, we introduce Instruct-Particulate, a model that takes a 3D mesh together with a target kinematic specification, including part descriptions, connectivity, joint types, and optional point prompts, and predicts the corresponding kinematic part segmentation and joint motion parameters. The kinematic specification disambiguates the task and allows the model to target annotations of different granularity, thereby making it possible to use more abundant heterogeneous training data. At test time, the kinematic specification can be obtained automatically from large-scale vision-language models, so the model can be applied to any input mesh. To train our model at scale, we construct a heterogeneous dataset of more than 150,000 articulated 3D objects, extending existing publicly available collections with data obtained by partially labelling other 3D models (monolithic or already decomposed into parts) with kinematic labels by means of vision-language models. Experiments show that our model generalizes better across categories and to AI-generated meshes, enabling articulated asset reconstruction from real-world images via image-to-3D models.

📄 PDF Abstract BibTeX arXiv:2606.14699

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Particulate: Feed-Forward 3D Object Articulation

2025-12-12 · Ruining Li, Yuxin Yao, Chuanxia Zheng, Christian Rupprecht 외 arxiv

We introduce Particulate, a feed-forward model that, given a 3D mesh of an object, infers its articulations, including its 3D parts, their kinematic structure, and the motion constraints. The model is based on a transfor…

Native 3D Editing with Full Attention

2025-11-21 · Weiwei Cai, Shuangkang Fang, Weicai Ye, Xin Dong 외 arxiv

Instruction-guided 3D editing is a rapidly emerging field with the potential to broaden access to 3D content creation. However, existing methods face critical limitations: optimization-based approaches are prohibitively …

Operator theory, kernels, and Feedforward Neural Networks

2023-01-03 · Palle E. T. Jorgensen, Myung-Sin Song, James Tian

In this paper we show how specific families of positive definite kernels serve as powerful tools in analyses of iteration algorithms for multiple layer feedforward Neural Network models. Our focus is on particular kernel…

SHAP-EDITOR: Instruction-guided Latent 3D Editing in Seconds

2023-12-14 · CVPR 2024 1 · Minghao Chen, Junyu Xie, Iro Laina, Andrea Vedaldi

We propose a novel feed-forward 3D editing framework called Shap-Editor. Prior research on editing 3D objects primarily concentrated on editing individual objects by leveraging off-the-shelf 2D image editing networks. Th…

SEFORA: Student Essays with Feedback Corpus and LLM Feedback Evaluation Framework

2026-06-30 · Shayan Peyghambari Oskoui, Norah Almousa, Zhaoyi Joey Hou, Carolina Gustafson 외 arxiv

Effective writing feedback is among the strongest drivers of student learning, yet producing it at scale is labor-intensive. LLMs offer a natural path to scaling writing support, but two gaps stand in the way: few public…

Semantic correspondence