paper-with-me

Papers

PhysGraph: A Physics-aware 3D Scene Graph for Perception and Reasoning

2026-06-07 · Haoyu Li, Aaron Thomas, Shuyan Zhou, Xianyi Cheng arxiv

To perform a wide range of daily tasks, robots need to construct a 3D representation that is semantically rich, physically grounded, and structured enough to support task planning and affordance prediction. However, existing approaches primarily focus on semantic retrieval, often overlooking physical and kinematic factors. Methods that attempt to model physical properties typically rely on narrow training sets or single-object modeling, limiting scalability and generalization across diverse object types. To address these challenges, we present PhysGraph, a framework that unifies symbolic reasoning with structured 3D geometry to model kinematic and physical properties in cluttered scenes. Given RGB-D observations, PhysGraph reconstructs object-centric 3D geometry and associates object instances across views. It then decomposes objects into functional parts and infers materials and articulations through visual reasoning. Evaluated on both synthetic and real-world datasets, PhysGraph achieves state-of-the-art results in semantic segmentation, multi-object mass estimation, and articulation prediction. With its simple yet effective design, PhysGraph produces physically consistent and semantically structured scene graphs, serving as a structured 3D representation for downstream tasks such as constraint-aware 3D affordance prediction and real-to-sim transfer, both of which are demonstrated in our experiments.

📄 PDF Abstract BibTeX arXiv:2606.08655

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic SegmentationSemantic RetrievalVisual Reasoning

Similar Papers 제목 키워드 기반

PhysGraph: Physically-Grounded Graph-Transformer Policies for Bimanual Dexterous Hand-Tool-Object Manipulation

2026-03-02 · Runfa Blark Li, David Kim, Xinshuang Liu, Keito Suzuki 외 arxiv

Bimanual dexterous manipulation for tool use remains a formidable challenge in robotics due to the high-dimensional state space and complicated contact dynamics. Existing methods naively represent the entire system state…

Learning Visual Commonsense for Robust Scene Graph Generation

2020-06-17 · ECCV 2020 8 · Alireza Zareian, Zhecan Wang, Haoxuan You, Shih-Fu Chang

Scene graph generation models understand the scene through object and predicate recognition, but are prone to mistakes due to the challenges of perception in the wild. Perception errors often lead to nonsensical composit…

Graph GenerationScene Graph GenerationScene Understanding

PhysGraph: Physics-Based Integration Using Graph Neural Networks

2023-01-27 · Oshri Halimi, Egor Larionov, Zohar Barzelay, Philipp Herholz 외

Physics-based simulation of mesh based domains remains a challenging task. State-of-the-art techniques can produce realistic results but require expert knowledge. A major bottleneck in many approaches is the step of inte…

Virtual Try-on

Learning to See Physics via Visual De-animation

2017-12-01 · NeurIPS 2017 12 · Jiajun Wu, Erika Lu, Pushmeet Kohli, Bill Freeman 외

We introduce a paradigm for understanding physical scenes without human annotations. At the core of our system is a physical world representation that is first recovered by a perception module and then utilized by physic…

Future predictionState Estimation

Sim2Radar: Toward Bridging the Radar Sim-to-Real Gap with VLM-Guided Scene Reconstruction

2026-02-10 · Emily Bejerano, Federico Tondolo, Ayaan Qayyum, Xiaofan Yu 외 arxiv

Millimeter-wave (mmWave) radar provides reliable perception in visually degraded indoor environments (e.g., smoke, dust, and low light), but learning-based radar perception is bottlenecked by the scarcity and cost of col…

Monocular Depth EstimationTransfer LearningObject Detection