paper-with-me

홈 › Papers

Hi-Dyna Graph: Hierarchical Dynamic Scene Graph for Robotic Autonomy in Human-Centric Environments

2025-05-30 · Jiawei Hou, xiangyang xue, Taiping Zeng

Autonomous operation of service robotics in human-centric scenes remains challenging due to the need for understanding of changing environments and context-aware decision-making. While existing approaches like topological maps offer efficient spatial priors, they fail to model transient object relationships, whereas dense neural representations (e.g., NeRF) incur prohibitive computational costs. Inspired by the hierarchical scene representation and video scene graph generation works, we propose Hi-Dyna Graph, a hierarchical dynamic scene graph architecture that integrates persistent global layouts with localized dynamic semantics for embodied robotic autonomy. Our framework constructs a global topological graph from posed RGB-D inputs, encoding room-scale connectivity and large static objects (e.g., furniture), while environmental and egocentric cameras populate dynamic subgraphs with object position relations and human-object interaction patterns. A hybrid architecture is conducted by anchoring these subgraphs to the global topology using semantic and spatial constraints, enabling seamless updates as the environment evolves. An agent powered by large language models (LLMs) is employed to interpret the unified graph, infer latent task triggers, and generate executable instructions grounded in robotic affordances. We conduct complex experiments to demonstrate Hi-Dyna Grap's superior scene representation effectiveness. Real-world deployments validate the system's practicality with a mobile manipulator: robotics autonomously complete complex tasks with no further training or complex rewarding in a dynamic scene as cafeteria assistant. See https://anonymous.4open.science/r/Hi-Dyna-Graph-B326 for video demonstration and more details.

📄 PDF Abstract BibTeX arXiv:2506.00083

Code (0)

등록된 구현이 없습니다.

Tasks

Graph GenerationHuman-Object Interaction DetectionNeRFScene Graph GenerationVideo scene graph generation

Methods 이 논문이 사용한 방법론

Golden Queue Managers 설명 없음

Similar Papers 제목 키워드 기반

Aion: Towards Hierarchical 4D Scene Graphs with Temporal Flow Dynamics

2025-12-10 · Iacopo Catalano, Eduardo Montijano, Javier Civera, Julio A. Placed 외 arxiv

Autonomous navigation in dynamic environments requires spatial representations that capture both semantic structure and temporal evolution. 3D Scene Graphs (3DSGs) provide hierarchical multi-resolution abstractions that …

CausalNav: A Long-term Embodied Navigation System for Autonomous Mobile Robots in Dynamic Outdoor Scenarios

2026-01-05 · Hongbo Duan, Shangyi Luo, Zhiyuan Deng, Yanbo Chen 외 arxiv

Autonomous language-guided navigation in large-scale outdoor environments remains a key challenge in mobile robotics, due to difficulties in semantic reasoning, dynamic conditions, and long-term stability. We propose Cau…

THYME: Temporal Hierarchical-Cyclic Interactivity Modeling for Video Scene Graphs in Aerial Footage

2025-07-12 · Trong-Thuan Nguyen, Pha Nguyen, Jackson Cothren, Alper Yilmaz 외 arxiv

The rapid proliferation of video in applications such as autonomous driving, surveillance, and sports analytics necessitates robust methods for dynamic scene understanding. Despite advances in static scene graph generati…

Video scene graph generationScene UnderstandingAutonomous Driving

Sketching Image Gist: Human-Mimetic Hierarchical Scene Graph Generation

2020-07-17 · ECCV 2020 8 · Wenbin Wang, Ruiping Wang, Shiguang Shan, Xilin Chen

Scene graph aims to faithfully reveal humans' perception of image content. When humans analyze a scene, they usually prefer to describe image gist first, namely major objects and key relations in a scene graph. This huma…

Graph GenerationScene Graph GenerationScene Parsing

Graph Canvas for Controllable 3D Scene Generation

2024-11-27 · Libin Liu, Shen Chen, Sen Jia, Jingzhe Shi 외

Spatial intelligence is foundational to AI systems that interact with the physical world, particularly in 3D scene generation and spatial comprehension. Current methodologies for 3D scene generation often rely heavily on…

In-Context LearningScene Generation