Exploiting Edge-Oriented Reasoning for 3D Point-based Scene Graph Analysis
Scene understanding is a critical problem in computer vision. In this paper, we propose a 3D point-based scene graph generation ($\mathbf{SGG_{point}}$) framework to effectively bridge perception and reasoning to achieve scene understanding via three sequential stages, namely scene graph construction, reasoning, and inference. Within the reasoning stage, an EDGE-oriented Graph Convolutional Network ($\texttt{EdgeGCN}$) is created to exploit multi-dimensional edge features for explicit relationship modeling, together with the exploration of two associated twinning interaction mechanisms between nodes and edges for the independent evolution of scene graph representations. Overall, our integrated $\mathbf{SGG_{point}}$ framework is established to seek and infer scene structures of interest from both real-world and synthetic 3D point-based scenes. Our experimental results show promising edge-oriented reasoning effects on scene graph generation studies. We also demonstrate our method advantage on several traditional graph representation learning benchmark datasets, including the node-wise classification on citation networks and whole-graph recognition problems for molecular analysis.
Code (1)
Tasks
3d scene graph generationgraph constructionGraph GenerationGraph Representation LearningRepresentation LearningScene Graph GenerationScene UnderstandingSimilar Papers 제목 키워드 기반
KLDrive: Fine-Grained 3D Scene Reasoning for Autonomous Driving based on Knowledge Graph
Autonomous driving requires reliable reasoning over fine-grained 3D scene facts. Fine-grained question answering over multi-modal driving observations provides a natural way to evaluate this capability, yet existing perc…
Autonomous DrivingQuestion AnsweringLearning to Reason and Use Tools through Unsupervised Fine-Tuning in Task-Oriented Dialog Systems
Current dialogue systems struggle with dynamic information retrieval, often leading to hallucinations and lower response accuracy. We address this by adapting the ReAct framework for Task-Oriented Dialogue, enabling Larg…
Domain GeneralizationInformation RetrievalGraphDialog: Integrating Graph Knowledge into End-to-End Task-Oriented Dialogue Systems
End-to-end task-oriented dialogue systems aim to generate system responses directly from plain text inputs. There are two challenges for such systems: one is how to effectively incorporate external knowledge bases (KBs) …
Dependency ParsingRepresentation LearningTask-Oriented Dialogue SystemsExploiting High Level Scene Cues in Stereo Reconstruction
We present a novel approach to 3D reconstruction which is inspired by the human visual system. This system unifies standard appearance matching and triangulation techniques with higher level reasoning and scene understan…
3D ReconstructionScene UnderstandingVocal Bursts Intensity PredictionAn Interpretable Neuro-Symbolic Reasoning Framework for Task-Oriented Dialogue Generation
We study the interpretability issue of task-oriented dialogue systems in this paper. Previously, most neural-based task-oriented dialogue systems employ an implicit reasoning strategy that makes the model predictions uni…
Dialogue GenerationTask-Oriented Dialogue Systemsvalid