TESGNN: Temporal Equivariant Scene Graph Neural Networks for Efficient and Robust Multi-View 3D Scene Understanding

Quang P. M. Pham; Khoi T. N. Nguyen; Lan C. Ngo; Truong Do; Dezhen Song; Truong-Son Hy

arXiv:2411.10509·cs.CV·November 4, 2025

TESGNN: Temporal Equivariant Scene Graph Neural Networks for Efficient and Robust Multi-View 3D Scene Understanding

Quang P. M. Pham, Khoi T. N. Nguyen, Lan C. Ngo, Truong Do, Dezhen Song, Truong-Son Hy

PDF

Open Access 1 Repo

TL;DR

TESGNN introduces a novel neural network architecture that preserves symmetry and models temporal relationships in 3D scene graphs, significantly improving accuracy, robustness, and efficiency for multi-view scene understanding tasks.

Contribution

It proposes TESGNN, combining symmetry-preserving scene graph generation with temporal modeling, addressing limitations of prior methods in robustness and dynamic scene understanding.

Findings

01

Achieves higher accuracy in scene graph generation.

02

Faster training convergence compared to existing methods.

03

Produces more stable and accurate global scene representations.

Abstract

Scene graphs have proven to be highly effective for various scene understanding tasks due to their compact and explicit representation of relational information. However, current methods often overlook the critical importance of preserving symmetry when generating scene graphs from 3D point clouds, which can lead to reduced accuracy and robustness, particularly when dealing with noisy, multi-view data. Furthermore, a major limitation of prior approaches is the lack of temporal modeling to capture time-dependent relationships among dynamically evolving entities in a scene. To address these challenges, we propose Temporal Equivariant Scene Graph Neural Network (TESGNN), consisting of two key components: (1) an Equivariant Scene Graph Neural Network (ESGNN), which extracts information from 3D point clouds to generate scene graph while preserving crucial symmetry properties, and (2) a…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

hysonlab/tesgraph
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

Topics3D Shape Modeling and Analysis · Human Pose and Action Recognition · Advanced Vision and Imaging

MethodsGraph Neural Network