跳到主要导航 跳到搜索 跳到主要内容

3D Semantic Gaussian via Geometric-Semantic Hypergraph Computation

  • Xinran Wang
  • , Zhiqiang Tian
  • , Dejian Guo
  • , Siqi Li
  • , Shaoyi Du
  • , Xiangmin Han
  • , Yue Gao
  • Xi'an Jiaotong University
  • Tsinghua University

科研成果: 期刊稿件文章同行评审

摘要

Semantic labels are inherently tied to geometry and luminance reconstruction, as entities with similar shapes and appearances often share categories. Traditional methods use synthesis-analysis, NeRF, or 3D Gaussian representations to encode semantics and geometry separately. However, 2D methods lack view consistency, NeRF extensions are slow, and faster 3D Gaussian methods risk spatial and channel inconsistencies between semantic and RGB. Moreover, these methods require costly manual dense semantic labels. To alleviate resource demands and achieve effective semantic reconstruction with sparse inputs while enhancing RGB rendering quality, we build upon 3D Gaussian by integrating semantic features from pre-trained models - requiring no additional ground truth input - into Gaussian features, and construct a hypergraph neural network to capture higher-order correlations across RGB and semantic information as well as between different frames. Hypergraphs use hyperedges to link multiple vertices, capturing complex relationships essential for cross-modal tasks. This higher-order structure addresses the limitations of NeRF and Gaussian methods, which lack the capacity for such advanced associations. This framework enables precise novel view synthesis and 2D semantic reconstruction without manual annotations, achieving state-of-the-art results for RGB and semantic tasks on room-scale scenes in the ScanNet and Replica datasets, while supporting real-time rendering speeds of 34 FPS.

源语言英语
页(从-至)3068-3080
页数13
期刊IEEE Transactions on Multimedia
28
DOI
出版状态已出版 - 2026

学术指纹

探究 '3D Semantic Gaussian via Geometric-Semantic Hypergraph Computation' 的科研主题。它们共同构成独一无二的指纹。

引用此