跳到主要导航 跳到搜索 跳到主要内容

Enhancing 3D Instance Segmentation With Dense Connection Decoder and Layer-Aware Fusion

  • Xidian University
  • Xi'an Jiaotong University

科研成果: 期刊稿件文章同行评审

摘要

3D instance segmentation (3DIS) aims to identify object instances in a 3D scene by predicting binary foreground masks with corresponding semantic labels. Transformer-based methods have demonstrated strong performance by effectively capturing global context information through attention mechanisms. However, existing approaches primarily focus on capturing external relationships between scene features and instance queries, while overlooking internal dependencies between queries across decoder layers. This limitation can lead to inconsistencies in query mask predictions across layers, ultimately hindering segmentation performance and slowing model convergence. To address this, we propose the Dense Connection Decoder (DCD), a novel architecture that explicitly models dependencies between instance queries across decoder layers. Our design introduces a Fusion Module and a Memory Module to construct layer-aware hybrid states, dynamically assigning information weights to previous queries based on their decoder distance. Additionally, a Selection Module refines query features through a gating mechanism, adaptively controlling the influence of upstream information. By enforcing prediction consistency across layers, DCD not only enhances segmentation accuracy, but also accelerates model convergence. Extensive experiments on ScanNetV2, ScanNet++V2, ScanNet200, and S3DIS demonstrate that DCD outperforms existing transformer-based baselines, achieving state-of-the-art performance and faster convergence.

源语言英语
页(从-至)10186-10193
页数8
期刊IEEE Robotics and Automation Letters
10
10
DOI
出版状态已出版 - 2025

学术指纹

探究 'Enhancing 3D Instance Segmentation With Dense Connection Decoder and Layer-Aware Fusion' 的科研主题。它们共同构成独一无二的指纹。

引用此