跳到主要导航 跳到搜索 跳到主要内容

CLEAN: Category Knowledge-Driven Compression Framework for Efficient 3D Object Detection

  • Haonan Zhang
  • , Longjun Liu
  • , Fei Hui
  • , Bo Zhang
  • , Hengmin Zhang
  • , Zhiyuan Zha
  • Chang'an University
  • Sichuan University
  • East China University of Science and Technology
  • Jilin University

科研成果: 期刊稿件文章同行评审

5 引用 (Scopus)

摘要

Deep neural networks (DNNs) are potent in LiDAR-based 3D object detection (LiDAR-3DOD), yet their deployment remains daunting due to their cumbersome parameters and computations. Knowledge distillation (KD) is promising for compressing DNNs in LiDAR-3DOD. However, most existing KD methods transfer inadequate knowledge between homogeneous detectors, and do not thoroughly explore optimal student architectures, resulting in insufficient gains for compact student detectors. To this end, we propose a category knowledge-driven compression framework to achieve efficient LiDAR-based 3D detectors. Firstly, we distill knowledge from two-stage teacher detectors to one-stage student detectors, overcoming the limitations of homogeneous pairs. To conduct KD in these heterogeneous pairs, we explore the gap between heterogeneous detectors, and introduce category knowledge-driven KD (CaKD), which includes both student-oriented distillation and two-stage-oriented label assignment distillation. Secondly, to search for the optimal architecture of compact student detectors, we introduce a masked category knowledge-driven structured pruning scheme. This scheme evaluates filter importance by analyzing the changes in category predictions related to foreground regions before and after filter removal, and prunes the less important filters accordingly. Finally, we propose a modified IoU-aware redundancy elimination module to remove redundant false positive samples, thereby further improving the accuracy of detectors. Experiments on various point cloud datasets demonstrate that our method delivers impressive results. For example, on KITTI, several compressed one-stage detectors outperform two-stage detectors in both efficiency and accuracy. Besides, on WOD-mini, our framework reduces the memory footprint of CenterPoint by 5.2× and improves the L2 mAPH by 0.55%.

源语言英语
页(从-至)8740-8755
页数16
期刊IEEE Transactions on Pattern Analysis and Machine Intelligence
47
10
DOI
出版状态已出版 - 2025

学术指纹

探究 'CLEAN: Category Knowledge-Driven Compression Framework for Efficient 3D Object Detection' 的科研主题。它们共同构成独一无二的学术指纹。

引用此