跳到主要导航 跳到搜索 跳到主要内容

SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation

  • Sea AI Lab
  • Skywork AI

科研成果: 期刊稿件文章同行评审

13 引用 (Scopus)

摘要

Existing Transformers for monocular 3D human shape and pose estimation typically have a quadratic computation and memory complexity with respect to the feature length, which hinders the exploitation of fine-grained information in high-resolution features that is beneficial for accurate reconstruction. In this work, we propose an SMPL-based Transformer framework (SMPLer) to address this issue. SMPLer incorporates two key ingredients: a decoupled attention operation and an SMPL-based target representation, which allow effective utilization of high-resolution features in the Transformer. In addition, based on these two designs, we also introduce several novel modules including a multi-scale attention and a joint-aware attention to further boost the reconstruction performance. Extensive experiments demonstrate the effectiveness of SMPLer against existing 3D human shape and pose estimation methods both quantitatively and qualitatively. Notably, the proposed algorithm achieves an MPJPE of 45.2mm on the Human3.6M dataset, improving upon the state-of-the-art approach (Lin et al., 2021) by more than 10% with fewer than one-third of the parameters.

源语言英语
文章编号10354384
页(从-至)3275-3289
页数15
期刊IEEE Transactions on Pattern Analysis and Machine Intelligence
46
5
DOI
出版状态已出版 - 1 5月 2024

学术指纹

探究 'SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation' 的科研主题。它们共同构成独一无二的指纹。

引用此