跳到主要导航 跳到搜索 跳到主要内容

End-to-end pedestrian trajectory prediction via Efficient Multi-modal Predictors

  • Xi'an Jiaotong University
  • Wormpex AI Research

科研成果: 期刊稿件文章同行评审

3 引用 (Scopus)

摘要

Pedestrian trajectory prediction plays a key role in understanding human behavior and guiding autonomous driving. It is a difficult task due to the multi-modal nature of human motion. Recent advances have mainly focused on modeling this multi-modality, either by using implicit generative models or explicit pre-defined anchors. However, the former is limited by the sampling problem, while the latter introduces strong prior to the data, both of which require extra tricks to achieve better performance. To address these issues, we propose a simple yet effective framework called Efficient Multi-modal Predictors (EMP), which casts off the generative paradigm and predicts multi-modal trajectories in an end-to-end style. It is achieved by combining a set of parallel predictors with a model error based sparse selector. During training, the entire set of parallel multi-modal predictors will converge into disjoint subsets, with each subset specializing in one mode, thus obtaining multi-modal prediction with no human prior and reducing the problems of above two genres. Experiments on SDD/ETH-UCY/NBA datasets show that EMP achieves state-of-the-art performance with the highest inference speed. Additionally, we show that by replacing multi-modal modules with EMP, state-of-the-art works outperform their baselines, which further validate the versatility of EMP. Moreover, we formally prove that EMP can alleviate the problem of modal collapse and has a low test error bound.

源语言英语
文章编号104107
期刊Computer Vision and Image Understanding
248
DOI
出版状态已出版 - 11月 2024

学术指纹

探究 'End-to-end pedestrian trajectory prediction via Efficient Multi-modal Predictors' 的科研主题。它们共同构成独一无二的指纹。

引用此