跳到主要导航 跳到搜索 跳到主要内容

Task-driven Image Fusion with Learnable Fusion Loss

  • Haowen Bai
  • , Jiangshe Zhang
  • , Zixiang Zhao
  • , Yichen Wu
  • , Lilun Deng
  • , Yukun Cui
  • , Tao Feng
  • , Shuang Xu
  • Xi'an Jiaotong University
  • Swiss Federal Institute of Technology Zurich
  • City University of Hong Kong
  • Tsinghua University
  • Northwestern Polytechnical University Xian

科研成果: 期刊稿件会议文章同行评审

63 引用 (Scopus)

摘要

Multi-modal image fusion aggregates information from multiple sensor sources, achieving superior visual quality and perceptual features compared to single-source images, often improving downstream tasks. However, current fusion methods for downstream tasks still use predefined fusion objectives that potentially mismatch the downstream tasks, limiting adaptive guidance and reducing model flexibility. To address this, we propose Task-driven Image Fusion (TDFusion), a fusion framework incorporating a learnable fusion loss guided by task loss. Specifically, our fusion loss includes learnable parameters modeled by a neural network called the loss generation module. This module is supervised by the downstream task loss in a meta-learning manner. The learning objective is to minimize the task loss of fused images after optimizing the fusion module with the fusion loss. Iterative updates between the fusion module and the loss module ensure that the fusion network evolves toward minimizing task loss, guiding the fusion process toward the task objectives. TDFusion's training relies entirely on the downstream task loss, making it adaptable to any specific task. It can be applied to any architecture of fusion and task networks. Experiments demonstrate TDFusion's performance through fusion experiments conducted on four different datasets, in addition to evaluations on semantic segmentation and object detection tasks. The code is available at https://github.com/HaowenBai/TDFusion.

源语言英语
页(从-至)7457-7468
页数12
期刊Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition
DOI
出版状态已出版 - 2025
活动2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2025 - Nashville, 美国
期限: 11 6月 202515 6月 2025

学术指纹

探究 'Task-driven Image Fusion with Learnable Fusion Loss' 的科研主题。它们共同构成独一无二的学术指纹。

引用此