跳到主要导航 跳到搜索 跳到主要内容

Multi-level graph self-supervised learning for multi-modal medical corpus construction

  • Yuping Lin
  • , Jingxi Feng
  • , Xudong Chen
  • , Rundong Xue
  • , Jue Jiang
  • , Zhiqiang Tian
  • , Juan Wang
  • Xi'an Jiaotong University
  • The Second Affiliated Hospital of Xi'an Jiaotong University

科研成果: 期刊稿件文章同行评审

3 引用 (Scopus)

摘要

Multi-modal medical corpus, as a novel tool for computer-aided medical diagnosis and learning research, hold significant value in exploring pathogenic mechanisms. However, in the field of neuroscience, brain imaging often lacks semantic labels, making the construction of multi-modal medical corpus challenging. Functional magnetic resonance imaging (fMRI), a commonly used brain imaging, can be employed to study functional brain networks and perform classification to obtain labels. Moreover, due to limited data samples, developing graph foundation model for brain network classification is important. We propose a multi-level graph self-supervised learning (MLGSL) method, which performs multi-level pretraining tasks to uncover disease-related correlation patterns. This graph foundation model is then used to classify brain networks to obtain semantic labels. Additionally, by integrating the potential biomarkers with relevant textual matched through large language model, a multi-modal medical corpus can be constructed. Specifically, MLGSL first conducts a functional brain network link prediction task at the individual level, and a population-level link prediction task on the population-associated network. In the encoder part of pretraining task, the proposed multi-channel enhanced attention graph convolution strengthens the attention mechanism via connection strength, while integrating visible node representations learned from different perspectives. Subsequently, MLGSL performs category prediction of functional brain networks through fine-tuning. Experiments demonstrate that MLGSL achieves optimal brain network classification performance, thereby enabling the acquisition of accurate semantic labels. By combining brain imaging data with key textual information, a multi-modal medical corpus can be constructed, providing important value for interpretation of pathological information and clinical diagnosis of diseases.

源语言英语
文章编号112113
期刊Pattern Recognition
171
DOI
出版状态已出版 - 3月 2026

学术指纹

探究 'Multi-level graph self-supervised learning for multi-modal medical corpus construction' 的科研主题。它们共同构成独一无二的学术指纹。

引用此