跳到主要导航 跳到搜索 跳到主要内容

Dual-Domain Visual Prompt Learning for Multi-Modal Medical Image Saliency Prediction

  • Ning Dai
  • , Mai Xu
  • , Xiaowan Hu*
  • , Ce Zheng
  • , Lai Jiang
  • *此作品的通讯作者
  • Beihang University

科研成果: 期刊稿件文章同行评审

摘要

Medical image saliency prediction plays a pivotal role in emulating clinician visual attention to prioritize diagnostically critical regions. Current methods remain constrained by their spatial-domain dependency and limited cross-modality generalizability, neglecting frequency-domain patterns critical for subtle pathology detection while suffering from over-specialization in specific imaging modalities. Therefore, we propose a dual-domain visual prompt network (DVPNet) that integrates cross-modality generalization with spectral pattern awareness. On the one hand, DVPNet establishes a dataset prompt branch that dynamically modulates spatial feature encoding through modality-specific priors, allowing adaptive interpretation of heterogeneous medical imaging domains. On the other hand, a spatial-frequency hybrid prompt module employs learnable wavelet filters to decompose images into multi-scale spectral components, preserving low-frequency anatomical context while enhancing discriminative high-frequency biomarkers that are typically obscured in previous pixel-level analysis. By seamlessly integrating these complementary representations, DVPNet optimally synthesizes spatial and spectral evidence, enabling robust generalization across diverse medical imaging modalities while sustaining computational efficiency. Extensive experimental results on two distinct datasets demonstrate that the proposed method outperforms state-of-the-art approaches, showing superior saliency prediction performance and enhanced generalizability across medical contexts.

源语言英语
页(从-至)4302-4315
页数14
期刊IEEE Journal of Biomedical and Health Informatics
30
5
DOI
出版状态已出版 - 1 5月 2026

学术指纹

探究 'Dual-Domain Visual Prompt Learning for Multi-Modal Medical Image Saliency Prediction' 的科研主题。它们共同构成独一无二的学术指纹。

引用此