计算机科学
稳健性(进化)
人工智能
情态动词
机器学习
图像(数学)
图像分割
模式识别(心理学)
分割
训练集
上下文图像分类
数据建模
桥接(联网)
最优化问题
计算机视觉
医学影像学
偏爱
强化学习
数据挖掘
作者
Junjie Shi,Zhaobin Sun,Li Yu,Xin Yang,Zengqiang Yan
标识
DOI:10.1109/tmi.2026.3667605
摘要
Despite the promising potential of multi-modal learning in medical image segmentation, real-world applications often encounter modal incompleteness sourced from diverse domains and institutions, sparking significant discussions on incomplete multi-modal learning. Existing approaches either train a unified model for all or develop individual models for specific multi-modal combinations to ensure model fairness and robustness during inference. However, the assumption of complete multi-modal data for training is unrealistic and infeasible in clinical practice. In this paper, we thoroughly formulate such a challenging setting and propose hierarchical gradient alignment (HGA) to address uni- and multi-modal imbalance. Specifically, gradient direction is aligned through sequential meta learning for multi-modal combinations and multi-level self-distillation for uni-modals within each combination. Gradient magnitude is aligned based on relative preference estimation to balance the dominance of each modal during training. Extensive experiments on five public benchmarks (BraTS2018, BraTS2020, BraTS2023, MyoPS2020, and MSSEG2016) demonstrate that HGA consistently outperforms state-of-the-art incomplete and imbalanced multi-modal learning methods, as well as representative multi-task learning optimization techniques. More importantly, HGA is validated to work as plug-and-play modules for consistent performance improvement across different backbones. Code is available at https://github.com/Jun-Jie-Shi/HGA.
科研通智能强力驱动
Strongly Powered by AbleSci AI