计算机科学
解码方法
编码(内存)
人工智能
一般化
模式识别(心理学)
代表(政治)
过程(计算)
水准点(测量)
脑电图
特征(语言学)
特征学习
特征提取
视觉感受
机器学习
可视化
任务分析
语义学(计算机科学)
构造(python库)
认知
神经编码
可视对象
认知建筑学
计算机视觉
编码(社会科学)
视觉推理
监督学习
人工神经网络
视皮层
作者
Haodong Jing,Yongqiang Ma,Panqi Yang,Haoyu Li,Shuai Huang,Badong Chen,Nanning Zheng
标识
DOI:10.1109/tip.2026.3666730
摘要
To efficiently assist humans in various tasks, it is crucial to accurately decode and understand the rich information embedded in brain's visual cognition. Existing brain-driven research often fails to overcome the challenge of small target data domains, and the lack of explicit semantic, spatial, and other information constraints on feature extractors prevents brain decoding models from learning uniform cross-domain representations, leading to degradation of their performance in unseen domains. To overcome these limitations, we propose DAMind, a multimodal EEG-based model for robust visual cross-domain alignment and decoding. Our approach integrates VLM with brain-inspired cognitive mechanisms, leveraging the strong image-text representation abilities to learn both fine-grained primary visual features and high-level semantic concepts from neural signals, provide effective visual fine-tuning using the visual guidance mechanism. DAMind introduces a stepwise EEG encoding process aligned with visual processing, and employs an instruction-based learning strategy for effective cross-domain zero-shot transfer. Its robust architecture efficiently achieves good generalization performance, enabling the mapping of EEG signals from multiple domains to a unified learning domain. We construct a comprehensive EEG decoding benchmark EBench, DAMind achieves state-of-the-art results on several visual tasks, and outperforms the baseline in zero-shot setting.
科研通智能强力驱动
Strongly Powered by AbleSci AI