人工智能
卷积神经网络
像素
卷积(计算机科学)
深度学习
模式识别(心理学)
计算机科学
特征提取
残余物
依赖关系(UML)
人工神经网络
计算机视觉
算法
作者
Jinze Song,Zexi Chen,Xianye Li,Xing Wang,Ting Yang,Wenjie Jiang,Baoqing Sun
出处
期刊:Optics Express
[Optica Publishing Group]
日期:2024-08-30
卷期号:32 (20): 34653-34653
被引量:6
摘要
Recent progress in single-pixel imaging (SPI) has exhibited remarkable performance using deep neural networks, e.g., convolutional neural networks (CNNs) and vision Transformers (ViTs). Nonetheless, it is challenging for existing methods to well model object image from single-pixel detections that have a long-range dependency, where CNNs are constrained by their local receptive fields, and ViTs suffer from high quadratic complexity of attention mechanism. Inspired by the Mamba architecture, known for its proficiency in handling long sequences and global contextual information with enhanced computational efficiency as state space models (SSMs), we propose a hybrid network of CNN and Mamba for SPI, named CMSPI. The proposed CMSPI integrates the local feature extraction capability of convolutional layers with the abilities of SSMs for efficiently capturing the long-range dependency, and the design of complementary split-concat structure, depthwise separable convolution, and residual connection enhance learning power of network model. Besides, CMSPI adopts a two-step training strategy, which makes reconstruction performance better and hardware-friendly. Simulations and real experiments demonstrate that CMSPI has higher imaging quality, lower memory consumption, and less computational burden than the state-of-the-art SPI methods.
科研通智能强力驱动
Strongly Powered by AbleSci AI