计算机科学
编码器
判别式
分割
人工智能
卷积神经网络
模式识别(心理学)
特征(语言学)
特征提取
变压器
特征学习
特征向量
人工神经网络
保险丝(电气)
图像分割
代表(政治)
桥接(联网)
计算机视觉
组分(热力学)
线性可变差动变压器
作者
Daniel Asefa Beyene,Kassahun Demissie Tola,Fitsum Emagnenehe Yigzew,Shuju Jing,Minsoo Park,Seunghee Park
标识
DOI:10.1016/j.rineng.2026.110145
摘要
• Proposed CrackHCT-Net, a hybrid CNN-Transformer for structural surface crack segmentation • Proposed an IDSC-based gated linear unit to further enhance contextual feature representation • Proposed an efficient feature fusion module to fuse local and global feature information • Validated the model on three public datasets to evaluate the model’s performance • Performed ablation studies to verify the effectiveness of each component Cracks are surface-level structural defects commonly found in built infrastructure at various scales and types, with distinct edges and textures. Additionally, the proportion of the crack surface is extremely small compared to the background surface, making accurate detection is challenging under diverse structural background conditions. Regular monitoring is therefore essential to maintain structural integrity and safety. This task requires a model capable of detecting cracks of varying shapes and backgrounds while remaining computationally efficient for automated anomaly detection. To address this, we propose CrackHCT-Net, a hybrid multi-scale network that combines Convolutional Neural Network (CNN) and Transformer architectures for crack segmentation. The model employs a lightweight CNN encoder to extract local features and a Transformer encoder incorporating lightweight attention and inverted depthwise separable convolution-based gated linear units to capture discriminative global contextual information. A multi-scale feature fusion module is introduced to aggregate features extracted by the CNN and Transformer encoders at the same semantic level while minimizing discrepancies in their feature representations, and reducing redundant features extracted by both encoders. Experiments on three public crack datasets: Crack3238, DeepCrack537, and CFD demonstrate that our method achieves strong segmentation performance, with IoU scores of 63.31%, 76.99%, and 58.33%, respectively.
科研通智能强力驱动
Strongly Powered by AbleSci AI