恶意软件
计算机科学
操作码
人工智能
机器学习
特征提取
特征(语言学)
深度学习
隐病毒学
一般化
模式识别(心理学)
数据挖掘
计算机安全
数学
哲学
语言学
数学分析
计算机硬件
作者
Yetao Jia,Yangyang Meng,Honglin Zhuang
标识
DOI:10.1109/qrs60937.2023.00071
摘要
The use of malware for illicit cyber activities, including network attacks and information theft, poses a severe threat to cybersecurity. In comparison to traditional malware detection methods based on signature and heuristics, machine learning and deep learning-based malware detection methods demonstrate superior generalization ability. However, existing research still faces challenges such as reliance on relatively single malware features, inadequate ability to describe malware features, and overdependence on labeled data. In this paper, we propose an image-based malware classification method using self-supervised and contrastive learning, named IMCSCL. We visualize malware using opcode semantic features, and then detect malware using a contrastive learning method with improved feature encoder network. Experimental results demonstrate that IMCSCL achieves higher detection accuracy compared to supervised malware detection methods, achieving 98.85% accuracy on the Microsoft Malware Classification Challenge dataset. Fine-tuning the model using randomly selected 5% labeled samples from the training set still achieved high accuracy of 94.22%. IMCSCL exhibits superior generalization ability, faster convergence speed, and better training stability. Moreover, contrastive learning significantly reduces malware labeling costs while effectively enhancing detection performance.
科研通智能强力驱动
Strongly Powered by AbleSci AI