A Vision Transformer network SeedViT for classification of maize seeds

卷积神经网络 人工智能 计算机科学 模式识别(心理学) 支持向量机 分类器(UML) 变压器 人工神经网络 机器学习 深度学习 上下文图像分类 图像(数学) 电压 量子力学 物理
作者
Jiqing Chen,Tian Luo,Jiahua Wu,Zhikui Wang,Hongdu Zhang
出处
期刊:Journal of Food Process Engineering [Wiley]
卷期号:45 (5) 被引量:21
标识
DOI:10.1111/jfpe.13998
摘要

Abstract Maize is a crop that is widely cultivated all over the world. Thus, the classification of maize seeds quality is important, while the traditional methods based on the texture, shape, and color which require repeated work is not efficient. Recently, deep learning reached the goal in the field of image processing, and a deep convolutional neural network (DCNN) is often used to do the image classification task. Here, we explored another neural network called Vision Transformer (ViT), which originally was applied to the natural language processing. Based on the self‐attention mechanism, ViT discards the convolutional structure. But when trained from scratch on medium‐sized datasets, ViT performed poorly compared to CNN. Due to the lack of local structure within the input image, tokenization cannot be used to generate a valid training set in the original ViT model. As a result, we proposed an improved ViT model SeedViT. Compared with the original ViT which could only train large datasets, SeedViT can train small and medium datasets to achieve SOTA (State of the Art) in vision classification with only 2,500 images in our study. The feasibility of SeedViT to classify maize seeds’ quality was studied in this article, and we compared it with DCNN and traditional machine learning algorithms. The accuracy, sensitivity, specificity, and precision were 97.6%, 94.1%, 98.9%, and 97%, respectively. In addition, we employed ViT and VGG (Visual Geometry Group, a convolutional neural network) to extract image features, and SVM (support vector machine) was used as the classifier to classify them, with the result that ViT‐SVM was stable around 96.6% on the test set and VGG‐SVM was stable around 94.6%. At last, a visual attention map was generated by visualization technology. It showed that SeedViT can be a new and novel way for maize seed manufacturing. Practical applications An algorithm for classifying and sorting out high‐quality maize seeds. Use the Transformer algorithm from the field of natural language processing instead of the convolutional neural network algorithm. Use GeLU function for activation and soft split with images to improve Vision Transformer model so that it can achieve great performance with small and medium datasets.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
科研通AI6.2应助yueliang采纳,获得10
1秒前
1秒前
1秒前
1秒前
2秒前
慕青应助花痴的如波采纳,获得10
2秒前
白白发布了新的文献求助10
2秒前
jackson发布了新的文献求助10
2秒前
sinFlee发布了新的文献求助10
2秒前
3秒前
温馨完成签到,获得积分10
4秒前
Hello应助汤圆软软软采纳,获得10
4秒前
充电宝应助汤圆软软软采纳,获得10
4秒前
星辰大海应助汤圆软软软采纳,获得10
4秒前
4秒前
在水一方应助汤圆软软软采纳,获得10
4秒前
NexusExplorer应助汤圆软软软采纳,获得10
4秒前
4秒前
醉熏的老师完成签到 ,获得积分10
4秒前
乐乐应助汤圆软软软采纳,获得10
5秒前
5秒前
酷波er应助汤圆软软软采纳,获得10
5秒前
充电宝应助汤圆软软软采纳,获得10
5秒前
sunny30发布了新的文献求助30
5秒前
njfu完成签到,获得积分10
5秒前
米粒完成签到,获得积分10
5秒前
温馨发布了新的文献求助10
6秒前
团子团子猪完成签到 ,获得积分10
6秒前
7秒前
7秒前
共享精神应助云天河采纳,获得10
8秒前
一7发布了新的文献求助100
8秒前
李爱国应助云曳采纳,获得10
11秒前
霜风款冬发布了新的文献求助10
11秒前
思源应助指甲刀19采纳,获得10
12秒前
陳陳陳发布了新的文献求助10
13秒前
Macs发布了新的文献求助30
13秒前
kuku完成签到,获得积分10
15秒前
Orange应助风趣的绿茶采纳,获得10
15秒前
coco完成签到,获得积分10
15秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Rosenblum, Global Change Biology 800
Essentials of Carbohydrate Chemistry and Biochemistry, 4th Edition 800
Organizational Behavior 510
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
Physiologic specialization in Peronospora manshurica 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 计算机科学 化学工程 工程类 有机化学 物理 复合材料 生物化学 内科学 细胞生物学 基因 遗传学 免疫学 冶金 光电子学 癌症研究
热门帖子
关注 科研通微信公众号,转发送积分 7777088
求助须知:如何正确求助?哪些是违规求助? 9318254
关于积分的说明 20363169
捐赠科研通 7364154
什么是DOI,文献DOI怎么找? 3318840
关于科研通互助平台的介绍 2466494
邀请新用户注册赠送积分活动 2334061