Encoder and decoder network with ResNet-50 and global average feature pooling for local change detection

计算机科学 编码器 人工智能 联营 背景减法 特征(语言学) 模式识别(心理学) 计算机视觉 RGB颜色模型 帧(网络) 特征向量 像素 电信 语言学 操作系统 哲学
作者
Manoj Kumar Panda,Akhilesh Sharma,Vatsalya Bajpai,Badri Narayan Subudhi,T. Veerakumar,Vinit Jakhetiya
出处
期刊:Computer Vision and Image Understanding [Elsevier BV]
卷期号:222: 103501-103501 被引量:25
标识
DOI:10.1016/j.cviu.2022.103501
摘要

Background subtraction is a prevalent way of dealing with detecting the local changes from video scenes. Background subtraction divides an image frame into foreground and background. The proposed scheme is a unique attempt to detect the local changes in video using a combination of the feature pooling module (FPM) with a ResNet-50 encoder–decoder network. In this context, we proposed a robust encoder–decoder structured deep learning network that is trained with limited training data. The proposed scheme has several folds of novelties including as mentioned below. The use of the feature pooling module with the ResNet-50 encoder–decoder network is the first attempt to use background subtraction in complex video scenes. In the proposed scheme the weights of the ResNet-50 network are learnt by using the transfer learning mechanism. Further, due to the use of a selected number of layers in ResNet-50 architecture with a fewer number of trainable parameters, the proposed architecture becomes less complex as compared to competitive architecture like VGG-16. The proposed ResNet-50 encoder with the FPM module is capable of extracting relevant multi-scale features for local change detection from complex videos. The said encoder uses residual connections between the layers and is hence capable of extracting meaningful multi-scale features with a fewer number of parameters and a higher number of layers. We finally used an up-sampling in the decoder to learn a mapping from the feature space to the image-frame space. The model takes an RGB image frame as the input and generates a foreground segmented probability mask for the corresponding image. To evaluate our model, we have tested it on the three popular benchmark databases. The robustness of the proposed scheme is evaluated by comparing its results with twenty-eight state-of-the-art techniques. The evaluation of the results is carried out using visual and eight quantitative evaluation measures.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
千早爱音发布了新的文献求助100
1秒前
小宝完成签到 ,获得积分10
1秒前
傻傻的初柳完成签到,获得积分10
1秒前
0411345完成签到,获得积分10
2秒前
乐观的哈密瓜完成签到,获得积分10
2秒前
炙热盼兰发布了新的文献求助10
3秒前
sanvva应助悦耳的乐松采纳,获得80
4秒前
5秒前
Akim应助安详的幻梦采纳,获得10
5秒前
小小青完成签到,获得积分10
6秒前
fyc发布了新的文献求助10
6秒前
坚定的问梅完成签到,获得积分10
8秒前
缥缈安荷完成签到,获得积分10
8秒前
橙子味完成签到,获得积分10
9秒前
10秒前
愈久弥新完成签到,获得积分10
10秒前
mayzee完成签到,获得积分10
11秒前
桐桐应助蔡宇滔采纳,获得10
14秒前
hunter完成签到 ,获得积分10
14秒前
大模型应助独特微笑采纳,获得30
14秒前
zzk完成签到,获得积分10
14秒前
15秒前
KangYe完成签到,获得积分10
15秒前
zhui发布了新的文献求助10
16秒前
phoenix完成签到,获得积分10
16秒前
17秒前
NexusExplorer应助勤劳的寄灵采纳,获得10
17秒前
18秒前
赘婿应助小五采纳,获得10
19秒前
容嬷嬷完成签到 ,获得积分10
19秒前
20秒前
20秒前
20秒前
炙热盼兰完成签到 ,获得积分10
22秒前
23秒前
笨小孩发布了新的文献求助10
23秒前
蔡宇滔发布了新的文献求助10
24秒前
wzy发布了新的文献求助30
24秒前
JamesPei应助里旺采纳,获得10
25秒前
丘比特应助我是个唐氏采纳,获得10
25秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Reducing Compassion Fatigue, Secondary Traumatic Stress and Burnout 600
Comparative Elite Sport Development Systems, Structures and Public Policy 600
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
Auslegungsgeschichte 500
Cosmos as Art Object: Studies in Plato's Timaeus and Other Dialogues 500
What is the Future of Psychotherapy in Digital Age? Technology, AI Bots, and Psychotherapy after Covid 444
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7637984
求助须知:如何正确求助?哪些是违规求助? 9211325
关于积分的说明 19758495
捐赠科研通 7204970
什么是DOI,文献DOI怎么找? 3275767
关于科研通互助平台的介绍 2437385
邀请新用户注册赠送积分活动 2272936