Fusion of tactile and visual information in deep learning models for object recognition

计算机科学 人工智能 对象(语法) 可靠性(半导体) 多模式学习 视觉对象识别的认知神经科学 一般化 机器学习 任务(项目管理) 深度学习 人工神经网络 模式 数学分析 管理 功率(物理) 经济 社会学 数学 物理 量子力学 社会科学
作者
Reza Pebdani Babadian,Karim Faez,Mahmood Amiri,Egidio Falotico
出处
期刊:Information Fusion [Elsevier BV]
卷期号:92: 313-325 被引量:57
标识
DOI:10.1016/j.inffus.2022.11.032
摘要

Humans use multimodal sensory information to understand the physical properties of their environment. Intelligent decision-making systems such as the ones used in robotic applications could also utilize the fusion of multimodal information to improve their performance and reliability. In recent years, machine learning and deep learning methods are used at the heart of such intelligent systems. Developing visuo-tactile models is a challenging task due to various problems such as performance, limited datasets, reliability, and computational efficiency. In this research, we propose four efficient models based on dynamic neural network architectures for unimodal and multimodal object recognition. For unimodal object recognition, TactileNet and VisionNet are proposed. For multimodal object recognition, the FusionNet-A and the FusionNet-B are designed to implement early and late fusion strategies, respectively. The proposed models have a flexible structure and are able to change at the train or test phase to accommodate the amount of available information. Model confidence calibration is employed to enhance the reliability and generalization of the models. The proposed models are evaluated on MIT CSAIL large-scale multimodal dataset. Our results demonstrate accurate performance in both unimodal and multimodal scenarios. It has been illustrated that by using different fusion strategies and augmenting the tactile-based models with visual information, the top-1 error rate of the single-frame tactile model was reduced by 78% and the mean average precision was increased by 2.19 times. Although the focus has been on the fusion of tactile and visual modalities, the proposed design methodology can be generalized to include more modalities.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
大模型应助刘真焊采纳,获得10
1秒前
无限萃完成签到,获得积分10
4秒前
8秒前
cy__完成签到,获得积分10
10秒前
刘真焊发布了新的文献求助10
12秒前
新帅完成签到,获得积分10
12秒前
molihuakai应助liyi采纳,获得10
12秒前
成功的强完成签到,获得积分10
14秒前
Tong完成签到 ,获得积分10
14秒前
bkagyin应助Hao采纳,获得30
15秒前
17秒前
飞儿完成签到 ,获得积分10
17秒前
17秒前
月儿完成签到 ,获得积分0
19秒前
古柳完成签到,获得积分10
23秒前
davidli发布了新的文献求助10
24秒前
crazy完成签到 ,获得积分10
29秒前
徐伟业完成签到 ,获得积分10
30秒前
30秒前
谓易ing完成签到 ,获得积分10
31秒前
xelloss完成签到,获得积分10
32秒前
瘦瘦白薇完成签到,获得积分10
33秒前
34秒前
四叶草完成签到 ,获得积分10
37秒前
Sofie发布了新的文献求助10
37秒前
仙女完成签到 ,获得积分10
37秒前
ran完成签到 ,获得积分10
40秒前
45秒前
Sofie完成签到,获得积分10
46秒前
49秒前
壮观的谷冬完成签到 ,获得积分0
49秒前
zouzh完成签到 ,获得积分10
50秒前
独钓寒江雪完成签到 ,获得积分10
52秒前
davidli完成签到,获得积分10
54秒前
orixero应助年轻龙猫采纳,获得10
56秒前
陈少华完成签到 ,获得积分10
56秒前
kk完成签到 ,获得积分10
1分钟前
搞怪白秋完成签到 ,获得积分10
1分钟前
刘真焊发布了新的文献求助10
1分钟前
ll完成签到,获得积分10
1分钟前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Single Cell Analysis of the Tumor Microenvironment Landscape Across the Disease Spectrum of Multiple Myeloma 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场现状调查及投资机会研判报告 1000
2026年中国辛酸癸酸聚乙二醇甘油酯行业市场规模及竞争格局分析报告 1000
模型平均及其应用 900
Fundamentals of Pharmaceutical and Biologics Regulations: A Global Perspective, Second Edition 700
The Cambridge History of China 英文版16册 600
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7331617
求助须知:如何正确求助?哪些是违规求助? 8946001
关于积分的说明 18975356
捐赠科研通 6985875
什么是DOI,文献DOI怎么找? 3216880
关于科研通互助平台的介绍 2383416
邀请新用户注册赠送积分活动 2196531