计算机科学
编码
编码器
变压器
点云
分割
频道(广播)
人工智能
计算机视觉
模式识别(心理学)
电信
工程类
生物化学
基因
操作系统
电气工程
电压
化学
作者
Guoquan Xu,Hezhi Cao,Yifan Zhang,Yanxin Ma,Jianwei Wan,Ke Xu
标识
DOI:10.1007/978-3-031-15934-3_1
摘要
AbstractTransformer plays an increasingly important role in various computer vision areas and has made remarkable achievements in point cloud analysis. Since existing methods mainly focus on point-wise transformer, an adaptive channel-wise Transformer is proposed in this paper. Specifically, a channel encoding Transformer called Transformer Channel Encoder (TCE) is designed to encode the coordinate channel. It can encode coordinate channels by capturing the potential relationship between coordinates and features. The encoded channel can extract features with stronger representation ability. Compared with simply assigning attention weight to each channel, our method aims to encode the channel adaptively. Moreover, our method can be extended to other frameworks to improve their preformance. Our network adopts the neighborhood search method of feature similarity semantic receptive fields to improve the performance. Extensive experiments show that our method is superior to state-of-the-art point cloud classification and segmentation methods on three benchmark datasets. KeywordsTransformerPoint cloud analysisAdaptive channel encoding
科研通智能强力驱动
Strongly Powered by AbleSci AI