计算机科学
人工智能
MNIST数据库
相似性(几何)
噪音(视频)
模式识别(心理学)
公制(单位)
代表(政治)
像素
比例(比率)
分割
计算机视觉
图像(数学)
深度学习
运营管理
物理
量子力学
政治
政治学
法学
经济
作者
V. S. R. Veeravasarapu,Abhishek Goel,Deepak Mittal,Maneesh Singh
出处
期刊:
日期:2020-06-01
卷期号:: 9668-9676
被引量:5
标识
DOI:10.1109/cvpr42600.2020.00969
摘要
Contour shape alignment is a fundamental but challenging problem in computer vision, especially when the observations are partial, noisy, and largely misaligned. Recent ConvNet-based architectures that were proposed to align image structures tend to fail with contour representation of shapes, mostly due to the use of proximity-insensitive pixel-wise similarity measures as loss functions in their training processes. This work presents a novel ConvNet, "ProAlignNet," that accounts for large scale misalignments and complex transformations between the contour shapes. It infers the warp parameters in a multi-scale fashion with progressively increasing complex transformations over increasing scales. It learns --without supervision-- to align contours, agnostic to noise and missing parts, by training with a novel loss function which is derived an upperbound of a proximity-sensitive and local shape-dependent similarity metric that uses classical Morphological Chamfer Distance Transform. We evaluate the reliability of these proposals on a simulated MNIST noisy contours dataset via some basic sanity check experiments. Next, we demonstrate the effectiveness of the proposed models in two real-world applications of (i) aligning geo-parcel data to aerial image maps and (ii) refining coarsely annotated segmentation labels. In both applications, the proposed models consistently perform superior to state-of-the-art methods.
科研通智能强力驱动
Strongly Powered by AbleSci AI