人工智能
计算机视觉
计算机科学
图像配准
光流
仿射变换
特征(语言学)
图像融合
镜面反射高光
图像(数学)
数学
镜面反射
物理
哲学
量子力学
语言学
纯数学
作者
Liang Li,Evangelos B. Mazomenos,James H. Chandler,Keith L. Obstein,Pietro Valdastri,Danail Stoyanov,Francisco Vasconcelos
标识
DOI:10.1016/j.media.2022.102709
摘要
We propose an endoscopic image mosaicking algorithm that is robust to light conditioning changes, specular reflections, and feature-less scenes. These conditions are especially common in minimally invasive surgery where the light source moves with the camera to dynamically illuminate close range scenes. This makes it difficult for a single image registration method to robustly track camera motion and then generate consistent mosaics of the expanded surgical scene across different and heterogeneous environments. Instead of relying on one specialised feature extractor or image registration method, we propose to fuse different image registration algorithms according to their uncertainties, formulating the problem as affine pose graph optimisation. This allows to combine landmarks, dense intensity registration, and learning-based approaches in a single framework. To demonstrate our application we consider deep learning-based optical flow, hand-crafted features, and intensity-based registration, however, the framework is general and could take as input other sources of motion estimation, including other sensor modalities. We validate the performance of our approach on three datasets with very different characteristics to highlighting its generalisability, demonstrating the advantages of our proposed fusion framework. While each individual registration algorithm eventually fails drastically on certain surgical scenes, the fusion approach flexibly determines which algorithms to use and in which proportion to more robustly obtain consistent mosaics.
科研通智能强力驱动
Strongly Powered by AbleSci AI