计算机科学
卷积神经网络
人工智能
稳健性(进化)
模式识别(心理学)
数据压缩
水准点(测量)
图像压缩
帧(网络)
视频质量
机器学习
图像处理
图像(数学)
公制(单位)
化学
基因
地理
经济
电信
生物化学
运营管理
大地测量学
作者
Pamela Johnston,Eyad Elyan,Chrisina Jayne
标识
DOI:10.1109/ijcnn.2018.8489370
摘要
A collection of computer vision applications reuse pre-learned features to analyse video frame-by-frame. Those features are classically learned by Convolutional Neural Networks (CNN) trained on high quality images. However, available video content is almost always subject to compression which is nearly never considered during the analysis process. In this paper, we present an empirical study to measure how the visual discrepancy of compressed data limit the learning performance of the CNN model. The learning performance is evaluated using a benchmark of synthetic datasets compressed at various levels using H.264/AVC. We measure the image quality quantitatively using classical evaluation metrics such as Peak Signal to Noise Ratio and Structural SIMilarity. A cross-evaluation is performed to measure the robustness of the CNN model in processing for a wide range of quality-varying visual data. Our experimental results have shown that the performance of the CNN depends on the compression rate. The results show that, in general, higher compression results in lower performance. However performance on lower quality test data can be improved by using lower quality data for CNN training. Finally, our work demonstrates that conditioning the CNN with the compression properties could potentially lead to better learning.
科研通智能强力驱动
Strongly Powered by AbleSci AI