Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Blind Image Quality Assessment Using a Deep Bilinear Convolutional Neural Network

作者:Weixia Zhang, Kede Ma, Jia Yan, Dexiang Deng, Zhou Wang · 发表于:IEEE Transactions on Circuits and Systems for Video Technology · 年份:2018 · DOI:10.1109/tcsvt.2018.2886771 · 被引用次数:950 · 研究领域:Image and Video Quality Assessment、Advanced Image Processing Techniques、Image Enhancement Techniques

We propose a deep bilinear model for blind image quality assessment that works for both synthetically and authentically distorted images. Our model constitutes two streams of deep convolutional neural networks (CNNs), specializing in two distortion scenarios separately. For synthetic distortions, we first pre-train a CNN to classify the distortion type and the level of an input image, whose ground truth label is readily available at a large scale. For authentic distortions, we make use of a pre-train CNN (VGG-16) for the image classification task. The two feature sets are bilinearly pooled into one representation for a final quality prediction. We fine-tune the whole network on the target databases using a variant of stochastic gradient descent. The extensive experimental results show that the proposed model achieves state-of-the-art performance on both synthetic and authentic IQA databases. Furthermore, we verify the generalizability of our method on the large-scale Waterloo Exploration Database, and demonstrate its competitiveness using the group maximum differentiation competition methodology.