<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Neurosci.</journal-id>
<journal-title>Frontiers in Neuroscience</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Neurosci.</abbrev-journal-title>
<issn pub-type="epub">1662-453X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fnins.2022.1118087</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Neuroscience</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Subjective and objective quality assessment of gastrointestinal endoscopy images: From manual operation to artificial intelligence</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name><surname>Yuan</surname> <given-names>Peng</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="author-notes" rid="fn002"><sup>&#x02020;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/2129440/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Bai</surname> <given-names>Ruxue</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="author-notes" rid="fn002"><sup>&#x02020;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1457119/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Yan</surname> <given-names>Yan</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="author-notes" rid="fn002"><sup>&#x02020;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/2135648/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Li</surname> <given-names>Shijie</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
</contrib>
<contrib contrib-type="author">
<name><surname>Wang</surname> <given-names>Jing</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
</contrib>
<contrib contrib-type="author">
<name><surname>Cao</surname> <given-names>Changqi</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Wu</surname> <given-names>Qi</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>The Key Laboratory of Carcinogenesis and Translational Research (Ministry of Education), Department of Endoscopy, Peking University Cancer Hospital and Institute</institution>, <addr-line>Beijing</addr-line>, <country>China</country></aff>
<aff id="aff2"><sup>2</sup><institution>Faculty of Information Technology, Beijing University of Technology</institution>, <addr-line>Beijing</addr-line>, <country>China</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Xiongkuo Min, Shanghai Jiao Tong University, China</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Ke Gu, Beijing University of Technology, China; Jiheng Wang, Apple, United States; Chengxu Zhou, Liaoning University of Technology, China</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Qi Wu &#x02709; <email>wuqi1973&#x00040;bjmu.edu.cn</email></corresp>
<fn fn-type="other" id="fn001"><p>This article was submitted to Perception Science, a section of the journal Frontiers in Neuroscience</p></fn>
<fn fn-type="equal" id="fn002"><p>&#x02020;These authors have contributed equally to this work and share first authorship</p></fn></author-notes>
<pub-date pub-type="epub">
<day>14</day>
<month>02</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>16</volume>
<elocation-id>1118087</elocation-id>
<history>
<date date-type="received">
<day>07</day>
<month>12</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>30</day>
<month>12</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2023 Yuan, Bai, Yan, Li, Wang, Cao and Wu.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Yuan, Bai, Yan, Li, Wang, Cao and Wu</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license></permissions>
<abstract>
<p>Gastrointestinal endoscopy has been identified as an important tool for cancer diagnosis and therapy, particularly for treating patients with early gastric cancer (EGC). It is well known that the quality of gastroscope images is a prerequisite for achieving a high detection rate of gastrointestinal lesions. Owing to manual operation of gastroscope detection, in practice, it possibly introduces motion blur and produces low-quality gastroscope images during the imaging process. Hence, the quality assessment of gastroscope images is the key process in the detection of gastrointestinal endoscopy. In this study, we first present a novel gastroscope image motion blur (GIMB) database that includes 1,050 images generated by imposing 15 distortion levels of motion blur on 70 lossless images and the associated subjective scores produced with the manual operation of 15 viewers. Then, we design a new artificial intelligence (AI)-based gastroscope image quality evaluator (GIQE) that leverages the newly proposed semi-full combination subspace to learn multiple kinds of human visual system (HVS) inspired features for providing objective quality scores. The results of experiments conducted on the GIMB database confirm that the proposed GIQE showed more effective performance compared with its state-of-the-art peers.</p></abstract>
<kwd-group>
<kwd>gastroscope images</kwd>
<kwd>motion blur</kwd>
<kwd>subjective and objective quality assessment</kwd>
<kwd>human visual system</kwd>
<kwd>semi-full combination subspace</kwd>
</kwd-group>
<counts>
<fig-count count="8"/>
<table-count count="1"/>
<equation-count count="21"/>
<ref-count count="52"/>
<page-count count="13"/>
<word-count count="8688"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>1. Introduction</title>
<p>Gastric cancer (GC) is the major cause of cancer death worldwide (Chen et al., <xref ref-type="bibr" rid="B3">2022</xref>). Recently, gastrointestinal endoscopy has been identified as an important tool for cancer diagnosis and therapy, particularly for treating patients with early gastric cancer (EGC) (Li Y.-D. et al., <xref ref-type="bibr" rid="B24">2021</xref>). A proper application of endoscopy could identify and treat gastric lesions better. The main purpose of medical image processing and analysis is to facilitate physicians to conduct diagnosis and therapy (Cai et al., <xref ref-type="bibr" rid="B1">2021</xref>; Xu et al., <xref ref-type="bibr" rid="B44">2022</xref>). It is well known that the quality of gastroscope images is a prerequisite for achieving a high detection rate of gastrointestinal lesions (Liu et al., <xref ref-type="bibr" rid="B26">2021</xref>). As gastroscope detection is operated manually, in practice, it possibly introduces motion blur and produces low-quality gastroscope images during the imaging process. These poor quality gastroscope images could lead to misdiagnosis, and thus patients must need a second examination that increases their pain one more time and even worse makes them miss the best time for treatment. Therefore, the image quality assessment (IQA) of gastroscope images is helpful to lead to more accurate and earlier detection, helping further in the development of image deblurring, enhancement, fusion, and denoising (Chen et al., <xref ref-type="bibr" rid="B4">2021</xref>; Qin et al., <xref ref-type="bibr" rid="B30">2021</xref>). To sum up, a good IQA method of gastroscope images is very important to determine lesions effectively.</p>
<p>In the field of image processing and computer vision, IQA is a crucial topic of research topic (Ye X. et al., <xref ref-type="bibr" rid="B47">2020</xref>; Sun et al., <xref ref-type="bibr" rid="B37">2021</xref>), including the subjective assessment and the objective assessment. The subjective assessment is widely perceived to be the most accurate IQA method because the measuring results of its image quality as the mean opinion score (MOS) are provided by human viewers. A few well-known and publicly available IQA databases with MOS or differential MOS (DMOS), such as Tampere Image Database 2013 (TID2013) (Ponomarenko et al., <xref ref-type="bibr" rid="B29">2013</xref>), Categorical image quality (CSIQ) (Larson and Chandler, <xref ref-type="bibr" rid="B22">2010</xref>), and Laboratory for Image and Video Engineering (LIVE) (Sheikh et al., <xref ref-type="bibr" rid="B35">2006</xref>), pave the way for the development of the IQA. Over the past decade, many scholars have built several IQA databases for more practical purposes. For example, a contrast-changed image database (CCID2014) was included in Gu et al. (<xref ref-type="bibr" rid="B13">2015b</xref>) to enable a study on the perceptual quality of images with contrast changes. Two tone mapping image databases were presented in Kundu et al. (<xref ref-type="bibr" rid="B21">2017</xref>) and Gu et al. (<xref ref-type="bibr" rid="B12">2016b</xref>) to facilitate research on the quality of evaluation of tone-mapped images with a high dynamic range (HDR). The IQA database for super-resolved images was designed in Fei (<xref ref-type="bibr" rid="B6">2020</xref>) for assessing the visual quality of super-resolution images. However, well-known IQA databases are improper in the case of gastroscope images. Specifically, there is no specific subjective IQA database of gastroscope images. Because a gastroscope is placed inside the body, many types of distortion in these databases, such as impulse noise, brightness change, and Joint Photographic Experts Group (JPEG) compression, are not included in gastroscope images, but a motion blur usually exists. Up to now, to the best of our knowledge, there has been not a publicly available database for the quality assessment of gastroscope images, so it is highly necessary to establish an IQA database of distorted gastroscope images.</p>
<p>The MOS values are obtained experiments that include different individuals and circumstances, but which are improper for the real-time IQA of gastroscope images. The MOS is obtained in a labor-intensive and time-consuming process, and is thus of very low reusability. Another strategy to evaluate the quality of images which is highly demanding is to develop objective assessment methods toward matching the characteristics of a human vision system (HVS). Recently, objective IQA method have achieved good results. The typical objective IQA metrics are based on the full-reference (FR), where a &#x0201C;clean&#x0201D; gastroscope image is available. The &#x0201C;clean&#x0201D; gastroscope image is the ground truth in the case of gastroscope images distorted with motion blur. The visual signal-to-noise ratio (VSNR) Chandler and Hemami (<xref ref-type="bibr" rid="B2">2007</xref>) takes advantage of the supra- and near-threshold characteristics of human vision. The peak signal-to-noise ratio (PSNR) and mean-squared errors (MSEs) are the most popular and commonly used FR IQA techniques, but their correlation with perceived quality is not ideal. The most apparent distortion (MAD) Larson and Chandler (<xref ref-type="bibr" rid="B22">2010</xref>) method adaptively extracts visual features from the reference and distorted images using the log-Gabor filtering and Fourier transform. The structural similarity (SSIM) Wang et al. (<xref ref-type="bibr" rid="B39">2004</xref>) compares three visual aspects including contrast, luminance, and structure. Later on, many variants were proposed, based on the SSIM (Wang et al., <xref ref-type="bibr" rid="B41">2003</xref>; Sampat et al., <xref ref-type="bibr" rid="B32">2009</xref>; Wang and Li, <xref ref-type="bibr" rid="B40">2010</xref>; Zhu et al., <xref ref-type="bibr" rid="B52">2018</xref>).</p>
<p>The FR IQA methods also make use of many other cues or features, except for covariance, variance, and mean. Mutual information between the distorted and the lossless images is used to evaluate the quality of visual perception in the information fidelity criterion (IFC) (Sheikh et al., <xref ref-type="bibr" rid="B34">2005</xref>) and its extended approach named the visual information fidelity (VIF) (Sheikh and Bovik, <xref ref-type="bibr" rid="B33">2006</xref>). In addition, since it is known that image gradients contain many types of significant visual information, some IQA approaches extract the gradient features. In Zhang et al. (<xref ref-type="bibr" rid="B50">2011</xref>), the feature similarity (FSIM) was proposed to incorporate gradient magnitudes with phase congruency. In Liu et al. (<xref ref-type="bibr" rid="B25">2011</xref>), the gradient similarity (GSIM) was developed by combining gradient features with masking effect and distortion visibility. In Xue et al. (<xref ref-type="bibr" rid="B45">2013</xref>), the gradient magnitude similarity deviation (GMSD) takes advantage of a new pooling strategy that is the global variation of a local gradient similarity. Both the pooling weights and local features represent visual saliency of the image in the IQA (Zhang et al., <xref ref-type="bibr" rid="B49">2014</xref>; Ye Y. et al., <xref ref-type="bibr" rid="B46">2020</xref>). A few existing IQA models utilize the predictability as a feature. The different strategies of the unpredicted and predicted parts in an image are employed in Wu et al. (<xref ref-type="bibr" rid="B42">2012</xref>) to measure the internal generative mechanism (IGM) index.</p>
<p>However, the scope of application of FR IQA is constrained by the dependence of lossless images. In recent years, the no-reference (NR) IQA models have been emphatically developed to solve the problem of the original image not being available in many cases (Hu et al., <xref ref-type="bibr" rid="B19">2021</xref>; Li T. et al., <xref ref-type="bibr" rid="B23">2021</xref>; Pan et al., <xref ref-type="bibr" rid="B28">2021</xref>; ur Rehman et al., <xref ref-type="bibr" rid="B38">2022</xref>). In Gu et al. (<xref ref-type="bibr" rid="B10">2017c</xref>), the authors extracted 17 features including brightness, sharpness, contrast, and so on, and then achieved a predictive quality score by a regression model. In Gu et al. (<xref ref-type="bibr" rid="B18">2017d</xref>), the authors developed a novel blind IQA model for evaluating the perceptual quality of screen content images with big data learning. In Gu et al. (<xref ref-type="bibr" rid="B16">2014b</xref>), the authors proposed a new blind IQA model using the classical HVS features and the free energy feature based on the image processing and brain theory. In Gu et al. (<xref ref-type="bibr" rid="B14">2015c</xref>), the authors designed an NR sharpness IQA metric that is built using the analysis of autoregressive (AR) parameters. However, some distortion types, such as motion blur that may appear in the gastroscope images, are not considered in the majority of the existing IQA methods, so these off-the-shelf methods do not suit gastroscope images the best.</p>
<p>In this study, we attempt to construct a novel image database and a specific IQA metric of gastroscope images to identify and treat gastric lesions better. Because motion blur easily takes place in a gastroscope image during the imaging process, we focus mainly on how it affects the quality of a gastroscope image. First, we build a gastroscope image motion blur (GIMB) database that encompasses 70 source images from 27 categories of the upper endoscopy anatomy is built and 1,050 corresponding motion blurred images derived from five pixel levels for three different motion angles. We adopt the single stimulus (SS) method to gather subjective ratings. Then, we properly integrate the existing FR IQA methods (Wu et al., <xref ref-type="bibr" rid="B43">2021</xref>) to design an artificial intelligence (AI)-based gastroscope image quality evaluator (GIQE). To define it more concretely, we learn multiple kinds of HVS inspired features from gastroscope motion blurred images by the newly proposed semi-full combination subspace. The results reveal that the proposed GIQE can achieve a superior performance relative to the state-of-the-art FR IQA metrics.</p>
<p>The remainder of this article is arranged as follows. In Section 2, the subjective assessment of gastroscope images and the establishment of the relevant GIMB database are introduced in detail. In Section 3, a detailed implementation of the proposed GIQE is presented. In Section 4, a comparison of the proposed GIQE with several mainstream FR IQA metrics is carried out using the GIMB database. In Section 5, some conclusions are finally drawn.</p>
</sec>
<sec id="s2">
<title>2. GIMB image database</title>
<p>In this section, we describe the proposed GIMB database. First, we introduce the formation and processing of source images. Then, the subjective methodology is leveraged to collect the MOS values from the viewers. Finally, the collected values of MOS are processed and analyzed.</p>
<sec>
<title>2.1. The formation of source images</title>
<p>It is nontrivial to select source images, because the content of source images has a strong effect on the IQA. According to the general theory, the source images ought to be undistorted, and their contents should be abundant and diverse. The GIMB database encompasses 70 source images that are taken from 27 categories of the upper endoscopy anatomy, such as the antrum anterior wall, the pharynx, the pylorus, and the fundus, as shown in <xref ref-type="fig" rid="F1">Figure 1</xref>. In this study, the patients were examined by gastroscopy at the Peking University Cancer Hospital from June 2020 to December 2021. The Ethics Committee approved the study at the Peking University Cancer Hospital on 15 May 2020 (ethics board protocol number 2020KT60). The source images were captured by endoscopes such as GIF-H290, GIF-HQ290, GIF-H260 (Olympus, Japan), EG-760Z, EG-760R, EG-L600ZW7, EGL600WR7, and EG-580R7 (Fujifilm, Japan). Areas around gastroscope images contain information on indicators that does not contribute to the IQA and should therefore be removed. We cropped the source images into the same resolution of 1,075 &#x000D7; 935 to remove unnecessary information and obtain a higher processing level of IQA.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>The source images in different gastrointestinal tract regions in the gastroscope image motion blur (GIMB) database.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0001.tif"/>
</fig>
</sec>
<sec>
<title>2.2. The processing of source images</title>
<p>From the perspective of an IQA database, the gastroscope blurred images are actually the images distorted by motion blur. The relative motion between the gastrointestinal tract regions and the probe during gastroscopy by artificial operation often leads to motion blur in gastroscope images. The motion blur is caused by the superposition of multiple images at different times. We set <italic>x</italic><sub>0</sub>(<italic>t</italic>) and <italic>y</italic><sub>0</sub>(<italic>t</italic>) as the motion components in <italic>x</italic> and <italic>y</italic>, and set <italic>T</italic> as the exposure time. The vague image adopted at time <italic>t</italic> is</p>
<disp-formula id="E1"><label>(1)</label><mml:math id="M1"><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:mrow><mml:msubsup><mml:mo>&#x0222B;</mml:mo><mml:mn>0</mml:mn><mml:mi>T</mml:mi></mml:msubsup><mml:mi>f</mml:mi></mml:mrow></mml:mstyle><mml:mo stretchy='false'>[</mml:mo><mml:mi>x</mml:mi><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mn>0</mml:mn></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>y</mml:mi><mml:mn>0</mml:mn></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>]</mml:mo><mml:mi>d</mml:mi><mml:mi>t</mml:mi><mml:mo>.</mml:mo></mml:math></disp-formula>
<p>We suppose that the motion between the gastrointestinal tract regions and the probe is a kind of uniform rectilinear movement. During time <italic>T</italic>, the moving distances are represented by <italic>a</italic> and <italic>b</italic> in <italic>x</italic> and <italic>y</italic>:</p>
<disp-formula id="E2"><label>(2)</label><mml:math id="M2"><mml:mrow><mml:mo>{</mml:mo><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mn>0</mml:mn></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mo>/</mml:mo><mml:mi>T</mml:mi></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>y</mml:mi><mml:mn>0</mml:mn></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mi>b</mml:mi><mml:mi>t</mml:mi><mml:mo>/</mml:mo><mml:mi>T</mml:mi></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mrow></mml:math></disp-formula>
<p>Combining with Equations (1), (2), the probe moves <italic>L</italic> pixels with uniform speed in a straight line at &#x003B8; angle in the <italic>x</italic>-<italic>y</italic> plane. The vague image is obtained by</p>
<disp-formula id="E3"><label>(3)</label><mml:math id="M3"><mml:mrow><mml:mtable columnalign='left'><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>L</mml:mi></mml:mfrac><mml:mstyle displaystyle='true'><mml:munderover><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>0</mml:mn></mml:mrow><mml:mrow><mml:mi>L</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:munderover><mml:mrow><mml:mi>f</mml:mi><mml:mo stretchy='false'>[</mml:mo><mml:msup><mml:mi>x</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo>&#x02212;</mml:mo><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:msup><mml:mi>y</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>]</mml:mo></mml:mrow></mml:mstyle></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>Where <italic>x</italic>&#x02032; &#x0003D; <italic>x</italic> cos&#x003B8; &#x0002B; <italic>y</italic> sin&#x003B8; and <italic>y</italic>&#x02032; &#x0003D; <italic>y</italic> cos&#x003B8; &#x02212; <italic>x</italic> sin&#x003B8;. <italic>i</italic> &#x02208; {1, 2, 3,..., <italic>L</italic>-1} is an integer.</p>
<p>Therefore, we define the point spread function (PSF) of the motion blurred image in any direction by</p>
<disp-formula id="E4"><label>(4)</label><mml:math id="M4"><mml:mrow><mml:mi>h</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mn>1</mml:mn><mml:mo>/</mml:mo><mml:mi>L</mml:mi><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mi>y</mml:mi><mml:mo>=</mml:mo><mml:mi>x</mml:mi><mml:mi>tan</mml:mi><mml:mi>&#x003B8;</mml:mi><mml:mo>,</mml:mo><mml:mn>0</mml:mn><mml:mo>&#x02264;</mml:mo><mml:mi>x</mml:mi><mml:mo>&#x02264;</mml:mo><mml:mi>L</mml:mi><mml:mi>cos</mml:mi><mml:mi>&#x003B8;</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mi>y</mml:mi><mml:mo>&#x02260;</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mi>tan</mml:mi><mml:mi>&#x003B8;</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>,</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>&#x0221E;</mml:mi><mml:mo>&#x0003C;</mml:mo><mml:mi>x</mml:mi><mml:mo>&#x0003C;</mml:mo><mml:mi>&#x0221E;</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mrow></mml:math></disp-formula>
<p>Two important parameters include the direction of the motion blur &#x003B8; and the distance from where the pixels <italic>L</italic> have blurred.</p>
<p>To obtain motion blurred images, we processed source images using the built-in function of MATLAB application. To be more specific, we used two key parameters, <italic>L</italic> and &#x003B8;, of motion blur aforementioned to process each lossless image. We set the direction of the motion blur to be at three different motion angles &#x003B8; &#x0003D; {30&#x000B0;, 60&#x000B0;, and 90&#x000B0;}. Because gastroscope images are different from natural images, their rotations have no impact on the diagnosis of doctors. In addition, we set the motion distance to be five pixel levels, that is, <italic>L</italic> &#x0003D; {5, 10, 15, 20, 25}, which directly affect the performance of the IQA and the detection rate of gastric lesions. <xref ref-type="fig" rid="F2">Figure 2</xref> shows five motion blur levels of a lossless image. For the five motion blur grades, doctors agree that <italic>L</italic> &#x0003D; {5, 10} is useful for diagnosis, while <italic>L</italic> &#x0003D; {5, 10} corresponds to poor quality gastroscope images, potentially contributing to the misdiagnosis. Moreover, <italic>L</italic> &#x0003D; 15 is the boundary between the availability and the unavailability as confirmed by most of the physicians. On this basis, we generated 15 motion blurred images from each source image. Overall, the proposed GIMB database contains 70 lossless images and 1,050 distorted images.</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p>The five motion blur levels of a lossless image <bold>(A)</bold>. <bold>(B&#x02013;E)</bold> Are the motion blurred image with five levels corresponding to lossless image. <bold>(D)</bold> Shows the boundary between available <bold>(B, C)</bold> and unavailable <bold>(E, F)</bold>.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0002.tif"/>
</fig>
</sec>
<sec>
<title>2.3. Subjective methodology</title>
<p>Subjective methodology is an important procedure in creating an IQA database, yet it is very labor-intensive and time-consuming. In the following, we present the subjective test method, subject, environment, and the apparatus.</p>
<sec>
<title>2.3.1. Method</title>
<p>The methodology for the subjective assessment of the quality of television pictures. Recommendation ITU-R BT.500-13 (Ritur, <xref ref-type="bibr" rid="B31">2002</xref>) has defined several subjective test methods that include SS, double-stimulus impairment scale (DSIS), and paired comparison. In this study, we used the SS method to conduct the subjective experiment. The order of all test images on the database was randomized to minimize the impact of subjects&#x00027; memories on MOS. The subjects were asked to score the quality of each gastroscope image from 1 to 5, according to their overall sensation to these images. The test was divided into four subsessions, each of which lasted &#x0003C;20 min. A subsession includes 18 min for scoring and 2 min for training, and the interval for each subsession lasts 5 min.</p>
</sec>
<sec>
<title>2.3.2. Subject</title>
<p>This subjective experiment involves experienced and inexperienced viewers, most of whom are physicians and postgraduates from the medical specialty. The inexperienced subjects are ignorant about distorted images and the corresponding terminology. Specific visual acuity tests including vision and color are not needed since the gastroscope image is a classical two-dimensional (2D) image. The subjects could wear their own glasses with suitable degree they wear every day. Before the test, we gave the viewers oral and written instructions, as specified in the International Telecommunication Union Telecommunication Standardization Sector (ITU-T) Recommendation. P.910. In the training phase, each subject is shown different pixel levels of motion blur, from the lowest to the highest as given in <xref ref-type="fig" rid="F3">Figure 3</xref>, and familiar with the scoring procedure. The images used in the training stage and testing stage are different.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>The interface window applied in the subjective assessment.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0003.tif"/>
</fig></sec>
<sec>
<title>2.3.3. Environment</title>
<p>To achieve reliable scoring results, we conducted the test in a fixed and controlled environment. Specifically, all the viewers were asked to perform their assessment in an indoor environment without any background light (Huang et al., <xref ref-type="bibr" rid="B20">2021</xref>; Shi et al., <xref ref-type="bibr" rid="B36">2021</xref>). We chose the suitable ambient luminance (Gu et al., <xref ref-type="bibr" rid="B8">2015a</xref>; Yu and Akita, <xref ref-type="bibr" rid="B48">2020</xref>). In the training phase, the viewing distance was set to approximately three times the image height. To get more precise scores, the viewers were able to modify the distance between the monitor and themselves slightly after a round testing.</p>
</sec>
<sec>
<title>2.3.4. Apparatus</title>
<p>Two interface windows are shown simultaneously in MATLAB application and are applied to subjective assessment, as illustrated in <xref ref-type="fig" rid="F3">Figure 3</xref>. The left window is used to score, while the right window is used to show the gastroscope motion gastroscope motion blurred image. The right window can be controlled by subjects during the test. The subjects can control which images should be shown in this window by pressing the key &#x0201C;c&#x0201D; or &#x0201C;d&#x0201D; on the keyboard. According to the psychovisual evaluation, we found that the viewers can make their decisions much more precisely and quickly by flipping the images at exactly the same position. Information about the psychovisual evaluation in detail is given in online materials section in Zhou et al. (<xref ref-type="bibr" rid="B51">2018</xref>). During scoring, the subjects were asked to give their scores as early as possible and were guided to click the button &#x0201C;1,&#x0201D; &#x0201C;2,&#x0201D; &#x0201C;3,&#x0201D; &#x0201C;4,&#x0201D; or &#x0201C;5&#x0201D; on the window, indicating their grading of motion blur from the lowest to the highest. The characteristics of the display device and system used in the experiment are described briefly in <xref ref-type="fig" rid="F4">Figure 4</xref>. We saved the final scores of all the gastroscope images given by all the viewers after the subjective test for further analysis.</p>
<fig id="F4" position="float">
<label>Figure 4</label>
<caption><p>The characteristics of the display device and system are used in the experiment.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0004.tif"/>
</fig>
</sec>
</sec>
<sec>
<title>2.4. Scores processing and analysis</title>
<p>According to the subjective test aforementioned, we gathered the viewers&#x00027; scores to be processed and analyzed as follows:</p>
<p>First, we analyzed 15 motion blurred images of a lossless image using the box plot to study the influence of inattentive subjects on an individual observer&#x00027;s rating. The box plot (i.e., box and whisker diagram) is used to analyze the distribution of data on the basis of five indicators, including minimum, maximum, median, and the 25th and 75th percentiles. The range of the first and third quartiles is obvious, and there are a few points that are outliers, as shown in <xref ref-type="fig" rid="F5">Figure 5</xref>. These results indicate that it is worthwhile to analyze the subjective score of an individual participant. Thus, we invited an experienced physician to screen the outliers.</p>
<fig id="F5" position="float">
<label>Figure 5</label>
<caption><p>Box plot of the subjects&#x00027; scores for 15 motion blurred images of a lossless image. On each image box, the central red line is the median score, the edges of the box are first quartile and third quartile, and the outliers are marked by a red cross individually.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0005.tif"/>
</fig>
<p>Then, we processed all the values within the normal range after elimination. We assigned <italic>m</italic><sub><italic>ij</italic></sub> as the raw subjective score obtained from the viewer&#x00027;s <italic>i</italic> evaluation of the gastroscope motion blurred image <italic>I</italic><sub><italic>j</italic></sub>, where <italic>i</italic> &#x0003D; {1, 2, 3, &#x02026;, 15}, <italic>j</italic> &#x0003D; {1, 2, 3, &#x02026;, <italic>N</italic>}, <italic>N</italic> &#x0003C; 1, 050. For a <italic>j</italic><sup><italic>th</italic></sup> image, the MOS value is calculated by the formula as follows:</p>
<disp-formula id="E5"><label>(5)</label><mml:math id="M5"><mml:mrow><mml:mtable columnalign='left'><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mrow><mml:mtext>MOS</mml:mtext></mml:mrow><mml:mi>j</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munderover><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>M</mml:mi></mml:munderover><mml:mrow><mml:mfrac><mml:mrow><mml:msub><mml:mi>m</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>j</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mi>M</mml:mi></mml:mfrac></mml:mrow></mml:mstyle></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>Where <italic>M</italic> &#x0003D; 15 represents the number of subjects. We draw on the distribution histogram of the MOS to display the viewers&#x00027; MOS scores as illustrated in <xref ref-type="fig" rid="F6">Figure 6</xref>. An important observation indicates that the MOS scores of most distorted gastroscope images are only around 3.5 in comparison. Hence, motion blur influences the original gastroscope images considerably, which leads to a misdiagnosis of GC.</p>
<fig id="F6" position="float">
<label>Figure 6</label>
<caption><p>Histogram of mean opinion score (MOS) values for gastroscope motion blurred images.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0006.tif"/>
</fig></sec></sec>
<sec id="s3">
<title>3. Proposed IQA metric</title>
<p>The existing FR and NR IQA models designed for a specific distortion category and an application scenario perform well, but are not suitable for gastroscope images. We explore an IQA metric for gastroscope motion blurred images using the semi-full combination subspace. Specifically, the image quality evaluation method of a gastroscope image is carried out in three steps.</p>
<sec>
<title>3.1. The first step</title>
<p>We extract five features of gastroscope images since the processing of IQA is to learn multiple kinds of HVS inspired features. We then fuse these features to train a regression module.</p>
<sec>
<title>3.1.1. The low-level similarity feature</title>
<p>The phase congruency (PC) principle postulates that the Fourier transform phase contains maximal perceptual information, which helps the HVS to detect and identify features, according to the psychophysical and physiological evidence. Hence, we compute the feature <italic>F</italic><sub><italic>PC</italic></sub>:</p>
<disp-formula id="E6"><label>(6)</label><mml:math id="M6"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>P</mml:mi><mml:mi>C</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo class="qopname">max</mml:mo></mml:mrow><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003D5;</mml:mi></mml:mrow><mml:mo class="qopname">&#x00304;</mml:mo></mml:mover><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>2</mml:mn><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow></mml:munder></mml:mstyle><mml:mrow><mml:mo stretchy="false">{</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:msub><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003BC;</mml:mi></mml:mrow></mml:msub><mml:mo>|</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003BC;</mml:mi></mml:mrow></mml:msub><mml:mo>|</mml:mo><mml:mo class="qopname">cos</mml:mo><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003D5;</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003BC;</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:mi>&#x003D5;</mml:mi></mml:mrow><mml:mo class="qopname">&#x00304;</mml:mo></mml:mover><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003BC;</mml:mi></mml:mrow></mml:msub><mml:mo>|</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003BC;</mml:mi></mml:mrow></mml:msub><mml:mo>|</mml:mo></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">}</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where |<italic>F</italic><sub>&#x003BC;</sub>| is the amplitude of an image and &#x003D5;<sub>&#x003BC;</sub>(<italic>i</italic>) represents the phase of <italic>F</italic><sub>&#x003BC;</sub> at pixel <italic>i</italic> on the scale &#x003BC;.</p>
<p>The gradient magnitude (GM) is a very classical and valid feature for improving the IQA performance. We employ the Scharr operator defined as <inline-formula><mml:math id="M7"><mml:mi>G</mml:mi><mml:mi>M</mml:mi><mml:mo>=</mml:mo><mml:msqrt><mml:mrow><mml:mi>G</mml:mi><mml:msubsup><mml:mrow><mml:mi>M</mml:mi></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>&#x0002B;</mml:mo><mml:mi>G</mml:mi><mml:msubsup><mml:mrow><mml:mi>M</mml:mi></mml:mrow><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:mrow></mml:msqrt></mml:math></inline-formula>, where <italic>GM</italic><sub><italic>x</italic></sub> and <italic>GM</italic><sub><italic>y</italic></sub> are the partial derivatives along <italic>x</italic> and <italic>y</italic> axis directions. This GM is regarded as</p>
<disp-formula id="E7"><label>(7)</label><mml:math id="M8"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>G</mml:mi><mml:mi>M</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>E</mml:mi><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mn>2</mml:mn><mml:mi>G</mml:mi><mml:mi>M</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>G</mml:mi><mml:mi>M</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>G</mml:mi><mml:mi>M</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:mi>G</mml:mi><mml:mi>M</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where <italic>I</italic><sub><italic>r</italic></sub> is the reference image. <italic>I</italic><sub><italic>d</italic></sub> and <italic>I</italic><sub><italic>p</italic></sub> represent the distorted and predicted versions of <italic>I</italic><sub><italic>r</italic></sub>, respectively. <italic>A</italic><sub>1</sub> is a fixed positive constant. Many recently proposed IQA algorithms have proven that PC and GM are very valid, since the HVS is very sensitive to them.</p>
<p>We then combine <italic>f</italic><sub><italic>PC</italic></sub> with <italic>f</italic><sub><italic>GM</italic></sub> to obtain the similarity feature <italic>F</italic><sub><italic>L</italic></sub>, which is defined by</p>
<disp-formula id="E8"><label>(8)</label><mml:math id="M9"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>L</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msubsup><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>P</mml:mi><mml:mi>C</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B1;</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x000B7;</mml:mo><mml:msubsup><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>G</mml:mi><mml:mi>M</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow></mml:msubsup></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where parameters &#x003B1; and &#x003B2; are applied to change the importance of <italic>f</italic><sub><italic>GM</italic></sub> and <italic>f</italic><sub><italic>PC</italic></sub>. Since the visual cortex is very sensitive to PC features, we use <italic>F</italic><sub><italic>PC</italic></sub> as a weight value to extract the low-level similarity feature <italic>F</italic><sub><italic>LSF</italic></sub> as the first feature:</p>
<disp-formula id="E9"><label>(9)</label><mml:math id="M10"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>L</mml:mi><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>L</mml:mi></mml:mrow></mml:msub><mml:mo>&#x000B7;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>P</mml:mi><mml:mi>C</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>P</mml:mi><mml:mi>C</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
</sec>
<sec>
<title>3.1.2. The visual saliency feature</title>
<p>Visual saliency (VS) areas of an image attract maximum attention of the HVS. We fuse VS, GM, and chrominance features to obtain the visual saliency feature (VSF) of images for IQA tasks. We extract VS maps of original and lossless images by a specific VS model. The similarity between them is defined as:</p>
<disp-formula id="E10"><label>(10)</label><mml:math id="M11"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>V</mml:mi><mml:mi>S</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>E</mml:mi><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mn>2</mml:mn><mml:mi>V</mml:mi><mml:mi>S</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>V</mml:mi><mml:mi>S</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>V</mml:mi><mml:mi>S</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:mi>V</mml:mi><mml:mi>S</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>The similarity between the chrominance featue components is simply defined as:</p>
<disp-formula id="E11"><label>(11)</label><mml:math id="M12"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>C</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>E</mml:mi><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mn>2</mml:mn><mml:mi>M</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>M</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>M</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:mi>M</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>E</mml:mi><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mn>2</mml:mn><mml:mi>N</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>N</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mi>N</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:mi>N</mml:mi><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where parameters <italic>M</italic> and <italic>N</italic> are the numbers of channels. <italic>A</italic><sub>2</sub> is another fixed positive constant.</p>
<p>We define <italic>F</italic><sub><italic>VSF</italic></sub> as the second feature:</p>
<disp-formula id="E12"><label>(12)</label><mml:math id="M13"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>V</mml:mi><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>V</mml:mi><mml:mi>S</mml:mi></mml:mrow></mml:msub><mml:mo>&#x000B7;</mml:mo><mml:msubsup><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>G</mml:mi><mml:mi>M</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B1;</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x000B7;</mml:mo><mml:msubsup><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>C</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow></mml:msubsup></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where two parameters &#x003B1; and &#x003B2; are used to adjust the relative importance of VS, GM, and chrominance features.</p>
</sec>
<sec>
<title>3.1.3. The log-gabor filter</title>
<p>The log-Gabor filter (LGF) has strong robustness for brightness and contrast changes of images, and it has been widely used to extract local features and texture analysis in computer vision. We use a log-Gabor filter bank to decompose the source and lossless images into a set of subbands. The subband&#x00027;s features are obtained by the inverse density functional theory (DFT) of the images&#x00027; DFT with the following multiplying 2D frequency response as the third feature:</p>
<disp-formula id="E13"><label>(13)</label><mml:math id="M14"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>L</mml:mi><mml:mi>G</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B8;</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mo class="qopname">exp</mml:mo><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>g</mml:mi><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub><mml:mo>/</mml:mo><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mn>2</mml:mn><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>g</mml:mi><mml:msub><mml:mrow><mml:mi>&#x003C3;</mml:mi></mml:mrow><mml:mrow><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mo>/</mml:mo><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow><mml:mo>&#x000D7;</mml:mo><mml:mo class="qopname">exp</mml:mo><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B8;</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003BC;</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mn>2</mml:mn><mml:msubsup><mml:mrow><mml:mi>&#x003C3;</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where <italic>F</italic><sub><italic>LGF</italic></sub>(<italic>f</italic><sub><italic>r</italic></sub>, <italic>f</italic><sub>&#x003B8;</sub>) is a log-Gabor filter by two indexes, which are the normalized radial frequency <inline-formula><mml:math id="M15"><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msqrt><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003BC;</mml:mi><mml:mo>/</mml:mo><mml:mi>M</mml:mi><mml:mo>/</mml:mo><mml:mn>2</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003BC;</mml:mi><mml:mo>/</mml:mo><mml:mi>N</mml:mi><mml:mo>/</mml:mo><mml:mn>2</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:msqrt></mml:math></inline-formula>, and the angle of orientation <italic>f</italic><sub>&#x003B8;</sub> &#x0003D; arctan(&#x003C5;/&#x003BC;). The parameter <italic>f</italic><sub><italic>rs</italic></sub> is the normalized center frequency of the scale, and the bandwidth of the filter is determined by &#x003C3;<sub><italic>s</italic></sub>/<italic>f</italic><sub><italic>rs</italic></sub>. The parameters &#x003BC;<sub>0</sub> and &#x003C3;<sub>0</sub> denote the orientation and angular spread of the filter, respectively. The parameters <italic>f</italic><sub><italic>rs</italic></sub>, &#x003C3;<sub><italic>s</italic></sub>, &#x003BC;<sub>0</sub>, and &#x003C3;<sub>0</sub> can be determined by the corresponding evaluation derived from the HVS, since it is known that the log-Gabor filter approximates cortical responses in the primary visual cortex.</p>
</sec>
<sec>
<title>3.1.4. The mutual information feature</title>
<p>The mutual information represents the amount of feature information that we can extract from the HVS output. For the source or lossless images, we define the mutual information to be the fourth feature by</p>
<disp-formula id="E14"><label>(14)</label><mml:math id="M16"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>M</mml:mi><mml:mi>I</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac><mml:mstyle displaystyle="true"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>N</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:msub><mml:mrow><mml:mi>g</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:msubsup><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>|</mml:mo><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>U</mml:mi></mml:mrow></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mtext class="textrm" mathvariant="normal">I</mml:mtext><mml:mo>|</mml:mo></mml:mrow><mml:mrow><mml:mo>|</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mtext class="textrm" mathvariant="normal">I</mml:mtext><mml:mo>|</mml:mo></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where <italic>x</italic> &#x0003D; {<italic>x</italic><sub><italic>i</italic></sub> : <italic>i</italic> &#x02208; <italic>I</italic>} is an RF of positive scalars and <italic>U</italic> is a Gaussian vector RF with mean zero and covariance <italic>B</italic><sub><italic>U</italic></sub>. <inline-formula><mml:math id="M17"><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>U</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>Q</mml:mi><mml:mtext>&#x0039B;</mml:mtext><mml:msup><mml:mrow><mml:mi>Q</mml:mi></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> is symmetric. The parameter <italic>Q</italic> is an orthonormal matrix and &#x0039B; is a diagonal matrix. <inline-formula><mml:math id="M18"><mml:msubsup><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> represents the variance of the visual noise. The RFs <italic>M</italic> and <italic>N</italic> are supposed to be independent of <italic>U</italic> and <inline-formula><mml:math id="M19"><mml:msub><mml:mrow><mml:mi>B</mml:mi></mml:mrow><mml:mrow><mml:mi>M</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup></mml:math></inline-formula> I. |&#x000B7;| denotes the determinant.</p>
</sec>
<sec>
<title>3.1.5. The novelty structural feature</title>
<p>The HVS is sensitive to structural distortion since natural images are highly structured. The structural feature of an image represents the structure of objects in the scene, different from the contrast and luminance. For example, as regards SSIM, Wang et al. (<xref ref-type="bibr" rid="B39">2004</xref>) calculate the differences in a few features (i.e., contrast, structural, and luminance) between <italic>I</italic><sub><italic>r</italic></sub> and <italic>I</italic><sub><italic>d</italic></sub>. Multiscale structural similarity (MS-SSIM) (Wang et al., <xref ref-type="bibr" rid="B41">2003</xref>) mainly incorporates contrast and structural similarities that are more effective than the luminance similarity in SSIM. We compute the contrast similarity:</p>
<disp-formula id="E15"><label>(15)</label><mml:math id="M20"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>E</mml:mi><mml:mrow><mml:mo stretchy="false">[</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mn>2</mml:mn><mml:msub><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:msubsup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>&#x0002B;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>I</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msubsup><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">]</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where &#x003B7;<sub>(<sub><italic>I</italic></sub><sub><italic>d</italic></sub>)</sub> and &#x003B7;<sub>(<sub><italic>I</italic></sub><sub><italic>r</italic></sub>)</sub> are the gradient values for the central pixel of images <italic>I</italic><sub><italic>d</italic></sub> and <italic>I</italic><sub><italic>r</italic></sub>, respectively. <italic>A</italic><sub>1</sub>, <italic>A</italic><sub>2</sub>, and <italic>A</italic><sub>3</sub> are all fixed. <italic>E</italic>(&#x000B7;) represents the expectation or the mean value. In the pixel version, we define <italic>F</italic><sub><italic>NFS</italic></sub> to be the fifth feature:</p>
<disp-formula id="E16"><label>(16)</label><mml:math id="M21"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>N</mml:mi><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>,</mml:mo><mml:mi>j</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mn>2</mml:mn><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:mi>R</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:mi>K</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:mi>R</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:mi>K</mml:mi></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where <inline-formula><mml:math id="M22"><mml:mi>R</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>|</mml:mo></mml:mrow><mml:mrow><mml:mstyle class="text"><mml:mtext class="textrm" mathvariant="normal">max</mml:mtext></mml:mstyle><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>,</mml:mo><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow></mml:mfrac></mml:math></inline-formula> and <inline-formula><mml:math id="M23"><mml:mi>K</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>A</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mo>/</mml:mo><mml:mstyle class="text"><mml:mtext class="textrm" mathvariant="normal">max</mml:mtext></mml:mstyle><mml:msup><mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>,</mml:mo><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>S</mml:mi><mml:mi>F</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula> of image blocks <italic>i</italic> and <italic>j</italic>.</p>
</sec>
</sec>
<sec>
<title>3.2. The second step</title>
<p>Inspired by Gu et al. (<xref ref-type="bibr" rid="B17">2020</xref>), we propose a semi-full combination subspace method, which is an elaborate integration of bootstrapping and aggregation applied to environmental factors. The semi-full combination subspace exerts bootstrapping on the input features. A high-dimensional feature vector or a small number of training samples is very likely to lead to an overfitting. Specifically, directly using all of the aforementioned five features is not always superior to the situation of using only a few of them. To address this issue, a new subset composed of a segment of the features is generated, which decreases the conformity between the length of the feature vector and the size of the training sample. Using the new semi-full combination subspace, we can obtain a component learner. By applying the aforementioned process to the feature space repeatedly through the feature selection, we can build multiple component learners with diversity of the environmental factors.</p>
</sec>
<sec>
<title>3.3. The third step</title>
<p>We use the semi-full combination subspace of five features to attain a single direct visual quality of gastroscope images. An efficient regression engine, namely support vector regression (SVR) (Mittal et al., <xref ref-type="bibr" rid="B27">2012</xref>), is used to reliably transform the semi-full combination subspace into a single objective quality score. Concretely, we implement the SVR by the radial basis function (RBF) kernel (Mittal et al., <xref ref-type="bibr" rid="B27">2012</xref>) included in the LibSVM package, as shown in <xref ref-type="fig" rid="F7">Figure 7</xref>.</p>
<fig id="F7" position="float">
<label>Figure 7</label>
<caption><p>The implementation of an efficient regression engine.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0007.tif"/>
</fig>
<sec>
<title>3.3.1. SVR training</title>
<p>We train an SVR to learn a regression model using the GIMB database. This database contains a number of different gastrointestinal tract regions and motion blur levels. To train our proposed model, we split the GIMB into 40% data for testing and 60% data for training. The SVR has significant advantages of high efficiency and flexibility.</p>
<p>We consider the GIMB training database <italic>T</italic> &#x0003D; {(<italic>f</italic><sub>1</sub>, <italic>m</italic><sub>1</sub>), (<italic>f</italic><sub>2</sub>, <italic>m</italic><sub>2</sub>), &#x02026;, (<italic>f</italic><sub><italic>r</italic></sub>, <italic>m</italic><sub><italic>r</italic></sub>)}, where <italic>f</italic><sub><italic>r</italic></sub> and <italic>m</italic><sub><italic>r</italic></sub>, <italic>r</italic> &#x0003D; {1, &#x02026;, <italic>N</italic>}. <italic>f</italic><sub><italic>r</italic></sub> indicates a feature vector of <italic>f</italic><sub>1</sub> &#x02212; <italic>f</italic><sub>5</sub> of the <italic>r</italic>th training image. The training labels <italic>m</italic><sub><italic>r</italic></sub> are subjective MOSs. We express the linear soft-margin SVR as</p>
<disp-formula id="E17"><label>(17)</label><mml:math id="M24"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:munder><mml:mrow><mml:mi>min</mml:mi></mml:mrow><mml:mrow><mml:mi>w</mml:mi><mml:mo>,</mml:mo><mml:mi>b</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:mover accent='true'><mml:mrow><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub></mml:mrow><mml:mo stretchy='false'>&#x0005E;</mml:mo></mml:mover></mml:mrow></mml:munder><mml:mtext>&#x02003;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mo>&#x02016;</mml:mo><mml:mi>W</mml:mi><mml:msup><mml:mo>&#x02016;</mml:mo><mml:mn>2</mml:mn></mml:msup><mml:mo>+</mml:mo><mml:mtext>&#x000A0;</mml:mtext><mml:mi>H</mml:mi><mml:mstyle displaystyle='true'><mml:munderover><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mi>N</mml:mi></mml:munderover><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:mover accent='true'><mml:mrow><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub></mml:mrow><mml:mo stretchy='false'>&#x0005E;</mml:mo></mml:mover><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mstyle></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>s</mml:mi><mml:mo>.</mml:mo><mml:mi>t</mml:mi><mml:mo>.</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mi>m</mml:mi><mml:mi>o</mml:mi><mml:mi>d</mml:mi><mml:mi>e</mml:mi><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>m</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>&#x02264;</mml:mo><mml:mi>&#x003B5;</mml:mi><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>,</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:msub><mml:mi>m</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>m</mml:mi><mml:mi>o</mml:mi><mml:mi>d</mml:mi><mml:mi>e</mml:mi><mml:mo>&#x02264;</mml:mo><mml:mi>&#x003B5;</mml:mi><mml:mo>+</mml:mo><mml:mover accent='true'><mml:mrow><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub></mml:mrow><mml:mo stretchy='false'>&#x0005E;</mml:mo></mml:mover><mml:mo>,</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>&#x02265;</mml:mo><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mover accent='true'><mml:mrow><mml:msub><mml:mi>&#x003BE;</mml:mi><mml:mi>r</mml:mi></mml:msub></mml:mrow><mml:mo stretchy='false'>&#x0005E;</mml:mo></mml:mover><mml:mo>&#x02265;</mml:mo><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mi>r</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>2</mml:mn><mml:mo>,</mml:mo><mml:mo>&#x02026;</mml:mo><mml:mo>,</mml:mo><mml:mi>N</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where we set kernel function <italic>K</italic>(<italic>f</italic><sub><italic>r</italic></sub>, <italic>f</italic><sub><italic>i</italic></sub>) to be the RBF kernel defined by</p>
<disp-formula id="E18"><label>(18)</label><mml:math id="M25"><mml:mrow><mml:mi>K</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:msup><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mi>T</mml:mi></mml:msup><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>f</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mi>e</mml:mi><mml:mi>x</mml:mi><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>k</mml:mi><mml:mo stretchy='false'>&#x02016;</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>r</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:msup><mml:mo stretchy='false'>&#x02016;</mml:mo><mml:mn>2</mml:mn></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:math></disp-formula>
<p>By training the SVR on the GIMB database, we want to determine the optimal parameters <italic>H</italic>, &#x003B5;, and <italic>k</italic> to obtain a fixed regression model, which is defined as</p>
<disp-formula id="E19"><label>(19)</label><mml:math id="M26"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mi>m</mml:mi><mml:mi>o</mml:mi><mml:mi>d</mml:mi><mml:mi>e</mml:mi><mml:mi>l</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mtext class="textrm" mathvariant="normal">SVR</mml:mtext></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mrow><mml:mi>r</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mrow><mml:mi>D</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mi>r</mml:mi><mml:mi>a</mml:mi><mml:mi>i</mml:mi><mml:mi>n</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where <italic>D</italic><sub><italic>train</italic></sub> is the training set. Five features are extracted to create a model named the gastroscope image quality evaluator (GIQE).</p>
</sec>
<sec>
<title>3.3.2. SVR prediction</title>
<p>Finally, the performance of the proposed GIQE metric is verified on testing the GIMB database with the obtained model. The perceived quality score <italic>Q</italic><sub><italic>j</italic></sub> of GIQE for gastroscope images is computed by</p>
<disp-formula id="E20"><label>(20)</label><mml:math id="M27"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>Q</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mtext class="textrm" mathvariant="normal">SVR</mml:mtext></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi><mml:mi>i</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:mi>m</mml:mi><mml:mi>o</mml:mi><mml:mi>d</mml:mi><mml:mi>e</mml:mi><mml:mi>l</mml:mi><mml:mo>,</mml:mo><mml:msub><mml:mrow><mml:mi>D</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi><mml:mi>i</mml:mi><mml:mi>c</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where <italic>D</italic><sub><italic>predict</italic></sub> is the testing set.</p>
</sec>
</sec>
</sec>
<sec id="s4">
<title>4. Comparison of objective quality assessment metrics</title>
<p>In this section, we investigate whether several existing FR IQA models can evaluate the quality of gastroscopic motion blurred images effectively. There are 20 traditional and mainstream FR IQA methods. Four commonly used performance indicators are adopted to compute the correlation between each MOS and FR IQA metric.</p>
<sec>
<title>4.1. Objective quality assessment models</title>
<p>We introduce some categories of FR IQA algorithms as follows:</p>
<list list-type="bullet">
<list-item><p>Analysis of distortion distribution-based SSIM (ADD-SSIM): Gu et al. (<xref ref-type="bibr" rid="B11">2016a</xref>) propose a high-performance fusion model based on the SSIM by analyzing the distortion distribution influenced by the image content and distortion.</p></list-item>
<list-item><p>MAD: Larson and Chandler (<xref ref-type="bibr" rid="B22">2010</xref>) evaluates the perceived quality of low- and high-quality images using two different strategies respectively.</p></list-item>
<list-item><p>Visual signal-to-noise ratio (VSNR): Chandler and Hemami (<xref ref-type="bibr" rid="B2">2007</xref>) uses image features to estimate the image quality in the wavelet domain, visual masking, near-threshold, and supra-threshold properties.</p></list-item>
<list-item><p>Analysis of distortion distribution GSIM (ADD-GSIM): Gu et al. (<xref ref-type="bibr" rid="B11">2016a</xref>) incorporate the frequency variation, distortion intensity, histogram changes, and distortion position distributions to infer the image quality.</p></list-item>
<list-item><p>IFC: Sheikh et al. (<xref ref-type="bibr" rid="B34">2005</xref>) use the natural scene statistics captured by sophisticated models to propose a novel information fidelity criterion (IFC).</p></list-item>
<list-item><p>VIF: Sheikh and Bovik (<xref ref-type="bibr" rid="B33">2006</xref>) considers it an information fidelity problem to quantify the loss of distorted images and explore the correlation between visual quality and images.</p></list-item>
<list-item><p>Visual information fidelity in pixel domain (VIFP): Sheikh and Bovik (<xref ref-type="bibr" rid="B33">2006</xref>) develops a novel version of VIF in the pixel domain to reduce computational complexity.</p></list-item>
<list-item><p>IGM: Wu et al. (<xref ref-type="bibr" rid="B42">2012</xref>) control the process of cognition according to the basic hypothesis of the free-energy-based brain theory.</p></list-item>
<list-item><p>Local-tuned-global model (LTG): Gu et al. (<xref ref-type="bibr" rid="B15">2014a</xref>) assume that the HVS draws on the prominent local distortion and global quality degradation to characterize the image quality.</p></list-item>
<list-item><p>Noise quality measure (NQM): Damera-Venkata et al. (<xref ref-type="bibr" rid="B5">2000</xref>) combine the local luminance mean, contrast pyramid of Peli, contrast sensitivity, contrast mask effects, and contrast interaction in spatial-frequency domain.</p></list-item>
<list-item><p>Reduced-reference image quality metric for contrast change (RIQMC): Gu et al. (<xref ref-type="bibr" rid="B11">2016a</xref>) design a novel pooling module by the analysis of distortion intensity, distortion position, histogram changes, and frequency changes to infer an overall quality measurement.</p></list-item>
<list-item><p>Structural variation-based quality index (SVQI): Gu et al. (<xref ref-type="bibr" rid="B9">2017b</xref>) evaluate the perceived quality of image based on the analysis of global and local structural variations on account of transmission, compression, etc.</p></list-item>
<list-item><p>Perceptual similarity (PSIM): Gu et al. (<xref ref-type="bibr" rid="B7">2017a</xref>) take into account the similarities of GM at two scales and color information, and an effective fusion based on perception.</p></list-item>
<list-item><p>GMSD: Xue et al. (<xref ref-type="bibr" rid="B45">2013</xref>) explore a novel fusion strategy according to the pixelwise gradient magnitude similarity (GMS) between the lossless image and the corresponding distorted image.</p></list-item>
<list-item><p>FSIM and Feature similarity in color domain (FSIMC): Zhang et al. (<xref ref-type="bibr" rid="B50">2011</xref>) compute the PC and the similarity of GM between the lossless image and the distorted image.</p></list-item>
<list-item><p>GSIM: Liu et al. (<xref ref-type="bibr" rid="B25">2011</xref>) combine gradient features with visual distortion and masking effect.</p></list-item>
<list-item><p>Visual saliency-induced index (VSI): Zhang et al. (<xref ref-type="bibr" rid="B49">2014</xref>) skillfully combine the GM variations and vision saliency to perceive the image quality.</p></list-item>
</list>
</sec>
<sec>
<title>4.2. Performance of the objective quality assessment models</title>
<p>After introducing the aforementioned objective quality assessment models, we first map the objective predictions of the IQA models by the five-parameter logistic function:</p>
<disp-formula id="E21"><label>(21)</label><mml:math id="M28"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>Q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mi>e</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mo>&#x000B7;</mml:mo><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>z</mml:mi><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:msub><mml:mo>&#x000B7;</mml:mo><mml:mi>z</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow><mml:mrow><mml:mn>5</mml:mn></mml:mrow></mml:msub></mml:mtd></mml:mtr></mml:mtable></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Where <italic>x</italic> and <italic>Q</italic>(<italic>x</italic>) represent the input scores and the mapped scores, respectively. <italic>z</italic> is the predicted score of the IQA. &#x003B2;<sub><italic>i</italic></sub>(<italic>i</italic> &#x0003D; 1, 2, &#x02026;, 5) are variable parameters that have to be defined in the fitting process.</p>
<p>Then, we draw on four statistical indicators, as detailed in Zhang et al. (<xref ref-type="bibr" rid="B49">2014</xref>), to compare the consistency of the predicted ratings from subjective MOSs and objective IQA models. The four indicators represent different meanings and evaluate the predicted performance in different ways. First, Pearson&#x00027;s linear correlation coefficient (PLCC) points out the accuracy by computing correlation of the subjective and objective scores. Second, Spearman&#x00027;s rank-order correlation coefficient (SROCC) reflects the predicted monotonicity of IQA, which does not dependent on any monotone nonlinear mapping between the objective scores and MOSs. Third, Kendall&#x00027;s rank-order correlation coefficient (KROCC) is a nonparametric rank correlation metric to measure the matching between the original scores and the converted objective ones. The last root mean-squared error (RMSE) indicates the predicted consistency, which is defined as the energy between two data sets. For the four indicators aforementioned, a superior IQA model means the values of PLCC, SROCC, and KROCC are close to 1, while the value of RMSE is close to 0. <xref ref-type="table" rid="T1">Table 1</xref> lists the performance of 20 FR IQA models on PLCC, SROCC, KROCC, and RMSE. The best performing objective methods are highlighted in boldface in each column.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Performance comparison of the proposed gastroscope image quality evaluator and the existing full-reference image quality assessment (FR IQA) metrics on the gastroscope image motion blur (GIMB) database.</p></caption>
<table frame="box" rules="all">
<thead>
<tr style="background-color:#919497; color:#ffffff;">
<th/>
<th valign="top" align="center"><bold>PLCC</bold></th>
<th valign="top" align="center"><bold>SROCC</bold></th>
<th valign="top" align="center"><bold>KROCC</bold></th>
<th valign="top" align="center"><bold>RMSE</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">PSNR</td>
<td valign="top" align="center">0.6463</td>
<td valign="top" align="center">0.5703</td>
<td valign="top" align="center">0.4042</td>
<td valign="top" align="center">0.6966</td>
</tr> <tr>
<td valign="top" align="left">MSE</td>
<td valign="top" align="center">0.4379</td>
<td valign="top" align="center">0.5833</td>
<td valign="top" align="center">0.4139</td>
<td valign="top" align="center">0.8207</td>
</tr> <tr>
<td valign="top" align="left">SSIM</td>
<td valign="top" align="center">0.5831</td>
<td valign="top" align="center">0.5313</td>
<td valign="top" align="center">0.3730</td>
<td valign="top" align="center">0.7416</td>
</tr> <tr>
<td valign="top" align="left">ADD-SSIM</td>
<td valign="top" align="center">0.7608</td>
<td valign="top" align="center">0.7022</td>
<td valign="top" align="center">0.5099</td>
<td valign="top" align="center">0.5924</td>
</tr> <tr>
<td valign="top" align="left">MS-SSIM</td>
<td valign="top" align="center">0.7826</td>
<td valign="top" align="center">0.7248</td>
<td valign="top" align="center">0.5493</td>
<td valign="top" align="center">0.5683</td>
</tr> <tr>
<td valign="top" align="left">FSIM</td>
<td valign="top" align="center">0.8571</td>
<td valign="top" align="center">0.8325</td>
<td valign="top" align="center">0.6343</td>
<td valign="top" align="center">0.4702</td>
</tr> <tr>
<td valign="top" align="left">FSIMC</td>
<td valign="top" align="center">0.8568</td>
<td valign="top" align="center">0.8319</td>
<td valign="top" align="center">0.6336</td>
<td valign="top" align="center">0.4708</td>
</tr> <tr>
<td valign="top" align="left">PSIM</td>
<td valign="top" align="center">0.7335</td>
<td valign="top" align="center">0.6780</td>
<td valign="top" align="center">0.4904</td>
<td valign="top" align="center">0.6205</td>
</tr> <tr>
<td valign="top" align="left">MAD</td>
<td valign="top" align="center">0.8585</td>
<td valign="top" align="center">0.8412</td>
<td valign="top" align="center">0.6404</td>
<td valign="top" align="center">0.4682</td>
</tr> <tr>
<td valign="top" align="left">VSNR</td>
<td valign="top" align="center">0.6027</td>
<td valign="top" align="center">0.5312</td>
<td valign="top" align="center">0.3707</td>
<td valign="top" align="center">0.7284</td>
</tr> <tr>
<td valign="top" align="left">GMSD</td>
<td valign="top" align="center">0.7813</td>
<td valign="top" align="center">0.7090</td>
<td valign="top" align="center">0.5266</td>
<td valign="top" align="center">0.5697</td>
</tr> <tr>
<td valign="top" align="left">GSIM</td>
<td valign="top" align="center">0.8486</td>
<td valign="top" align="center">0.8205</td>
<td valign="top" align="center">0.6237</td>
<td valign="top" align="center">0.4829</td>
</tr> <tr>
<td valign="top" align="left">ADD-GSIM</td>
<td valign="top" align="center">0.7657</td>
<td valign="top" align="center">0.7026</td>
<td valign="top" align="center">0.5 135</td>
<td valign="top" align="center">0.5871</td>
</tr> <tr>
<td valign="top" align="left">VIF</td>
<td valign="top" align="center">0.8392</td>
<td valign="top" align="center">0.8285</td>
<td valign="top" align="center">0.6329</td>
<td valign="top" align="center">0.4964</td>
</tr> <tr>
<td valign="top" align="left">VIFP</td>
<td valign="top" align="center">0.8575</td>
<td valign="top" align="center">0.8269</td>
<td valign="top" align="center">0.6319</td>
<td valign="top" align="center">0.4697</td>
</tr> <tr>
<td valign="top" align="left">IGM</td>
<td valign="top" align="center">0.7836</td>
<td valign="top" align="center">0.7111</td>
<td valign="top" align="center">0.5263</td>
<td valign="top" align="center">0.5671</td>
</tr> <tr>
<td valign="top" align="left">LTG</td>
<td valign="top" align="center">0.7690</td>
<td valign="top" align="center">0.7038</td>
<td valign="top" align="center">0.522 1</td>
<td valign="top" align="center">0.5836</td>
</tr> <tr>
<td valign="top" align="left">NQM</td>
<td valign="top" align="center">0.8533</td>
<td valign="top" align="center">0.8376</td>
<td valign="top" align="center">0.6375</td>
<td valign="top" align="center">0.4760</td>
</tr> <tr>
<td valign="top" align="left">VSI</td>
<td valign="top" align="center">0.8316</td>
<td valign="top" align="center">0.8027</td>
<td valign="top" align="center">0.6050</td>
<td valign="top" align="center">0.5070</td>
</tr> <tr>
<td valign="top" align="left">IFC</td>
<td valign="top" align="center">0.8630</td>
<td valign="top" align="center">0.8501</td>
<td valign="top" align="center">0.6578</td>
<td valign="top" align="center">0.4612</td>
</tr> <tr>
<td valign="top" align="left">GIQE</td>
<td valign="top" align="center"><bold>0.8883</bold></td>
<td valign="top" align="center"><bold>0.8849</bold></td>
<td valign="top" align="center"><bold>0.6988</bold></td>
<td valign="top" align="center"><bold>0.7766</bold></td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p>The best results of each indicator are highlighted in bold.</p>
</table-wrap-foot>
</table-wrap>
<p>We compared the performance of 20 commonly used FR IQA models for gastroscope motion blurred images. From <xref ref-type="table" rid="T1">Table 1</xref>, we derive some important conclusions as follows:</p>
<p>(1) The top two IQA models are highlighted in different bold colors to compare our method with those of other competitors straightaway. It is obvious that the proposed GIQE model, whose PLCC, SROCC, KROCC, and RMSE reach 0.8883, 0.8849, 0.6988, and 0.7766, respectively, shows a better performance than the existing FR IQA models.</p>
<p>Specifically, we concentrate only on the PLCC indicator, and similar conclusions can be drawn from the other three indicators. IFC is the second best performing model, achieving 0.8630 on PLCC. Compared with the IFC, the performance of the proposed IQA metric GIQE has improved by 2.9%. The performance gains of the proposed GIQE models are 13.5 and 16.7% higher than those of MS-SSIM and ADD-SSIM, respectively.</p>
<p>(2) We can see that a few aforementioned FR IQA models do not exhibit a remarkably high correlation with subjective quality. For example, the performance of PSNR and VSNR for the gastroscope motion blurred images is low. It means that the assessment model is not suitable for the study of gastroscope images. Since the gastroscope is placed inside the body, the images it produces do not contain most types of distortion found in natural images, such as pulse noise, brightness changes, and JPEG. It causes the PSNR to be inferior to the traditional successful methods for natural images.</p>
<p>(3) We study the performance of the SSIM and SSIM-based FR IQA models for gastroscope motion blurred images. The performance of all SSIM-based IQA models has showed an improvement compared with that of SSIM, indicating that they can promote the analysis of motion blurred distortion in gastroscope images. Both ADD-SSIM and MS-SSIM analyze the influence of motion blurred distortion on image&#x00027;s structure. MS-SSIM performs the best among these SSIM-based IQA methods and it obtains values 0.7826, 0.7248, 0.5493, and 0.5683 of PLCC, SROCC, KROCC, and RMSE, respectively. Yet, SSIM obtains values 0.5831, 0.5313, 0.3730, and 0.7416 of PLCC, SROCC, KROCC, and RMSE, respectively, which is the worst performing model among all FR IQA tested methods. It shows that the motion blurred distortion caused by the superposition of multiple images at different times has a great effect on the structure of images.</p>
<p>(4) We find that the methods based on the image gradients, such as FSIM, FSIMC, and GSIM, achieve a high performance in terms of these traditional FR IQA metrics, as image gradients are significant in gastroscope images. VIF, VIFP, and VSI metrics achieve superior performance than most of the tested FR IQAs, which indicate that the features extracted by VIF, VIFP, and VSI metrics are less affected by the motion blur. It brought to light the fact that the visual saliency features of VIF, VIFP, and VSI models are useful for assessing the quality of gastroscope motion blurred images. In addition, the saliency models used in VIF, VIFP, and VSI models are not specially devised for motion blurred images.</p>
<p>(5) Among the existing FR IQA methods, IFC shows the best performance, which achieves values 0.8630, 0.8501, 0.6578, and 0.4612 of PLCC, SROCC, KROCC, and RMSE, respectively. This observation indicates that IFC has the highest correlation with the perceptual scores for gastroscope images. However, the performance of IFC is far from satisfactory. All the existing FR IQA methods do not take into consideration the distorion-specific category of the gastroscope image. The objective algorithm for gastroscope images needs to be studied further.</p>
<p>The scatter plot is a common manifestation of comparison in the IQA study, which can show some direct-viewing illustrations of different IQA models. In <xref ref-type="fig" rid="F8">Figure 8</xref>, we provide the scatter plots of MOS vs. 20 existing objective FR IQA methods tested on the proposed GIMB database. These representative models are composed of PSNR, SSIM, ADD-SIMM, MS-SSIM, FSIM, FSIMC, PSIM, MAD, VSNR, GMSD, GSIM, ADD-GSIM, IFC, VIF, VIFP, VSI, VIF, VIFP, IGM, LTG, and NQM. It can be seen that the sample points of IFC, MAD, NQM, and VIF present better convergence and linearity, which illustrates that these models can deliver more consistent results between the objective scores and the subjective scores. From <xref ref-type="fig" rid="F8">Figure 8</xref>, we find that the proposed GIQE method (i.e., the last scatter plot) is more robust and shows a better performance with regard to correlation than the existing FR IQA models (including IFC, MAD, NQM, and VIF). Particularly, the sample points of the proposed GIQE metric are quite close to the centerline, whereas those of the majority of other tested FR IQA models are far from the centerline. According to this, we assume that the proposed GIQE method demonstrates higher consistency in prediction performance.</p>
<fig id="F8" position="float">
<label>Figure 8</label>
<caption><p>Scatter plots of mean opinion scores (MOS) vs. the proposed <bold>(T)</bold> gastroscope image quality evaluator (GIQE) and traditional Full-Reference Image Quality Assessment (IQA) models [<bold>(A)</bold> peak signal-to-noise ratio (PSNR), <bold>(B)</bold> structural similarity (SSIM), <bold>(C)</bold> analysis of distortion distribution-based SSIM (ADD-SSIM), <bold>(D)</bold> multi-scale structural similarity (MS-SSIM), <bold>(E)</bold> feature similarity (FSIM), <bold>(F)</bold> feature similarity in color domain (FSIMC), <bold>(G)</bold> perceptual similarity (PSIM), <bold>(H)</bold> most apparent distortion (MAD), <bold>(I)</bold> gradient magnitude similarity deviation (GMSD), <bold>(J)</bold> gradient similarity (GSIM), <bold>(K)</bold> analysis of distortion distribution GSIM (ADD-GSIM), <bold>(L)</bold> information fidelity criterion (IFC), <bold>(M)</bold> local-tuned-global model (LTG), <bold>(N)</bold> internal generative mechanism (IGM), <bold>(O)</bold> visual signal-to-noise ratio (VSNR), <bold>(P)</bold> visual saliency-induced index (VSI), <bold>(Q)</bold> visual information fidelity (VIF), <bold>(R)</bold> visual information fidelity in pixel domain (VIFP), and <bold>(S)</bold> noise quality measure (NQM)].</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnins-16-1118087-g0008.tif"/>
</fig>
</sec>
</sec>
<sec sec-type="conclusions" id="s5">
<title>5. Conclusions</title>
<p>In this study, we have investigated comprehensively a significant quality assessment problem of gastroscope motion blurred images in EGC diagnosis and therapy systems. We built a carefully devised GIMB database to facilitate the image quality evaluation of the gastroscope motion blurred images. This database is composed of 1,050 distorted images under five pixel levels for three different motion angles. It associates MOS values scored by 15 experienced and inexperienced viewers. What&#x00027;s more, we compared 20 FR IQA models by combining different features of images. The IFC, VIF, FSIM, and NQM achieved high consistency with the subjective scores. The results of the comparison show that visual saliency information, structure information, and image gradients are crucial features when devising objective IQA algorithms for gastroscope images. We then extracted and learned these features to design a novel IQA metric GIQE by adopting semi-full combination subspace. The results of the experiments imply that the proposed GIQE has always achieved a superior performance (i.e., better consistency) than the 20 existing FR IQA metrics. In the future, we would like to choose more lossless images to increase the capacity of the database. In addition, we would like to develop a higher performance objective IQA model for gastroscope images.</p>
</sec>
<sec sec-type="data-availability" id="s6">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec sec-type="ethics-statement" id="s7">
<title>Ethics statement</title>
<p>Written informed consent was obtained from the individual(s) for the publication of any potentially identifiable images or data included in this article.</p>
</sec>
<sec sec-type="author-contributions" id="s8">
<title>Author contributions</title>
<p>PY completed the first draft of the paper and confirmed the idea. RB completed the follow-up correction and modification of the paper. YY participated in the algorithm design of the paper. SL completed the experimental part of the paper. JW completed the summary part of the paper. CC completed the data collection part of the paper. QW completed the text correction and data collection part of the paper. All authors contributed to the article and approved the submitted version.</p>
</sec>
</body>
<back>
<sec sec-type="funding-information" id="s9">
<title>Funding</title>
<p>This work was supported in part by the Beijing Hospitals Authority Clinical Medicine Development of special funding support (XMLX202143), Capital&#x00027;s Funds for Health Improvement and Research (2020-2-2155), the Beijing Municipal Administration of Hospitals Incubating Program (PX2020047), the Science Foundation of Peking University Cancer Hospital (No. 202207), the Hygiene and Health Development Scientific Research Fostering Plan of Haidian District Beijing (HP2022-19-503002), and the Beijing Hospitals Authority Youth Programme (QML20211103).</p>
</sec>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest. The reviewer KG declared a shared affiliation with the author RB to the handling editor at the time of review.</p>
</sec>
<sec sec-type="disclaimer" id="s10">
<title>Publisher&#x00027;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cai</surname> <given-names>W.</given-names></name> <name><surname>Zhai</surname> <given-names>B.</given-names></name> <name><surname>Liu</surname> <given-names>Y.</given-names></name> <name><surname>Liu</surname> <given-names>R.</given-names></name> <name><surname>Ning</surname> <given-names>X.</given-names></name></person-group> (<year>2021</year>). <article-title>Quadratic polynomial guided fuzzy c-means and dual attention mechanism for medical image segmentation</article-title>. <source>Displays</source> <volume>70</volume>, <fpage>102106</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102106</pub-id></citation>
</ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chandler</surname> <given-names>D. M.</given-names></name> <name><surname>Hemami</surname> <given-names>S. S.</given-names></name></person-group> (<year>2007</year>). <article-title>Vsnr: a wavelet-based visual signal-to-noise ratio for natural images</article-title>. <source>IEEE Trans. Image Process</source>. <volume>16</volume>, <fpage>2284</fpage>&#x02013;<lpage>2298</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2007.901820</pub-id><pub-id pub-id-type="pmid">17784602</pub-id></citation></ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chen</surname> <given-names>D.</given-names></name> <name><surname>Fu</surname> <given-names>M.</given-names></name> <name><surname>Chi</surname> <given-names>L.</given-names></name> <name><surname>Lin</surname> <given-names>L.</given-names></name> <name><surname>Cheng</surname> <given-names>J.</given-names></name> <name><surname>Xue</surname> <given-names>W.</given-names></name> <etal/></person-group>. (<year>2022</year>). <article-title>Prognostic and predictive value of a pathomics signature in gastric cancer</article-title>. <source>Nat. Commun</source>. <volume>13</volume>, <fpage>1</fpage>&#x02013;<lpage>13</lpage>. <pub-id pub-id-type="doi">10.1038/s41467-022-34703-w</pub-id><pub-id pub-id-type="pmid">36371443</pub-id></citation></ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chen</surname> <given-names>Q.</given-names></name> <name><surname>Fan</surname> <given-names>J.</given-names></name> <name><surname>Chen</surname> <given-names>W.</given-names></name></person-group> (<year>2021</year>). <article-title>An improved image enhancement framework based on multiple attention mechanism</article-title>. <source>Displays</source> <volume>70</volume>, <fpage>102091</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102091</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Damera-Venkata</surname> <given-names>N.</given-names></name> <name><surname>Kite</surname> <given-names>T. D.</given-names></name> <name><surname>Geisler</surname> <given-names>W. S.</given-names></name> <name><surname>Evans</surname> <given-names>B. L.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name></person-group> (<year>2000</year>). <article-title>Image quality assessment based on a degradation model</article-title>. <source>IEEE Trans. Image Process</source>. <volume>9</volume>, <fpage>636</fpage>&#x02013;<lpage>650</lpage>. <pub-id pub-id-type="doi">10.1109/83.841940</pub-id><pub-id pub-id-type="pmid">18255436</pub-id></citation></ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fei</surname> <given-names>L. B. Z.</given-names></name></person-group> (<year>2020</year>). <article-title>Visual quality assessment for super-resolved images: database and method</article-title>. <source>Peng Cheng Lab. Commum</source>. <volume>1</volume>, <fpage>120</fpage>. <pub-id pub-id-type="doi">10.1109/TIP.2019.2898638</pub-id><pub-id pub-id-type="pmid">30762547</pub-id></citation></ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Li</surname> <given-names>L.</given-names></name> <name><surname>Lu</surname> <given-names>H.</given-names></name> <name><surname>Min</surname> <given-names>X.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name></person-group> (<year>2017a</year>). <article-title>A fast reliable image quality predictor by fusing micro-and macro-structures</article-title>. <source>IEEE Trans. Ind. Electron</source>. <volume>64</volume>, <fpage>3903</fpage>&#x02013;<lpage>3912</lpage>. <pub-id pub-id-type="doi">10.1109/TIE.2017.2652339</pub-id></citation>
</ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Liu</surname> <given-names>M.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Yang</surname> <given-names>X.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name></person-group> (<year>2015a</year>). <article-title>Quality assessment considering viewing distance and image resolution</article-title>. <source>IEEE Trans. Broadcast</source>. <volume>61</volume>, <fpage>520</fpage>&#x02013;<lpage>531</lpage>. <pub-id pub-id-type="doi">10.1109/TBC.2015.2459851</pub-id><pub-id pub-id-type="pmid">35062460</pub-id></citation></ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Qiao</surname> <given-names>J.</given-names></name> <name><surname>Min</surname> <given-names>X.</given-names></name> <name><surname>Yue</surname> <given-names>G.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <name><surname>Thalmann</surname> <given-names>D.</given-names></name></person-group> (<year>2017b</year>). <article-title>Evaluating quality of screen content images via structural variation analysis</article-title>. <source>IEEE Trans. Vis. Comput. Graph</source>. <volume>24</volume>, <fpage>2689</fpage>&#x02013;<lpage>2701</lpage>. <pub-id pub-id-type="doi">10.1109/TVCG.2017.2771284</pub-id><pub-id pub-id-type="pmid">29990169</pub-id></citation></ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Tao</surname> <given-names>D.</given-names></name> <name><surname>Qiao</surname> <given-names>J.-F.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name></person-group> (<year>2017c</year>). <article-title>Learning a no-reference quality assessment model of enhanced images with big data</article-title>. <source>IEEE Trans. Neural Netw. Learn. Syst</source>. <volume>29</volume>, <fpage>1301</fpage>&#x02013;<lpage>1313</lpage>. <pub-id pub-id-type="doi">10.1109/TNNLS.2017.2649101</pub-id><pub-id pub-id-type="pmid">28287984</pub-id></citation></ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Wang</surname> <given-names>S.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <name><surname>Yang</surname> <given-names>X.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name></person-group> (<year>2016a</year>). <article-title>Analysis of distortion distribution for pooling in image quality prediction</article-title>. <source>IEEE Trans. Broadcast</source>. <volume>62</volume>, <fpage>446</fpage>&#x02013;<lpage>456</lpage>. <pub-id pub-id-type="doi">10.1109/TBC.2015.2511624</pub-id></citation>
</ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Wang</surname> <given-names>S.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Ma</surname> <given-names>S.</given-names></name> <name><surname>Yang</surname> <given-names>X.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <etal/></person-group>. (<year>2016b</year>). <article-title>Blind quality assessment of tone-mapped images via analysis of information, naturalness, and structure</article-title>. <source>IEEE Trans. Multimedia</source> <volume>18</volume>, <fpage>432</fpage>&#x02013;<lpage>443</lpage>. <pub-id pub-id-type="doi">10.1109/TMM.2016.2518868</pub-id></citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <name><surname>Liu</surname> <given-names>M.</given-names></name></person-group> (<year>2015b</year>). <article-title>The analysis of image contrast: From quality assessment to automatic enhancement</article-title>. <source>IEEE Trans. Cybern</source>. <volume>46</volume>, <fpage>284</fpage>&#x02013;<lpage>297</lpage>. <pub-id pub-id-type="doi">10.1109/TCYB.2015.2401732</pub-id><pub-id pub-id-type="pmid">25775503</pub-id></citation></ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <name><surname>Yang</surname> <given-names>X.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name></person-group> (<year>2015c</year>). <article-title>No-reference image sharpness assessment in autoregressive parameter space</article-title>. <source>IEEE Trans. Image Process</source>. <volume>24</volume>, <fpage>3218</fpage>&#x02013;<lpage>3231</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2015.2439035</pub-id><pub-id pub-id-type="pmid">26054063</pub-id></citation></ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Yang</surname> <given-names>X.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name></person-group> (<year>2014a</year>). <article-title>&#x0201C;An efficient color image quality metric with local-tuned-global model,&#x0201D;</article-title> in <source>Conference on Image Processing</source>, <fpage>506</fpage>&#x02013;<lpage>510</lpage>.<pub-id pub-id-type="pmid">17324546</pub-id></citation></ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Yang</surname> <given-names>X.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name></person-group> (<year>2014b</year>). <article-title>Using free energy principle for blind image quality assessment</article-title>. <source>IEEE Trans. Multimedia</source> <volume>17</volume>, <fpage>50</fpage>&#x02013;<lpage>63</lpage>. <pub-id pub-id-type="doi">10.1109/TMM.2014.2373812</pub-id><pub-id pub-id-type="pmid">29185989</pub-id></citation></ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Zhang</surname> <given-names>Y.</given-names></name> <name><surname>Qiao</surname> <given-names>J.</given-names></name></person-group> (<year>2020</year>). <article-title>Random forest ensemble for river turbidity measurement from space remote sensing data</article-title>. <source>IEEE Trans. Instrum. Meas</source>. <volume>69</volume>, <fpage>9028</fpage>&#x02013;<lpage>9036</lpage>. <pub-id pub-id-type="doi">10.1109/TIM.2020.2998615</pub-id></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>K.</given-names></name> <name><surname>Zhou</surname> <given-names>J.</given-names></name> <name><surname>Qiao</surname> <given-names>J.-F.</given-names></name> <name><surname>Zhai</surname> <given-names>G.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name></person-group> (<year>2017d</year>). <article-title>No-reference quality assessment of screen content pictures</article-title>. <source>IEEE Trans. Image Process</source>. <volume>26</volume>, <fpage>4005</fpage>&#x02013;<lpage>4018</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2017.2711279</pub-id></citation>
</ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hu</surname> <given-names>R.</given-names></name> <name><surname>Liu</surname> <given-names>Y.</given-names></name> <name><surname>Wang</surname> <given-names>Z.</given-names></name> <name><surname>Li</surname> <given-names>X.</given-names></name></person-group> (<year>2021</year>). <article-title>Blind quality assessment of night-time image</article-title>. <source>Displays</source> <volume>69</volume>, <fpage>102045</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102045</pub-id></citation>
</ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Huang</surname> <given-names>Y.</given-names></name> <name><surname>Xu</surname> <given-names>H.</given-names></name> <name><surname>Ye</surname> <given-names>Z.</given-names></name></person-group> (<year>2021</year>). <article-title>Image quality evaluation for oled-based smart-phone displays at various lighting conditions</article-title>. <source>Displays</source> <volume>70</volume>, <fpage>102115</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102115</pub-id></citation>
</ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kundu</surname> <given-names>D.</given-names></name> <name><surname>Ghadiyaram</surname> <given-names>D.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name> <name><surname>Evans</surname> <given-names>B. L.</given-names></name></person-group> (<year>2017</year>). <article-title>Large-scale crowdsourced study for tone-mapped hdr pictures</article-title>. <source>IEEE Trans. Image Process</source>. <volume>26</volume>, <fpage>4725</fpage>&#x02013;<lpage>4740</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2017.2713945</pub-id><pub-id pub-id-type="pmid">28613173</pub-id></citation></ref>
<ref id="B22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Larson</surname> <given-names>E. C.</given-names></name> <name><surname>Chandler</surname> <given-names>D. M.</given-names></name></person-group> (<year>2010</year>). <article-title>Most apparent distortion: full-reference image quality assessment and the role of strategy</article-title>. <source>J. Electron. Imaging</source> <volume>19</volume>, <fpage>011006</fpage>. <pub-id pub-id-type="doi">10.1117/1.3267105</pub-id></citation>
</ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>T.</given-names></name> <name><surname>Min</surname> <given-names>X.</given-names></name> <name><surname>Zhu</surname> <given-names>W.</given-names></name> <name><surname>Xu</surname> <given-names>Y.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name></person-group> (<year>2021</year>). <article-title>No-reference screen content video quality assessment</article-title>. <source>Displays</source> <volume>69</volume>, <fpage>102030</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102030</pub-id></citation>
</ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>Y.-D.</given-names></name> <name><surname>Zhu</surname> <given-names>S.-W.</given-names></name> <name><surname>Yu</surname> <given-names>J.-P.</given-names></name> <name><surname>Ruan</surname> <given-names>R.-W.</given-names></name> <name><surname>Cui</surname> <given-names>Z.</given-names></name> <name><surname>Li</surname> <given-names>Y.-T.</given-names></name> <etal/></person-group>. (<year>2021</year>). <article-title>Intelligent detection endoscopic assistant: an artificial intelligence-based system for monitoring blind spots during esophagogastroduodenoscopy in real-time</article-title>. <source>Digest. Liver Dis</source>. <volume>53</volume>, <fpage>216</fpage>&#x02013;<lpage>223</lpage>. <pub-id pub-id-type="doi">10.1016/j.dld.2020.11.017</pub-id><pub-id pub-id-type="pmid">33272862</pub-id></citation></ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>A.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <name><surname>Narwaria</surname> <given-names>M.</given-names></name></person-group> (<year>2011</year>). <article-title>Image quality assessment based on gradient similarity</article-title>. <source>IEEE Trans. Image Process</source>. <volume>21</volume>, <fpage>1500</fpage>&#x02013;<lpage>1512</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2011.2175935</pub-id><pub-id pub-id-type="pmid">22106145</pub-id></citation></ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>Y.</given-names></name> <name><surname>Lin</surname> <given-names>D.</given-names></name> <name><surname>Li</surname> <given-names>L.</given-names></name> <name><surname>Chen</surname> <given-names>Y.</given-names></name> <name><surname>Wen</surname> <given-names>J.</given-names></name> <name><surname>Lin</surname> <given-names>Y.</given-names></name> <etal/></person-group>. (<year>2021</year>). <article-title>Using machine-learning algorithms to identify patients at high risk of upper gastrointestinal lesions for endoscopy</article-title>. <source>J. Gastroenterol. Hepatol</source>. <volume>36</volume>, <fpage>2735</fpage>&#x02013;<lpage>2744</lpage>. <pub-id pub-id-type="doi">10.1111/jgh.15530</pub-id><pub-id pub-id-type="pmid">33929063</pub-id></citation></ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mittal</surname> <given-names>A.</given-names></name> <name><surname>Moorthy</surname> <given-names>A. K.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name></person-group> (<year>2012</year>). <article-title>No-reference image quality assessment in the spatial domain</article-title>. <source>IEEE Trans. Image Process</source>. <volume>21</volume>, <fpage>4695</fpage>&#x02013;<lpage>4708</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2012.2214050</pub-id><pub-id pub-id-type="pmid">22910118</pub-id></citation></ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pan</surname> <given-names>D.</given-names></name> <name><surname>Wang</surname> <given-names>X.</given-names></name> <name><surname>Shi</surname> <given-names>P.</given-names></name> <name><surname>Yu</surname> <given-names>S.</given-names></name></person-group> (<year>2021</year>). <article-title>No-reference video quality assessment based on modeling temporal-memory effects</article-title>. <source>Displays</source> <volume>70</volume>, <fpage>102075</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102075</pub-id></citation>
</ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ponomarenko</surname> <given-names>N.</given-names></name> <name><surname>Ieremeiev</surname> <given-names>O.</given-names></name> <name><surname>Lukin</surname> <given-names>V.</given-names></name> <name><surname>Egiazarian</surname> <given-names>K.</given-names></name> <name><surname>Jin</surname> <given-names>L.</given-names></name> <name><surname>Astola</surname> <given-names>J.</given-names></name> <etal/></person-group>. (<year>2013</year>). <article-title>&#x0201C;Color image database tid2013: peculiarities and preliminary results,&#x0201D;</article-title> in <source>European Workshop on Visual Information Processing</source>, <fpage>106</fpage>&#x02013;<lpage>111</lpage>.</citation>
</ref>
<ref id="B30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Qin</surname> <given-names>Z.</given-names></name> <name><surname>Zeng</surname> <given-names>Q.</given-names></name> <name><surname>Zong</surname> <given-names>Y.</given-names></name> <name><surname>Xu</surname> <given-names>F.</given-names></name></person-group> (<year>2021</year>). <article-title>Image inpainting based on deep learning: a review</article-title>. <source>Displays</source> <volume>69</volume>, <fpage>102028</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102028</pub-id></citation>
</ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ritur</surname> <given-names>B. T.</given-names></name></person-group> (<year>2002</year>). <article-title>&#x0201C;Methodology for the subjective assessment of the quality of television pictures,&#x0201D;</article-title> in <source>International Telecommunication Union</source>.</citation>
</ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sampat</surname> <given-names>M. P.</given-names></name> <name><surname>Wang</surname> <given-names>Z.</given-names></name> <name><surname>Gupta</surname> <given-names>S.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name> <name><surname>Markey</surname> <given-names>M. K.</given-names></name></person-group> (<year>2009</year>). <article-title>Complex wavelet structural similarity: a new image similarity index</article-title>. <source>IEEE Trans. Image Process</source>. <volume>18</volume>, <fpage>2385</fpage>&#x02013;<lpage>2401</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2009.2025923</pub-id><pub-id pub-id-type="pmid">19556195</pub-id></citation></ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sheikh</surname> <given-names>H. R.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name></person-group> (<year>2006</year>). <article-title>Image information and visual quality</article-title>. <source>IEEE Trans. Image Process</source>. <volume>15</volume>, <fpage>430</fpage>&#x02013;<lpage>444</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2005.859378</pub-id><pub-id pub-id-type="pmid">16479813</pub-id></citation></ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sheikh</surname> <given-names>H. R.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name> <name><surname>De Veciana</surname> <given-names>G.</given-names></name></person-group> (<year>2005</year>). <article-title>An information fidelity criterion for image quality assessment using natural scene statistics</article-title>. <source>IEEE Trans. Image Process</source>. <volume>14</volume>, <fpage>2117</fpage>&#x02013;<lpage>2128</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2005.859389</pub-id><pub-id pub-id-type="pmid">16370464</pub-id></citation></ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sheikh</surname> <given-names>H. R.</given-names></name> <name><surname>Sabir</surname> <given-names>M. F.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name></person-group> (<year>2006</year>). <article-title>A statistical evaluation of recent full reference image quality assessment algorithms</article-title>. <source>IEEE Trans. Image Process</source>. <volume>15</volume>, <fpage>3440</fpage>&#x02013;<lpage>3451</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2006.881959</pub-id><pub-id pub-id-type="pmid">17076403</pub-id></citation></ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shi</surname> <given-names>Y.</given-names></name> <name><surname>Tu</surname> <given-names>Y.</given-names></name> <name><surname>Wang</surname> <given-names>L.</given-names></name> <name><surname>Zhang</surname> <given-names>Y.</given-names></name> <name><surname>Zhang</surname> <given-names>Y.</given-names></name> <name><surname>Wang</surname> <given-names>B.</given-names></name></person-group> (<year>2021</year>). <article-title>Spectral influence of the normal lcd, blue-shifted lcd, and oled smartphone displays on visual fatigue: a comparative study</article-title>. <source>Displays</source> <volume>69</volume>, <fpage>102066</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102066</pub-id></citation>
</ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sun</surname> <given-names>K.</given-names></name> <name><surname>Tang</surname> <given-names>L.</given-names></name> <name><surname>Qian</surname> <given-names>J.</given-names></name> <name><surname>Wang</surname> <given-names>G.</given-names></name> <name><surname>Lou</surname> <given-names>C.</given-names></name></person-group> (<year>2021</year>). <article-title>A deep learning-based pm2.5 concentration estimator</article-title>. <source>Displays</source> <volume>69</volume>, <fpage>102072</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102072</pub-id></citation>
</ref>
<ref id="B38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>ur Rehman</surname> <given-names>M.</given-names></name> <name><surname>Nizami</surname> <given-names>I. F.</given-names></name> <name><surname>Majid</surname> <given-names>M.</given-names></name></person-group> (<year>2022</year>). <article-title>Deeprpn-biqa: Deep architectures with region proposal network for natural-scene and screen-content blind image quality assessment</article-title>. <source>Displays</source> <volume>71</volume>, <fpage>102101</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102101</pub-id></citation>
</ref>
<ref id="B39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>Z.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name> <name><surname>Sheikh</surname> <given-names>H. R.</given-names></name> <name><surname>Simoncelli</surname> <given-names>E. P.</given-names></name></person-group> (<year>2004</year>). <article-title>Image quality assessment: from error visibility to structural similarity</article-title>. <source>IEEE Trans. Image Process</source>. <volume>13</volume>, <fpage>600</fpage>&#x02013;<lpage>612</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2003.819861</pub-id><pub-id pub-id-type="pmid">15376593</pub-id></citation></ref>
<ref id="B40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>Z.</given-names></name> <name><surname>Li</surname> <given-names>Q.</given-names></name></person-group> (<year>2010</year>). <article-title>Information content weighting for perceptual image quality assessment</article-title>. <source>IEEE Trans. Image Process</source>. <volume>20</volume>, <fpage>1185</fpage>&#x02013;<lpage>1198</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2010.2092435</pub-id><pub-id pub-id-type="pmid">21078577</pub-id></citation></ref>
<ref id="B41">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>Z.</given-names></name> <name><surname>Simoncelli</surname> <given-names>E. P.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name></person-group> (<year>2003</year>). <article-title>Multiscale structural similarity for image quality assessment</article-title>. <source>Conf. Signals Syst. Comput</source>. <volume>2</volume>, <fpage>1398</fpage>&#x02013;<lpage>1402</lpage>. <pub-id pub-id-type="doi">10.1109/ACSSC.2003.1292216</pub-id></citation>
</ref>
<ref id="B42">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wu</surname> <given-names>J.</given-names></name> <name><surname>Lin</surname> <given-names>W.</given-names></name> <name><surname>Shi</surname> <given-names>G.</given-names></name> <name><surname>Liu</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). <article-title>Perceptual quality metric with internal generative mechanism</article-title>. <source>IEEE Trans. Image Process</source>. <volume>22</volume>, <fpage>43</fpage>&#x02013;<lpage>54</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2012.2214048</pub-id><pub-id pub-id-type="pmid">22910116</pub-id></citation></ref>
<ref id="B43">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wu</surname> <given-names>Y.</given-names></name> <name><surname>Wang</surname> <given-names>Z.</given-names></name> <name><surname>Chen</surname> <given-names>W.</given-names></name> <name><surname>Lin</surname> <given-names>L.</given-names></name> <name><surname>Wei</surname> <given-names>H.</given-names></name> <name><surname>Zhao</surname> <given-names>T.</given-names></name></person-group> (<year>2021</year>). <article-title>Perceptual vvc quantization refinement with ensemble learning</article-title>. <source>Displays</source> <volume>70</volume>, <fpage>102103</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102103</pub-id></citation>
</ref>
<ref id="B44">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>T.</given-names></name> <name><surname>Zhu</surname> <given-names>Y.</given-names></name> <name><surname>Peng</surname> <given-names>L.</given-names></name> <name><surname>Cao</surname> <given-names>Y.</given-names></name> <name><surname>Zhao</surname> <given-names>X.</given-names></name> <name><surname>Meng</surname> <given-names>F.</given-names></name> <etal/></person-group>. (<year>2022</year>). <article-title>Artificial intelligence assisted identification of therapy history from periapical films for dental root canal</article-title>. <source>Displays</source> <volume>71</volume>, <fpage>102119</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2021.102119</pub-id></citation>
</ref>
<ref id="B45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xue</surname> <given-names>W.</given-names></name> <name><surname>Zhang</surname> <given-names>L.</given-names></name> <name><surname>Mou</surname> <given-names>X.</given-names></name> <name><surname>Bovik</surname> <given-names>A. C.</given-names></name></person-group> (<year>2013</year>). <article-title>Gradient magnitude similarity deviation: a highly efficient perceptual image quality index</article-title>. <source>IEEE Trans. Image Process</source>. <volume>23</volume>, <fpage>684</fpage>&#x02013;<lpage>695</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2013.2293423</pub-id><pub-id pub-id-type="pmid">26270911</pub-id></citation></ref>
<ref id="B46">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ye</surname> <given-names>P.</given-names></name> <name><surname>Wu</surname> <given-names>X.</given-names></name> <name><surname>Gao</surname> <given-names>D.</given-names></name> <name><surname>Deng</surname> <given-names>S.</given-names></name> <name><surname>Xu</surname> <given-names>N.</given-names></name> <name><surname>Chen</surname> <given-names>J.</given-names></name></person-group> (<year>2020</year>). <article-title>Dp3 signal as a neuro-indictor for attentional processing of stereoscopic contents in varied depths within the &#x02018;comfort zone&#x02019;</article-title>. <source>Displays</source>. 63, 101953. <pub-id pub-id-type="doi">10.1016/j.displa.2020.101953</pub-id></citation>
</ref>
<ref id="B47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ye</surname> <given-names>X.</given-names></name> <name><surname>Chen</surname> <given-names>Y.</given-names></name> <name><surname>Sang</surname> <given-names>X.</given-names></name> <name><surname>Liu</surname> <given-names>B.</given-names></name> <name><surname>Chen</surname> <given-names>D.</given-names></name> <name><surname>Wang</surname> <given-names>P.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>An optimization method for parameters configuration of the light field display based on subjective evaluation</article-title>. <source>Displays</source> <volume>62</volume>, <fpage>101945</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2020.101945</pub-id></citation>
</ref>
<ref id="B48">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yu</surname> <given-names>H.</given-names></name> <name><surname>Akita</surname> <given-names>T.</given-names></name></person-group> (<year>2020</year>). <article-title>Influence of ambient-tablet pc luminance ratio on legibility and visual fatigue during long-term reading in low lighting environment</article-title>. <source>Displays</source> <volume>62</volume>, <fpage>101943</fpage>. <pub-id pub-id-type="doi">10.1016/j.displa.2020.101943</pub-id></citation>
</ref>
<ref id="B49">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>L.</given-names></name> <name><surname>Shen</surname> <given-names>Y.</given-names></name> <name><surname>Li</surname> <given-names>H.</given-names></name></person-group> (<year>2014</year>). <article-title>Vsi: a visual saliency-induced index for perceptual image quality assessment</article-title>. <source>IEEE Trans. Image Process</source>. <volume>23</volume>, <fpage>4270</fpage>&#x02013;<lpage>4281</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2014.2346028</pub-id><pub-id pub-id-type="pmid">25122572</pub-id></citation></ref>
<ref id="B50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>L.</given-names></name> <name><surname>Zhang</surname> <given-names>L.</given-names></name> <name><surname>Mou</surname> <given-names>X.</given-names></name> <name><surname>Zhang</surname> <given-names>D.</given-names></name></person-group> (<year>2011</year>). <article-title>Fsim: a feature similarity index for image quality assessment</article-title>. <source>IEEE Trans. Image Process</source>. <volume>20</volume>, <fpage>2378</fpage>&#x02013;<lpage>2386</lpage>. <pub-id pub-id-type="doi">10.1109/TIP.2011.2109730</pub-id><pub-id pub-id-type="pmid">21292594</pub-id></citation></ref>
<ref id="B51">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Zhou</surname> <given-names>F.</given-names></name> <name><surname>Yao</surname> <given-names>R.</given-names></name> <name><surname>Zhang</surname> <given-names>B.</given-names></name> <collab>others</collab></person-group>. (<year>2018</year>). <source>Quality Assessment Database for Super-Resolved Images: QADS</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="http://www.vista.ac.cn/super-resolution/">http://www.vista.ac.cn/super-resolution/</ext-link><pub-id pub-id-type="pmid">30762547</pub-id></citation></ref>
<ref id="B52">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhu</surname> <given-names>R.</given-names></name> <name><surname>Zhou</surname> <given-names>F.</given-names></name> <name><surname>Xue</surname> <given-names>J.-H.</given-names></name></person-group> (<year>2018</year>). <article-title>Mvssim: a quality assessment index for hyperspectral images</article-title>. <source>Neurocomputing</source> <volume>272</volume>, <fpage>250</fpage>&#x02013;<lpage>257</lpage>. <pub-id pub-id-type="doi">10.1016/j.neucom.2017.06.073</pub-id></citation>
</ref>
</ref-list> 
</back>
</article> 