<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2021.754898</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>A Wilcoxon&#x02013;Mann&#x02013;Whitney Test for Latent Variables</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Dehaene</surname> <given-names>Heidelinde</given-names></name>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1422849/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>De Neve</surname> <given-names>Jan</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/765767/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Rosseel</surname> <given-names>Yves</given-names></name>
</contrib>
</contrib-group>
<aff><institution>Department of Data Analysis, Ghent University</institution>, <addr-line>Ghent</addr-line>, <country>Belgium</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Alexander Robitzsch, IPN-Leibniz Institute for Science and Mathematics Education, Germany</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Daniel Kasper, University of Hamburg, Germany; Christoph Koenig, Goethe University Frankfurt, Germany; Max Auerswald, University of Ulm, Germany</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Heidelinde Dehaene  <email>heidelinde.dehaene&#x00040;ugent.be</email></corresp>
<fn fn-type="other" id="fn001"><p>This article was submitted to Quantitative Psychology and Measurement, a section of the journal Frontiers in Psychology</p></fn></author-notes>
<pub-date pub-type="epub">
<day>15</day>
<month>11</month>
<year>2021</year>
</pub-date>
<pub-date pub-type="collection">
<year>2021</year>
</pub-date>
<volume>12</volume>
<elocation-id>754898</elocation-id>
<history>
<date date-type="received">
<day>07</day>
<month>08</month>
<year>2021</year>
</date>
<date date-type="accepted">
<day>11</day>
<month>10</month>
<year>2021</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2021 Dehaene, De Neve and Rosseel.</copyright-statement>
<copyright-year>2021</copyright-year>
<copyright-holder>Dehaene, De Neve and Rosseel</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license> </permissions>
<abstract><p>We propose an extension of the Wilcoxon&#x02013;Mann&#x02013;Whitney test to compare two groups when the outcome variable is latent. We empirically demonstrate that the test can have superior power properties relative to tests based on Structural Equation Modeling for a variety of settings. In addition, several other advantages of the Wilcoxon&#x02013;Mann&#x02013;Whitney test are retained such as robustness to outliers and good small sample performance. We demonstrate the proposed methodology on a case study.</p></abstract>
<kwd-group>
<kwd>rank test</kwd>
<kwd>measurement error</kwd>
<kwd>indicators</kwd>
<kwd>robustness</kwd>
<kwd>nonparametric inference</kwd>
<kwd>group comparison</kwd>
</kwd-group>
<contract-sponsor id="cn001">Bijzonder Onderzoeksfonds UGent<named-content content-type="fundref-id">10.13039/501100007229</named-content></contract-sponsor>
<counts>
<fig-count count="3"/>
<table-count count="5"/>
<equation-count count="13"/>
<ref-count count="33"/>
<page-count count="13"/>
<word-count count="8928"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>1. Introduction</title>
<p>Consider a study where the interest is in the association between employment (yes/no) and the construct depression. The latter is quantified by the score on three questionnaires: the Patient Health Questionnaire-9 (PHQ-9; Spitzer et al., <xref ref-type="bibr" rid="B29">1999</xref>; Kroenke et al., <xref ref-type="bibr" rid="B18">2001</xref>; Kroenke and Spitzer, <xref ref-type="bibr" rid="B17">2002</xref>), the Center for Epidemiological Studies Depression Scale-10 (CESD-10; Andresen et al., <xref ref-type="bibr" rid="B2">1994</xref>), and the eight-item PROMIS Depression Short Form (PROMIS D-8; 8b short form; Pilkonis et al., <xref ref-type="bibr" rid="B24">2011</xref>). The data originate from Amtmann et al. (<xref ref-type="bibr" rid="B1">2014</xref>), where the psychometric properties of these questionnaires were examined given that they all aim to screen for high depressive symptoms or major depressive disorder (MDD). <xref ref-type="fig" rid="F1">Figure 1</xref> displays these scores for 455 employed and unemployed individuals living with multiple sclerosis (MS). Scores for the PHQ-9 can range from 0 to 27, for the CESD-10 from 0 to 30 and the scores for PROMIS D-8 are reported on a standardized scale with mean 50 and standard deviation 10.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>Boxplot of the scores for employed and unemployed subjects on three questionnaires measuring depression: <bold>A</bold>, the Patient Health Questionnaire-9 (PHQ-9, scale [0, 27]), <bold>B</bold>, the Center for Epidemiological Studies Depression Scale-10 (CESD-10, scale [0, 30]) and, <bold>C</bold>, the eight-item PROMIS Depression Short Form (PROMIS D-8, scores reported on a standardized scale with mean 50 and standard deviation 10). Higher scores indicate more symptoms of depression.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-12-754898-g0001.tif"/>
</fig>
<p>At first sight, the Wilcoxon&#x02013;Mann&#x02013;Whitney (WMW) test seems an appropriate choice to compare these two groups given the skewness and presence of outliers depicted in <xref ref-type="fig" rid="F1">Figure 1</xref>. For each questionnaire, a WMW test can be carried out and for the sake of illustration, the WMW test will be introduced below for the PHQ-9 scores. The test considers the null hypothesis claiming that the distribution of the PHQ-9 scores is the same for both groups against the alternative stating that the probability that a subject of the unemployed group has a higher PHQ-9 score as compared to a subject of the employed group is different from 50%. The test statistic reflects this probability, also denoted as the probabilistic index, and the sole use of ranks hereby results in robustness against outliers. Under the null hypothesis and in the absence of ties, the sampling distribution of the test statistic only depends on the sample sizes of both groups, hence resulting in a distribution-free test (Thas, <xref ref-type="bibr" rid="B30">2010</xref>). The WMW test often has superior power compared with the <italic>t</italic>-test for heavy tailed distributions (Blair and Higgins, <xref ref-type="bibr" rid="B4">1980</xref>). For an exponential distribution, for example, the <italic>t</italic>-test needs approximately three times more observations than the WMW test to attain the same power (Van der Vaart, <xref ref-type="bibr" rid="B32">2000</xref>; Hollander et al., <xref ref-type="bibr" rid="B14">2013</xref>). Albeit both the WMW and the <italic>t</italic>-test aim to compare two samples, they apply different hypotheses and according effect sizes (i.e. probabilistic index vs. comparison of means) hence leading to different test properties.</p>
<p>Although the WMW test seems to be an appropriate choice in the current example, there is a complication as the outcome of interest, depression, is a latent variable. Stated otherwise, it is a variable that is not directly observed, but rather theoretically postulated or empirically inferred from observed variables (i.e. the indicators or proxies). Switching to latent variables, one cannot directly apply the methods that were designed for observed variables, because a measurement model that connects the latent variable to observed variables has to be postulated. Typically, continuous data with latent variables are analyzed via Factor Analysis (FA) or Structural Equation Modeling (SEM). The classical SEM has an optimal performance with respect to hypothesis testing when data are multivariate normally distributed. Given the skewness of the data, it is our interest to study whether the WMW test can be used in the context of latent variables while maintaining the attractive properties mentioned before. Applying the WMW test naively on each questionnaire results in a significant difference between the two groups on the first and second questionnaire (<italic>p</italic> &#x0003D; 0.003 and <italic>p</italic> &#x0003D; 0.04 respectively) but not on the third questionnaire (<italic>p</italic> &#x0003D; 0.12), making it difficult to obtain a global conclusion whether there is a significant association between depression and employment or not. In addition, these <italic>p</italic>-values do not take the measurement error into account. To the best of our knowledge, an extension of the WMW that takes into account those two essential aspects (i.e. combining the information of multiple indicators where measurement error is inherent) is not available yet and therefore the objective of this paper is to extend the WMW test in the context of latent variables with the main focus on hypothesis testing. The paper is organized as follows. In Section 2 we propose a WMW test for latent variables. Section 3 empirically contrasts the new methodology with existing methods in different settings by the use of a simulation study. Section 4 presents a case study and Section 5 provides a conclusion and discussion.</p>
</sec>
<sec id="s2">
<title>2. Method</title>
<p>We first introduce a measurement model that relates the latent variable to multiple indicators (or proxies). We then formulate the hypotheses of interest and demonstrate how they can be tested.</p>
<sec>
<title>2.1. Measurement Model</title>
<p>The measurement model that relates the latent variable &#x003B7; to a set of indicators <italic>Y</italic><sub><italic>p</italic></sub> (<italic>p</italic> &#x0003D; 1, &#x02026;, <italic>P</italic>), is given by</p>
<disp-formula id="E1"><label>(1)</label><mml:math id="M1"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) denotes a strictly monotone function and &#x003B5;<sub><italic>p</italic></sub> denotes the measurement error. To distinguish between the two groups, we use the notation (<italic>Y</italic><sub><italic>p</italic></sub>, &#x003B7;, &#x003B5;<sub><italic>p</italic></sub>) for the first group and (<inline-formula><mml:math id="M2"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>,</mml:mo><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup><mml:mo>,</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula>) for the second. We define the reliability of an indicator <italic>Y</italic><sub><italic>p</italic></sub> as the proportion of variance in the indicator <italic>Y</italic><sub><italic>p</italic></sub> that is not explained by the measurement error &#x003B5;<sub><italic>p</italic></sub> and thus reflecting the quality of that indicator (Nunnally and Bernstein, <xref ref-type="bibr" rid="B22">1994</xref>; Kline, <xref ref-type="bibr" rid="B16">2015</xref>). Specifically,</p>
<disp-formula id="E2"><label>(2)</label><mml:math id="M3"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mtext class="textrm" mathvariant="normal">Rel</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mtext class="textrm" mathvariant="normal">var</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mtext class="textrm" mathvariant="normal">var</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mtext class="textrm" mathvariant="normal">var</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:mfrac><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>The variance of the measurement error &#x003B5;<sub><italic>p</italic></sub> can be replaced by its empirical counterpart in order to obtain an estimate of the reliability of an indicator. We further elaborate on this estimation in the subsequent subsection.</p>
<p>Similar to classic factor analysis and SEM, comparing the latent means between two groups by the use of indicators is only feasible when assuming measurement invariance. Hence, the function <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) is assumed to be equal for the two groups and this needs to be confirmed by using e.g. SEM. In contrast with FA and SEM, where <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) is assumed to be linear, we impose a less stringent specification, i.e. monotone without imposing linearity. Garcia-Marques et al. (<xref ref-type="bibr" rid="B13">2014</xref>) pointed out that a curvelinear trend between a latent variable and its indicator may exist in psychological research due to e.g. ceiling or floor effects. An example of a floor effect is observed by Amtmann et al. (<xref ref-type="bibr" rid="B1">2014</xref>) for the PROMIS D-8 questionnaire and hence portrayed in the right panel of <xref ref-type="fig" rid="F1">Figure 1</xref>. This demonstrates that assuming linearity can sometimes be an incorrect representation of reality.</p>
<p>The focus in this paper does not lie on estimating the measurement model and will thus be treated as a nuisance, in contrast with classic FA and SEM.</p>
</sec>
<sec>
<title>2.2. Hypotheses and Testing Procedure</title>
<p>We are interested in testing the following hypotheses which are expressed in terms of the distribution of the latent variable &#x003B7; (i.e. <italic>F</italic><sub>&#x003B7;</sub>):</p>
<disp-formula id="E3"><label>(3)</label><mml:math id="M4"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>If &#x003B7;<sub><italic>i</italic></sub> (<italic>i</italic> &#x0003D; 1, &#x02026;, <italic>m</italic>) and <inline-formula><mml:math id="M5"><mml:msubsup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> (<italic>j</italic> &#x0003D; 1, &#x02026;, <italic>n</italic>) would be observable, the WMW test statistic (in the absence of ties) is given by <inline-formula><mml:math id="M6"><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover><mml:mo>-</mml:mo><mml:mn>0</mml:mn><mml:mo>.</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>/</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C3;</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub></mml:math></inline-formula> where <inline-formula><mml:math id="M7"><mml:msub><mml:mrow><mml:mi>&#x003C3;</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msqrt><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>m</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:mi>n</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>/</mml:mo><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>m</mml:mi><mml:mi>n</mml:mi><mml:mn>12</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:msqrt></mml:math></inline-formula> and</p>
<disp-formula id="E4"><label>(4)</label><mml:math id="M8"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mover accent="true"><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi><mml:mi>n</mml:mi></mml:mrow></mml:mfrac><mml:mstyle displaystyle="true"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:mstyle displaystyle="true"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:mtext>I</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x0003C;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>with I(&#x000B7;) the indicator function (Wilcoxon, <xref ref-type="bibr" rid="B33">1945</xref>; Mann and Whitney, <xref ref-type="bibr" rid="B21">1947</xref>). Here, <inline-formula><mml:math id="M9"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula> is an unbiased estimator for <italic>P</italic>(&#x003B7; &#x0003C; &#x003B7;<sup>&#x0002A;</sup>), i.e. the probability that a randomly selected subject from group 1 has a lower outcome than a randomly selected subject from group 2.</p>
<p>In practice, &#x003B7;<sub><italic>i</italic></sub> and <inline-formula><mml:math id="M10"><mml:msubsup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> are unobservable and therefore Equation (4) can not be computed. Instead, we only observe <italic>Y</italic><sub><italic>pi</italic></sub> and <inline-formula><mml:math id="M11"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> which are related to &#x003B7;<sub><italic>i</italic></sub> and <inline-formula><mml:math id="M12"><mml:msubsup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> via measurement model (1). Let WMW(<inline-formula><mml:math id="M13"><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>;</mml:mo><mml:mi>&#x003B1;</mml:mi></mml:math></inline-formula>) denote the WMW test applied to the indicators <italic>Y</italic><sub><italic>p</italic></sub> and <inline-formula><mml:math id="M14"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> and where &#x003B1; denotes the level of significance. Under location-shift, i.e. <inline-formula><mml:math id="M15"><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mover class="overset"><mml:mrow><mml:mo>=</mml:mo></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:mover><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>&#x0002B;</mml:mo><mml:mo>&#x00394;</mml:mo></mml:math></inline-formula>, WMW(<inline-formula><mml:math id="M16"><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>;</mml:mo><mml:mi>&#x003B1;</mml:mi></mml:math></inline-formula>) is an unbiased test for <inline-formula><mml:math id="M17"><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow></mml:msub></mml:math></inline-formula> vs. <inline-formula><mml:math id="M18"><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow></mml:msub></mml:math></inline-formula>, meaning that the rejection level does not exceed &#x003B1; when <inline-formula><mml:math id="M19"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow></mml:msub></mml:math></inline-formula> and that it is at least &#x003B1; when <inline-formula><mml:math id="M20"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow></mml:msub></mml:math></inline-formula> (Lehmann, <xref ref-type="bibr" rid="B19">1951</xref>).</p>
<p>Under the model</p>
<disp-formula id="E5"><label>(5)</label><mml:math id="M21"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;and&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>,</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;with&#x000A0;</mml:mtext><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mover class="overset"><mml:mrow><mml:mo>=</mml:mo></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:mover><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>WMW(<inline-formula><mml:math id="M22"><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>;</mml:mo><mml:mi>&#x003B1;</mml:mi></mml:math></inline-formula>) is also an unbiased test for <inline-formula><mml:math id="M23"><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula> vs. <inline-formula><mml:math id="M24"><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula>. Indeed, when assuming equal distributions for the measurement error over the two groups per indicator, it follows that when <inline-formula><mml:math id="M25"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula> then <inline-formula><mml:math id="M26"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow></mml:msub></mml:math></inline-formula> so that the rejection level of WMW(<inline-formula><mml:math id="M27"><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>;</mml:mo><mml:mi>&#x003B1;</mml:mi></mml:math></inline-formula>) does not exceed &#x003B1; when <inline-formula><mml:math id="M28"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula>. Secondly, when <inline-formula><mml:math id="M29"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula> then <inline-formula><mml:math id="M30"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow></mml:msub></mml:math></inline-formula> so that the rejection level is at least &#x003B1; when <inline-formula><mml:math id="M31"><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula> and under the assumption of location-shift (i.e. <inline-formula><mml:math id="M32"><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mover class="overset"><mml:mrow><mml:mo>=</mml:mo></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:mover><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>&#x0002B;</mml:mo><mml:mo>&#x00394;</mml:mo></mml:math></inline-formula>). Consequently, the test statistic</p>
<disp-formula id="E6"><label>(6)</label><mml:math id="M33"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mfrac><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>m</mml:mi><mml:mi>n</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mo>-</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msup><mml:mstyle displaystyle="false"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:mstyle displaystyle="false"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:mtext>I</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>&#x0003C;</mml:mo><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mn>0</mml:mn><mml:mo>.</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C3;</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>results in an unbiased test for <inline-formula><mml:math id="M34"><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub><mml:mstyle class="text"><mml:mtext class="textrm" mathvariant="normal">vs.</mml:mtext></mml:mstyle><mml:msub><mml:mrow><mml:mi>H</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi></mml:mrow></mml:msub><mml:mo>:</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02260;</mml:mo><mml:msub><mml:mrow><mml:mi>F</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula>. In summary, we can apply the WMW test on the observed data to test hypotheses concerning the latent variable. This testing procedure does not require that <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;), var(&#x003B5;<sub><italic>p</italic></sub>), var(<inline-formula><mml:math id="M35"><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula>) have to be estimated nor that <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) needs to be linear. The strong assumption of equal distributions for the measurement error is required for the theoretical validation of this method. However, violations of this assumption will appear to have no detrimental consequences with respect to hypothesis testing in our simulation study (see Section 3).</p>
<p>When multiple indicators are available, we propose to aggregate them to obtain a new indicator where the above-mentioned test rationale still holds. This aggregated indicator can be superior in terms of its reliability in comparison with the original indicators, but it however requires estimates of the measurement error variance. The aggregated indicator is obtained by making a linear combination of the original indicators while preserving measurement invariance as mentioned in Section 2.1. Therefore, the construction of this linear combination is based on all data of an indicator <italic>p</italic>, i.e. both <inline-formula><mml:math id="M36"><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mstyle class="text"><mml:mtext class="textrm" mathvariant="normal">and</mml:mtext></mml:mstyle><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula>. Let <italic>Z</italic><sub><italic>pk</italic></sub> denote an observation from the pooled sample <italic>Y</italic><sub><italic>p</italic></sub> and <inline-formula><mml:math id="M37"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> where <italic>k</italic> &#x0003D; 1, &#x02026;, (<italic>m</italic>&#x0002B;<italic>n</italic>). We define the aggregated indicator as</p>
<disp-formula id="E7"><label>(7)</label><mml:math id="M38"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>P</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:msub><mml:mrow><mml:mi>a</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mi>Z</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mi>k</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where the weights a<sub><italic>p</italic></sub> in Equation (7) can be chosen according to two strategies. A first strategy is to simplify the aggregation to an unweighted mean of the standardized indicators. However, treating all indicators as equally important is a reduction of the complexity in reality where some indicators are more reliable than others. Therefore, we propose a second strategy where the weights a<sub><italic>p</italic></sub> in Equation (7) are chosen so that the estimated reliability of <inline-formula><mml:math id="M39"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> is maximized (Bentler, <xref ref-type="bibr" rid="B3">1968</xref>; Li, <xref ref-type="bibr" rid="B20">1997</xref>; Penev and Raykov, <xref ref-type="bibr" rid="B23">2006</xref>). In other words, a maximally reliable composite is constructed and the estimated reliability of <inline-formula><mml:math id="M40"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> is by construction at least equal to the highest estimated reliability of the separate indicators used in the aggregation.</p>
<p>In order to obtain the weights <italic>a</italic><sub><italic>p</italic></sub>, one needs to estimate the variance of the measurement error of the accompanying indicators, comprising the data of both groups. Given the data structure of our simulation study (Section 3) and case study (Section 4), we briefly elaborate on how this variance can be estimated when at least three indicators are available. For a more exhaustive explanation on how to estimate the variance of measurement error under different settings, we refer to De Neve and Dehaene (<xref ref-type="bibr" rid="B9">2021</xref>). Imposing a linear relationship among the indicators <italic>Y</italic><sub>1<italic>i</italic></sub> &#x0003D; <italic>h</italic>(&#x003B7;<sub><italic>i</italic></sub>)&#x0002B;&#x003B5;<sub>1<italic>i</italic></sub>, <italic>Y</italic><sub>2<italic>i</italic></sub> &#x0003D; <italic>a</italic><sub>2</sub>&#x0002B;<italic>b</italic><sub>2</sub><italic>h</italic>(&#x003B7;<sub><italic>i</italic></sub>)&#x0002B;&#x003B5;<sub>2<italic>i</italic></sub> and <italic>Y</italic><sub>3<italic>i</italic></sub> &#x0003D; <italic>a</italic><sub>3</sub>&#x0002B;<italic>b</italic><sub>3</sub><italic>h</italic>(&#x003B7;<sub><italic>i</italic></sub>)&#x0002B;&#x003B5;<sub>3<italic>i</italic></sub>, it follows that whenever cov(<italic>Y</italic><sub>1</sub>, <italic>Y</italic><sub>2</sub>) &#x02260;0, cov(<italic>Y</italic><sub>1</sub>, <italic>Y</italic><sub>3</sub>) &#x02260;0 and cov(<italic>Y</italic><sub>2</sub>, <italic>Y</italic><sub>3</sub>) &#x02260;0 that var(&#x003B5;<sub>1</sub>) = var(<italic>Y</italic><sub>1</sub>) - cov(<italic>Y</italic><sub>1</sub>, <italic>Y</italic><sub>2</sub>)cov(<italic>Y</italic><sub>1</sub>, <italic>Y</italic><sub>3</sub>)/cov(<italic>Y</italic><sub>2</sub>, <italic>Y</italic><sub>3</sub>). Although the assumption of a linear relationship among the indicators is required for these formulas, the results of our simulation study show no detrimental consequences of violations.</p>
<p>The test rationale that has been put forward earlier on, i.e. employ the WMW test on observed data to make inference with respect to latent variables, is still valid when using this aggregated indicator <inline-formula><mml:math id="M41"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula>. Moreover, this aggregation does not complicate the justification due to the flexibility of the measurement model, because the aggregated indicator can be rewritten as</p>
<disp-formula id="E8"><mml:math id="M42"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:msup><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msup><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup><mml:mo>,</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:msup><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msup><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mo>&#x000B7;</mml:mo></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>P</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:msub><mml:mrow><mml:mi>a</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mo>&#x000B7;</mml:mo></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mstyle displaystyle="true"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>P</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:msub><mml:mrow><mml:mi>a</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mi>k</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>and hence the same structure as described in measurement model (1) still holds. Therefore, we obtain the following test statistic:</p>
<disp-formula id="E9"><label>(8)</label><mml:math id="M43"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msup><mml:mrow><mml:mi>U</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>m</mml:mi><mml:mi>n</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mo>-</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msup><mml:mstyle displaystyle="false"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:mstyle displaystyle="false"><mml:munderover accentunder="false" accent="false"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>j</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>n</mml:mi></mml:mrow></mml:munderover></mml:mstyle><mml:mtext>I</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x0003C;</mml:mo><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mn>0</mml:mn><mml:mo>.</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C3;</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub></mml:mrow></mml:mfrac><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where the data of the two groups for the aggregated indicator are distinguished by using the notation <inline-formula><mml:math id="M44"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> and <inline-formula><mml:math id="M45"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mi>j</mml:mi></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:math></inline-formula> for respectively the first and second group. Similar with the standard WMW test, a <italic>p</italic>-value can be obtained by either using a permutation null distribution or a standard normal approximation (Wilcoxon, <xref ref-type="bibr" rid="B33">1945</xref>; Mann and Whitney, <xref ref-type="bibr" rid="B21">1947</xref>; Thas, <xref ref-type="bibr" rid="B30">2010</xref>).</p>
</sec>
<sec>
<title>2.3. Sample Size and Power Calculation</title>
<p>An approximate total sample size <italic>N</italic> (<italic>N</italic> &#x0003D; <italic>m</italic>&#x0002B;<italic>n</italic>) for a one-sided test with significance level &#x003B1; can be determined by using the formula from Hollander et al. (<xref ref-type="bibr" rid="B14">2013</xref>):</p>
<disp-formula id="E10"><label>(9)</label><mml:math id="M46"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mi>N</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B1;</mml:mi></mml:mrow></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mn>12</mml:mn><mml:mi>c</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:mi>c</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003B4;</mml:mi><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfrac><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where <italic>c</italic> reflects the ratio of the sample sizes of the two groups, i.e. <inline-formula><mml:math id="M47"><mml:mi>c</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>m</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:mi>n</mml:mi></mml:mrow></mml:mfrac></mml:math></inline-formula>, and &#x003B4; denotes the effect size under the alternative hypothesis, i.e. <inline-formula><mml:math id="M48"><mml:mi>P</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup><mml:mo>&#x0003C;</mml:mo><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow><mml:mrow><mml:mi>A</mml:mi><mml:mi>G</mml:mi><mml:mi>G</mml:mi></mml:mrow></mml:msubsup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>. Subsequently, the expected power can be deduced via</p>
<disp-formula id="E11"><label>(10)</label><mml:math id="M49"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:mi>&#x003B2;</mml:mi><mml:mo>=</mml:mo><mml:mo>&#x003A6;</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:msqrt><mml:mrow><mml:mi>N</mml:mi><mml:mn>12</mml:mn><mml:mi>c</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:mi>c</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003B4;</mml:mi><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:msqrt><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B1;</mml:mi></mml:mrow></mml:msub><mml:mo>)</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
</sec>
</sec>
<sec id="s3">
<title>3. Simulation Study</title>
<p>In order to assess the finite sample performance of the extended Wilcoxon&#x02013;Mann&#x02013;Whitney test, a simulation study is performed. All simulations and analyses are performed with R version 3.5.1 (R Core Team, <xref ref-type="bibr" rid="B25">2020</xref>).</p>
<p>Different scenarios are explored, all based on the following data generating process:</p>
<p><bold>Latent variables</bold>:</p>
<disp-formula id="E12"><label>(11)</label><mml:math id="M50"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>&#x003B7;</mml:mi></mml:mtd><mml:mtd><mml:mo>=</mml:mo><mml:mi>&#x003B2;</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:mi>&#x003B6;</mml:mi></mml:mtd></mml:mtr><mml:mtr></mml:mtr></mml:mtable><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mo>=</mml:mo><mml:msup><mml:mrow><mml:mi>&#x003B2;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mi>&#x003B6;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr></mml:mtr></mml:mtable></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p><bold>Observed variables/indicators</bold>:</p>
<disp-formula id="E13"><label>(12)</label><mml:math id="M51"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mtable columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>&#x003B7;</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub></mml:mtd></mml:mtr></mml:mtable><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mtable columnalign="left"><mml:mtr><mml:mtd><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>=</mml:mo><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup><mml:mo>&#x0002B;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>h</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:mtd></mml:mtr><mml:mtr></mml:mtr></mml:mtable></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>To explore the performance of the suggested methodology under different scenarios, the simulation study covers both normal, heavily tailed and skewed distributions for &#x003B6; and &#x003B6;<sup>&#x0002A;</sup> (and consequently &#x003B7; and &#x003B7;<sup>&#x0002A;</sup>): <inline-formula><mml:math id="M52"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">N</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>, <italic>t</italic><sub>5</sub>, <italic>Laplace</italic>(0, 1.25) and the standard exponential centered around zero. In the wide range of possible non-normal distributions, these distributions correspond to kurtosis and skewness values that can be encountered in practice and that are leptokurtic (Chou et al., <xref ref-type="bibr" rid="B5">1991</xref>). This positive kurtosis enables the examination of whether the superior power of the WMW test in heavier tailed distributions is carried over to the context of latent variables (Van der Vaart, <xref ref-type="bibr" rid="B32">2000</xref>; Hollander et al., <xref ref-type="bibr" rid="B14">2013</xref>). The superiority in heavier tailed distributions is also observed in small samples and therefore sample sizes in this simulation study are varied between 15, 50, and 100 observations in each group. As a result, exploration of the properties of different methods is possible without running into a ceiling effect with respect to the empirical power. By varying the variances of the measurement error, the influence of the reliability of the indicators on the different methods can be studied. We consider reliabilities of 60% and 80% which corresponds to indicators that can be considered as having a relatively weak and an adequate reliability respectively (Nunnally and Bernstein, <xref ref-type="bibr" rid="B22">1994</xref>; Jackson, <xref ref-type="bibr" rid="B15">2001</xref>). The function <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) is either linear or non-linear. In the non-linear case, the transformation <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) equals &#x003A6;<sup>&#x02212;1</sup>[<italic>F</italic>(&#x000B7;)] where F equals the <italic>t</italic>-distribution with 1 or 3 degrees of freedom. This relationship can be seen as a modified inverse logit function and was chosen since it recreates a curvilinear trend between a latent variable and its indicator, in accordance with the trend mentioned in Garcia-Marques et al. (<xref ref-type="bibr" rid="B13">2014</xref>). <xref ref-type="fig" rid="F2">Figure 2</xref> shows an example of such a curvilinear trend between a latent variable &#x003B7; and <italic>h</italic><sub><italic>p</italic></sub>(&#x003B7;).</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p>Illustration of the curvilinear trend between a latent variable &#x003B7; and the function <italic>h</italic><sub><italic>p</italic></sub>(&#x003B7;) as used in the simulation study.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-12-754898-g0002.tif"/>
</fig>
<p>Taking into account all parameters discussed up till now, the simulation study involves linear or non-linear functions <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;), four error distributions, three sample sizes and two reliabilities, resulting in 48 simulation scenarios. For all these 48 combinations, four settings are considered to gain insight with respect to the empirical consequences of violating the assumption of equality in distribution of the measurement error as postulated under model (5) in Section 2.2.</p>
<list list-type="bullet">
<list-item><p>Setting 1: <inline-formula><mml:math id="M53"><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mover class="overset"><mml:mrow><mml:mo>=</mml:mo></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:mover><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula>, i.e. a correctly specified model.</p></list-item>
</list>
<p>In setting 2 and 3, the model is misspecified.</p>
<list list-type="bullet">
<list-item><p>Setting 2: <inline-formula><mml:math id="M54"><mml:mi>c</mml:mi><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mover class="overset"><mml:mrow><mml:mo>=</mml:mo></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:mover><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> with <italic>c</italic>&#x02260;1, i.e. the variance of the measurement error differs across groups, where <italic>c</italic> is chosen in such a way that the reliability of all indicators in the second group is consistently about 5% lower than in group 1.</p></list-item>
<list-item><p>Setting 3: <inline-formula><mml:math id="M55"><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mover class="overset"><mml:mrow><mml:mo>&#x02260;</mml:mo></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:mover><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> but <inline-formula><mml:math id="M56"><mml:mrow><mml:mtext>Var(</mml:mtext><mml:msub><mml:mi>&#x003B5;</mml:mi><mml:mi>p</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mtext>Var</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:msubsup><mml:mi>&#x003B5;</mml:mi><mml:mi>p</mml:mi><mml:mo>*</mml:mo></mml:msubsup><mml:mtext>)</mml:mtext></mml:mrow></mml:math></inline-formula>, i.e. the distribution of the measurement error differs across groups (a normal distribution and Laplace distribution respectively), but the variance is equal.</p></list-item>
</list>
<p>In the last setting, we consider a correctly specified model but with an indicator with very low reliability:</p>
<list list-type="bullet">
<list-item><p>Setting 4: <inline-formula><mml:math id="M57"><mml:msub><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mover class="overset"><mml:mrow><mml:mo>=</mml:mo></mml:mrow><mml:mrow><mml:mi>d</mml:mi></mml:mrow></mml:mover><mml:msubsup><mml:mrow><mml:mi>&#x003B5;</mml:mi></mml:mrow><mml:mrow><mml:mi>p</mml:mi></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula>, <italic>Y</italic><sub>3</sub> and <inline-formula><mml:math id="M58"><mml:msubsup><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow><mml:mrow><mml:mo>*</mml:mo></mml:mrow></mml:msubsup></mml:math></inline-formula> have a reliability of only 20%.</p></list-item>
</list>
<p>In order to assess the impact of all these parameters on both the empirical Type I error rate and power, the parameter &#x003B2;<sup>&#x0002A;</sup> as defined in Equation (11) is varied. The exact parameter value depends on the error distribution, but it is chosen so that the probabilistic index <italic>P</italic>(&#x003B7; &#x0003C; &#x003B7;<sup>&#x0002A;</sup>) equals 50% and 65% under the null and alternative hypotheses respectively. The rationale for this probabilistic index can be traced back to the simple relationship between a probabilistic index and the standardized difference as described in Cohen (<xref ref-type="bibr" rid="B8">1988</xref>) and De Schryver and De Neve (<xref ref-type="bibr" rid="B10">2019</xref>). By exploiting this relationship, a probabilistic index of 65% coincides with a standardized effect of 0.55 and is hence defined as a medium effect according to Cohen (<xref ref-type="bibr" rid="B8">1988</xref>).</p>
<p>For each setting, 1,000 Monte Carlo simulation runs are used to evaluate and contrast the performance of six methods. In the first method, further referred to as <bold>WMW&#x02013;max rel</bold>, a maximally reliable composite is used as input for the WMW test. In this simulation study, the weights <italic>a</italic><sub><italic>p</italic></sub> are obtained via an optimisation function, but an analytic solution based on Bentler (<xref ref-type="bibr" rid="B3">1968</xref>) is also possible. The variance of the measurement error of each indicator needed for this optimization process was estimated by using the formulas presented in Section 2.2. The second method, further referred to as <bold>WMW&#x02013;mean</bold>, also applies the WMW test on an aggregated indicator, but here the aggregation is simplified to the mean of the standardized indicators. This comparison enables us to study the influence of estimating weights according to the quality of the individual indicators. The third method uses the maximally reliable composite as input for a Welch <italic>t</italic>-test while the fourth method has the unweighted mean as input variable for a Welch <italic>t</italic>-test. These two methods are accordingly further referred to as <italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold> and <italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold> and form a parametric alternative for the first and second method. The fifth and sixth method are based on SEM with equal group loadings and intercepts and are currently the most common methods to compare a latent variable between two groups. We include both SEM with and without correction for non-normal data, further referred as <bold>SEM</bold> and <bold>SEM&#x02013;correction</bold>. The correction for non-normal data refers to the use of robust Satorra-Bentler standard errors (Satorra and Bentler, <xref ref-type="bibr" rid="B27">1994</xref>). <bold>SEM</bold> and <bold>SEM&#x02013;correction</bold> can also be seen as a parametric counterpart of the proposed WMW extension, but unlike the Welch <italic>t</italic>-test, SEM does take into account the measurement error of the indicators. The SEMs are implemented by using lavaan (Rosseel, <xref ref-type="bibr" rid="B26">2012</xref>).</p>
<p>For sample size <italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15, <italic>p</italic>-values for the methods based on SEM and the Welch <italic>t</italic>-test were obtained by using a permutation null model, to ensure a fair comparison between the six different methods. For larger samples, inference was conducted by relying on the asymptotic distribution of the respective test statistics.</p>
<p>The R code to recreate the simulation study is available in the <xref ref-type="supplementary-material" rid="SM1">Supplementary Material</xref>.</p>
<sec>
<title>3.1. Results</title>
<p>Because the Type I error rate is correctly controlled for almost all methods in almost all settings, we do not display these tables in the main text but provide them in the <xref ref-type="supplementary-material" rid="SM2">Supplementary Material</xref> (i.e., Tables 1&#x02013;4 in <xref ref-type="supplementary-material" rid="SM2">Supplementary Data Sheet 2</xref>). For 3 cases, it can be noted that the methods relying on SEM are too liberal, i.e. the empirical Type I error rate is 6.9%. On the other hand, one can observe that the methods relying on the WMW test are too conservative in 4 cases for setting 2 with an empirical Type I error rate of 3.0 or 3.2%. These 7 deviations are indicated in bold in the corresponding tables in the <xref ref-type="supplementary-material" rid="SM2">Supplementary Material</xref>. It is noteworthy that even under model misspecification (i.e. setting 2 and 3) the Type I error rate is controlled for by all six methods. Hence, the assumption of distributionally homogeneous errors as postulated in Equation (5) is in practice less stringent. The results with respect to the empirical power are summarized in <xref ref-type="table" rid="T1">Tables 1</xref>&#x02013;<xref ref-type="table" rid="T4">4</xref>.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Empirical power from the simulation study with the indicators having an overall reliability of 80% and a linear relationship with the latent variable.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th/>
<th valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Linear</bold></th>
</tr>
<tr>
<th/>
</tr>
</thead>
<tbody>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><inline-formula><mml:math id="M59"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">N</mml:mi></mml:mrow><mml:mtext>&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">29.7</td>
<td valign="top" align="center">31.0</td>
<td valign="top" align="center"><bold>33.2</bold></td>
<td valign="top" align="center">32.9</td>
<td valign="top" align="center">32.4</td>
<td valign="top" align="center">32.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">28.4</td>
<td valign="top" align="center">29.2</td>
<td valign="top" align="center">30.7</td>
<td valign="top" align="center"><bold>31.3</bold></td>
<td valign="top" align="center">30.0</td>
<td valign="top" align="center">30.3</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">29.5</td>
<td valign="top" align="center">29.3</td>
<td valign="top" align="center">31.8</td>
<td valign="top" align="center"><bold>32.3</bold></td>
<td valign="top" align="center">32.0</td>
<td valign="top" align="center">31.9</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">25.8</td>
<td valign="top" align="center">26.2</td>
<td valign="top" align="center">28.5</td>
<td valign="top" align="center">27.4</td>
<td valign="top" align="center"><bold>29.4</bold></td>
<td valign="top" align="center">28.6</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">78.7</td>
<td valign="top" align="center">79.1</td>
<td valign="top" align="center">81.7</td>
<td valign="top" align="center">81.4</td>
<td valign="top" align="center"><bold>83.2</bold></td>
<td valign="top" align="center">83.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">79.5</td>
<td valign="top" align="center">78.8</td>
<td valign="top" align="center">81.0</td>
<td valign="top" align="center">80.4</td>
<td valign="top" align="center">82.0</td>
<td valign="top" align="center"><bold>82.1</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">79.2</td>
<td valign="top" align="center">79.9</td>
<td valign="top" align="center">81.2</td>
<td valign="top" align="center">81.8</td>
<td valign="top" align="center"><bold>82.2</bold></td>
<td valign="top" align="center">82.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">74.8</td>
<td valign="top" align="center">72.3</td>
<td valign="top" align="center">77.6</td>
<td valign="top" align="center">74.4</td>
<td valign="top" align="center"><bold>80.0</bold></td>
<td valign="top" align="center"><bold>80.0</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">98.2</td>
<td valign="top" align="center">98.1</td>
<td valign="top" align="center">98.5</td>
<td valign="top" align="center">98.6</td>
<td valign="top" align="center"><bold>98.7</bold></td>
<td valign="top" align="center"><bold>98.7</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">97.7</td>
<td valign="top" align="center">97.7</td>
<td valign="top" align="center">98.2</td>
<td valign="top" align="center"><bold>98.3</bold></td>
<td valign="top" align="center"><bold>98.3</bold></td>
<td valign="top" align="center"><bold>98.3</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">97.6</td>
<td valign="top" align="center">97.5</td>
<td valign="top" align="center">98.1</td>
<td valign="top" align="center"><bold>98.2</bold></td>
<td valign="top" align="center"><bold>98.2</bold></td>
<td valign="top" align="center"><bold>98.2</bold></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">97.4</td>
<td valign="top" align="center">95.1</td>
<td valign="top" align="center">98.1</td>
<td valign="top" align="center">95.8</td>
<td valign="top" align="center">98.1</td>
<td valign="top" align="center"><bold>98.2</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>t</italic><sub>5</sub></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">23.5</td>
<td valign="top" align="center">23.9</td>
<td valign="top" align="center">24.2</td>
<td valign="top" align="center">24.1</td>
<td valign="top" align="center"><bold>24.4</bold></td>
<td valign="top" align="center">23.9</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">24.0</td>
<td valign="top" align="center"><bold>24.1</bold></td>
<td valign="top" align="center">23.1</td>
<td valign="top" align="center">22.6</td>
<td valign="top" align="center">22.0</td>
<td valign="top" align="center">22.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">23.7</td>
<td valign="top" align="center"><bold>23.8</bold></td>
<td valign="top" align="center">22.4</td>
<td valign="top" align="center">23.1</td>
<td valign="top" align="center">22.5</td>
<td valign="top" align="center">23.3</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">22.1</td>
<td valign="top" align="center">22.4</td>
<td valign="top" align="center">22.2</td>
<td valign="top" align="center">22.4</td>
<td valign="top" align="center">22.4</td>
<td valign="top" align="center"><bold>23.5</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">65.8</td>
<td valign="top" align="center"><bold>66.1</bold></td>
<td valign="top" align="center">60.8</td>
<td valign="top" align="center">60.9</td>
<td valign="top" align="center">61.8</td>
<td valign="top" align="center">61.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>68.7</bold></td>
<td valign="top" align="center"><bold>68.7</bold></td>
<td valign="top" align="center">62.2</td>
<td valign="top" align="center">62.2</td>
<td valign="top" align="center">63.7</td>
<td valign="top" align="center">63.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">67.3</td>
<td valign="top" align="center"><bold>67.6</bold></td>
<td valign="top" align="center">61.1</td>
<td valign="top" align="center">61.4</td>
<td valign="top" align="center">62.7</td>
<td valign="top" align="center">62.2</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>63.9</bold></td>
<td valign="top" align="center">58.4</td>
<td valign="top" align="center">57.6</td>
<td valign="top" align="center">53.7</td>
<td valign="top" align="center">60.2</td>
<td valign="top" align="center">60.0</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">92.3</td>
<td valign="top" align="center"><bold>92.4</bold></td>
<td valign="top" align="center">87.6</td>
<td valign="top" align="center">87.6</td>
<td valign="top" align="center">88.0</td>
<td valign="top" align="center">88.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>91.7</bold></td>
<td valign="top" align="center"><bold>91.7</bold></td>
<td valign="top" align="center">87.0</td>
<td valign="top" align="center">87.4</td>
<td valign="top" align="center">87.9</td>
<td valign="top" align="center">87.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">93.2</td>
<td valign="top" align="center"><bold>93.4</bold></td>
<td valign="top" align="center">88.9</td>
<td valign="top" align="center">88.7</td>
<td valign="top" align="center">89.4</td>
<td valign="top" align="center">89.5</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>90.6</bold></td>
<td valign="top" align="center">86.8</td>
<td valign="top" align="center">85.7</td>
<td valign="top" align="center">83.4</td>
<td valign="top" align="center">87.4</td>
<td valign="top" align="center">87.4</td>
</tr>  <tr>
<td/>
<td valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Linear</bold></td>
</tr>
 <tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>Laplace</italic> (0, 1.25)</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">21.5</td>
<td valign="top" align="center"><bold>22.3</bold></td>
<td valign="top" align="center">19.6</td>
<td valign="top" align="center">20.4</td>
<td valign="top" align="center">19.3</td>
<td valign="top" align="center">19.0</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>21.9</bold></td>
<td valign="top" align="center">21.4</td>
<td valign="top" align="center">21.0</td>
<td valign="top" align="center">21.1</td>
<td valign="top" align="center">19.8</td>
<td valign="top" align="center">21.0</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">22.0</td>
<td valign="top" align="center"><bold>22.5</bold></td>
<td valign="top" align="center">20.6</td>
<td valign="top" align="center">20.3</td>
<td valign="top" align="center">20.1</td>
<td valign="top" align="center">20.3</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">17.3</td>
<td valign="top" align="center"><bold>19.8</bold></td>
<td valign="top" align="center">16.5</td>
<td valign="top" align="center">18.8</td>
<td valign="top" align="center">19.4</td>
<td valign="top" align="center">19.3</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">61.8</td>
<td valign="top" align="center"><bold>62.0</bold></td>
<td valign="top" align="center">53.2</td>
<td valign="top" align="center">53.2</td>
<td valign="top" align="center">55.0</td>
<td valign="top" align="center">54.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">62.2</td>
<td valign="top" align="center"><bold>63.2</bold></td>
<td valign="top" align="center">52.7</td>
<td valign="top" align="center">53.3</td>
<td valign="top" align="center">54.5</td>
<td valign="top" align="center">53.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">64.6</td>
<td valign="top" align="center"><bold>66.0</bold></td>
<td valign="top" align="center">52.9</td>
<td valign="top" align="center">52.8</td>
<td valign="top" align="center">54.7</td>
<td valign="top" align="center">54.2</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>58.9</bold></td>
<td valign="top" align="center">56.9</td>
<td valign="top" align="center">51.7</td>
<td valign="top" align="center">49.8</td>
<td valign="top" align="center">56.0</td>
<td valign="top" align="center">56.1</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>92.3</bold></td>
<td valign="top" align="center">92.2</td>
<td valign="top" align="center">83.2</td>
<td valign="top" align="center">82.9</td>
<td valign="top" align="center">83.5</td>
<td valign="top" align="center">83.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>88.1</bold></td>
<td valign="top" align="center">87.9</td>
<td valign="top" align="center">77.7</td>
<td valign="top" align="center">77.5</td>
<td valign="top" align="center">78.2</td>
<td valign="top" align="center">78.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center"><bold>91.0</bold></td>
<td valign="top" align="center">90.7</td>
<td valign="top" align="center">82.6</td>
<td valign="top" align="center">82.9</td>
<td valign="top" align="center">82.9</td>
<td valign="top" align="center">83.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>85.5</bold></td>
<td valign="top" align="center">80.4</td>
<td valign="top" align="center">77.7</td>
<td valign="top" align="center">74.2</td>
<td valign="top" align="center">79.3</td>
<td valign="top" align="center">79.3</td>
</tr> <tr>
<td valign="top" align="left" colspan="7">Exp</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">22.2</td>
<td valign="top" align="center"><bold>23.1</bold></td>
<td valign="top" align="center">18.1</td>
<td valign="top" align="center">18.2</td>
<td valign="top" align="center">18.7</td>
<td valign="top" align="center">18.3</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">19.7</td>
<td valign="top" align="center"><bold>19.8</bold></td>
<td valign="top" align="center">17.0</td>
<td valign="top" align="center">17.5</td>
<td valign="top" align="center">17.2</td>
<td valign="top" align="center">17.5</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">19.4</td>
<td valign="top" align="center"><bold>20.2</bold></td>
<td valign="top" align="center">16.9</td>
<td valign="top" align="center">17.6</td>
<td valign="top" align="center">16.9</td>
<td valign="top" align="center">17.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">15.7</td>
<td valign="top" align="center"><bold>16.0</bold></td>
<td valign="top" align="center">13.4</td>
<td valign="top" align="center">14.9</td>
<td valign="top" align="center">14.8</td>
<td valign="top" align="center">14.9</td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">61.9</td>
<td valign="top" align="center"><bold>62.0</bold></td>
<td valign="top" align="center">40.8</td>
<td valign="top" align="center">41.8</td>
<td valign="top" align="center">42.6</td>
<td valign="top" align="center">42.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">56.2</td>
<td valign="top" align="center"><bold>57.6</bold></td>
<td valign="top" align="center">38.0</td>
<td valign="top" align="center">38.2</td>
<td valign="top" align="center">39.9</td>
<td valign="top" align="center">39.3</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">63.0</td>
<td valign="top" align="center"><bold>63.5</bold></td>
<td valign="top" align="center">42.6</td>
<td valign="top" align="center">42.6</td>
<td valign="top" align="center">44.9</td>
<td valign="top" align="center">44.8</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>50.1</bold></td>
<td valign="top" align="center">45.0</td>
<td valign="top" align="center">36.9</td>
<td valign="top" align="center">36.7</td>
<td valign="top" align="center">40.6</td>
<td valign="top" align="center">40.4</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">86.7</td>
<td valign="top" align="center"><bold>87.2</bold></td>
<td valign="top" align="center">67.2</td>
<td valign="top" align="center">67.5</td>
<td valign="top" align="center">67.7</td>
<td valign="top" align="center">67.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">85.4</td>
<td valign="top" align="center"><bold>85.6</bold></td>
<td valign="top" align="center">65.5</td>
<td valign="top" align="center">66.0</td>
<td valign="top" align="center">66.6</td>
<td valign="top" align="center">66.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">88.2</td>
<td valign="top" align="center"><bold>88.4</bold></td>
<td valign="top" align="center">66.4</td>
<td valign="top" align="center">66.3</td>
<td valign="top" align="center">67.3</td>
<td valign="top" align="center">67.6</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>83.5</bold></td>
<td valign="top" align="center">76.1</td>
<td valign="top" align="center">65.2</td>
<td valign="top" align="center">61.1</td>
<td valign="top" align="center">66.9</td>
<td valign="top" align="center">67.7</td>
</tr>
</tbody> 
</table>
<table-wrap-foot>
<p><italic>The highest power per setting is indicated in bold</italic>.</p>
</table-wrap-foot>
</table-wrap>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Empirical power from the simulation study with the indicators having an overall reliability of 60% and a linear relationship with the latent variable.</p></caption>
<table frame="hsides" rules="groups">
 <thead><tr>
<th/>
<th valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Linear</bold></th>
</tr>
<tr>
<th/>
</tr>
</thead>
<tbody>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><inline-formula><mml:math id="M60"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">N</mml:mi></mml:mrow><mml:mtext>&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">26.4</td>
<td valign="top" align="center">27.7</td>
<td valign="top" align="center">28.8</td>
<td valign="top" align="center"><bold>29.4</bold></td>
<td valign="top" align="center">28.5</td>
<td valign="top" align="center">28.9</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">23.9</td>
<td valign="top" align="center">25.4</td>
<td valign="top" align="center">26.0</td>
<td valign="top" align="center"><bold>26.9</bold></td>
<td valign="top" align="center">25.4</td>
<td valign="top" align="center">24.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">26.4</td>
<td valign="top" align="center">27.4</td>
<td valign="top" align="center">28.5</td>
<td valign="top" align="center"><bold>29.9</bold></td>
<td valign="top" align="center">28.9</td>
<td valign="top" align="center">28.3</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">21.1</td>
<td valign="top" align="center">24.1</td>
<td valign="top" align="center">23.5</td>
<td valign="top" align="center"><bold>25.3</bold></td>
<td valign="top" align="center">24.0</td>
<td valign="top" align="center">23.7</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">74.6</td>
<td valign="top" align="center">74.2</td>
<td valign="top" align="center">76.4</td>
<td valign="top" align="center">76.4</td>
<td valign="top" align="center">77.9</td>
<td valign="top" align="center"><bold>78.1</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">73.8</td>
<td valign="top" align="center">74.2</td>
<td valign="top" align="center">75.7</td>
<td valign="top" align="center">75.8</td>
<td valign="top" align="center"><bold>77.0</bold></td>
<td valign="top" align="center">76.7</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">75.4</td>
<td valign="top" align="center">74.7</td>
<td valign="top" align="center">76.7</td>
<td valign="top" align="center"><bold>77.8</bold></td>
<td valign="top" align="center"><bold>77.8</bold></td>
<td valign="top" align="center">77.6</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">67.8</td>
<td valign="top" align="center">68.0</td>
<td valign="top" align="center">71.1</td>
<td valign="top" align="center">70.0</td>
<td valign="top" align="center"><bold>74.5</bold></td>
<td valign="top" align="center">74.2</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">96.7</td>
<td valign="top" align="center">97.0</td>
<td valign="top" align="center">97.5</td>
<td valign="top" align="center">97.6</td>
<td valign="top" align="center"><bold>97.7</bold></td>
<td valign="top" align="center"><bold>97.7</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">95.9</td>
<td valign="top" align="center">96.2</td>
<td valign="top" align="center"><bold>97.7</bold></td>
<td valign="top" align="center"><bold>97.7</bold></td>
<td valign="top" align="center"><bold>97.7</bold></td>
<td valign="top" align="center"><bold>97.7</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">95.8</td>
<td valign="top" align="center">95.7</td>
<td valign="top" align="center">96.3</td>
<td valign="top" align="center">96.3</td>
<td valign="top" align="center"><bold>97.6</bold></td>
<td valign="top" align="center"><bold>97.6</bold></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">94.4</td>
<td valign="top" align="center">92.3</td>
<td valign="top" align="center">95.1</td>
<td valign="top" align="center">94.0</td>
<td valign="top" align="center">95.6</td>
<td valign="top" align="center"><bold>95.6</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>t</italic><sub>5</sub></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">19.9</td>
<td valign="top" align="center">20.8</td>
<td valign="top" align="center">21.2</td>
<td valign="top" align="center"><bold>21.8</bold></td>
<td valign="top" align="center">20.2</td>
<td valign="top" align="center">19.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">18.3</td>
<td valign="top" align="center"><bold>19.6</bold></td>
<td valign="top" align="center">18.3</td>
<td valign="top" align="center">19.4</td>
<td valign="top" align="center">18.2</td>
<td valign="top" align="center">18.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">20.1</td>
<td valign="top" align="center"><bold>20.8</bold></td>
<td valign="top" align="center">20.1</td>
<td valign="top" align="center">20.7</td>
<td valign="top" align="center">20.3</td>
<td valign="top" align="center">20.7</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">19.0</td>
<td valign="top" align="center">19.9</td>
<td valign="top" align="center">19.8</td>
<td valign="top" align="center"><bold>21.4</bold></td>
<td valign="top" align="center">18.0</td>
<td valign="top" align="center">18.3</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">56.4</td>
<td valign="top" align="center"><bold>57.8</bold></td>
<td valign="top" align="center">54.2</td>
<td valign="top" align="center">54.8</td>
<td valign="top" align="center">55.4</td>
<td valign="top" align="center">55.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">61.6</td>
<td valign="top" align="center"><bold>62.2</bold></td>
<td valign="top" align="center">56.7</td>
<td valign="top" align="center">57.0</td>
<td valign="top" align="center">58.2</td>
<td valign="top" align="center">58.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center"><bold>60.6</bold></td>
<td valign="top" align="center"><bold>60.6</bold></td>
<td valign="top" align="center">57.4</td>
<td valign="top" align="center">58.1</td>
<td valign="top" align="center">59.0</td>
<td valign="top" align="center">58.2</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>52.3</bold></td>
<td valign="top" align="center">51.7</td>
<td valign="top" align="center">49.1</td>
<td valign="top" align="center">48.4</td>
<td valign="top" align="center">51.2</td>
<td valign="top" align="center">51.5</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>87.7</bold></td>
<td valign="top" align="center"><bold>87.7</bold></td>
<td valign="top" align="center">83.5</td>
<td valign="top" align="center">83.8</td>
<td valign="top" align="center">84.1</td>
<td valign="top" align="center">83.7</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">85.3</td>
<td valign="top" align="center"><bold>86.3</bold></td>
<td valign="top" align="center">82.9</td>
<td valign="top" align="center">83.2</td>
<td valign="top" align="center">83.2</td>
<td valign="top" align="center">83.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center"><bold>89.1</bold></td>
<td valign="top" align="center">88.9</td>
<td valign="top" align="center">84.9</td>
<td valign="top" align="center">84.8</td>
<td valign="top" align="center">84.8</td>
<td valign="top" align="center">85.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>83.7</bold></td>
<td valign="top" align="center">82.0</td>
<td valign="top" align="center">80.3</td>
<td valign="top" align="center">79.1</td>
<td valign="top" align="center">81.2</td>
<td valign="top" align="center">81.1</td>
</tr>  <tr>
<td/>
<td valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Linear</bold></td>
</tr>
 <tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>Laplace</italic> (0, 1.25)</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">18.3</td>
<td valign="top" align="center"><bold>19.0</bold></td>
<td valign="top" align="center">17.7</td>
<td valign="top" align="center">18.2</td>
<td valign="top" align="center">17.7</td>
<td valign="top" align="center">17.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">18.3</td>
<td valign="top" align="center">18.5</td>
<td valign="top" align="center">18.5</td>
<td valign="top" align="center"><bold>19.0</bold></td>
<td valign="top" align="center">17.5</td>
<td valign="top" align="center">17.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">16.8</td>
<td valign="top" align="center">18.7</td>
<td valign="top" align="center">17.4</td>
<td valign="top" align="center"><bold>18.9</bold></td>
<td valign="top" align="center">17.9</td>
<td valign="top" align="center">17.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">14.9</td>
<td valign="top" align="center"><bold>17.7</bold></td>
<td valign="top" align="center">14.5</td>
<td valign="top" align="center">16.3</td>
<td valign="top" align="center">17.4</td>
<td valign="top" align="center">16.8</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">51.6</td>
<td valign="top" align="center"><bold>51.8</bold></td>
<td valign="top" align="center">47.4</td>
<td valign="top" align="center">48.0</td>
<td valign="top" align="center">48.4</td>
<td valign="top" align="center">48.7</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">54.0</td>
<td valign="top" align="center"><bold>54.1</bold></td>
<td valign="top" align="center">48.7</td>
<td valign="top" align="center">49.0</td>
<td valign="top" align="center">49.9</td>
<td valign="top" align="center">50.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">55.9</td>
<td valign="top" align="center"><bold>56.0</bold></td>
<td valign="top" align="center">47.2</td>
<td valign="top" align="center">47.2</td>
<td valign="top" align="center">48.9</td>
<td valign="top" align="center">49.2</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">49.0</td>
<td valign="top" align="center">48.0</td>
<td valign="top" align="center">46.7</td>
<td valign="top" align="center">45.5</td>
<td valign="top" align="center"><bold>49.7</bold></td>
<td valign="top" align="center"><bold>49.7</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>85.5</bold></td>
<td valign="top" align="center"><bold>85.5</bold></td>
<td valign="top" align="center">78.1</td>
<td valign="top" align="center">78.1</td>
<td valign="top" align="center">79.0</td>
<td valign="top" align="center">79.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">80.1</td>
<td valign="top" align="center"><bold>80.8</bold></td>
<td valign="top" align="center">72.7</td>
<td valign="top" align="center">72.7</td>
<td valign="top" align="center">73.0</td>
<td valign="top" align="center">73.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">85.5</td>
<td valign="top" align="center"><bold>85.9</bold></td>
<td valign="top" align="center">78.7</td>
<td valign="top" align="center">79.3</td>
<td valign="top" align="center">78.8</td>
<td valign="top" align="center">79.1</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">85.5</td>
<td valign="top" align="center"><bold>85.9</bold></td>
<td valign="top" align="center">78.7</td>
<td valign="top" align="center">79.3</td>
<td valign="top" align="center">78.8</td>
<td valign="top" align="center">79.1</td>
</tr> <tr>
<td valign="top" align="left" colspan="7">Exp</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">16.1</td>
<td valign="top" align="center"><bold>16.9</bold></td>
<td valign="top" align="center">14.6</td>
<td valign="top" align="center">16.4</td>
<td valign="top" align="center">15.7</td>
<td valign="top" align="center">14.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">15.5</td>
<td valign="top" align="center"><bold>15.6</bold></td>
<td valign="top" align="center">13.9</td>
<td valign="top" align="center">15.4</td>
<td valign="top" align="center">14.5</td>
<td valign="top" align="center">14.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">14.9</td>
<td valign="top" align="center"><bold>16.3</bold></td>
<td valign="top" align="center">14.3</td>
<td valign="top" align="center">14.5</td>
<td valign="top" align="center">14.3</td>
<td valign="top" align="center">13.8</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">12.9</td>
<td valign="top" align="center"><bold>14.1</bold></td>
<td valign="top" align="center">12.0</td>
<td valign="top" align="center">13.9</td>
<td valign="top" align="center">12.1</td>
<td valign="top" align="center">12.2</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>48.8</bold></td>
<td valign="top" align="center">48.2</td>
<td valign="top" align="center">36.5</td>
<td valign="top" align="center">37.3</td>
<td valign="top" align="center">37.7</td>
<td valign="top" align="center">37.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">43.6</td>
<td valign="top" align="center"><bold>44.7</bold></td>
<td valign="top" align="center">33.3</td>
<td valign="top" align="center">33.4</td>
<td valign="top" align="center">34.7</td>
<td valign="top" align="center">34.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">48.8</td>
<td valign="top" align="center"><bold>50.1</bold></td>
<td valign="top" align="center">38.1</td>
<td valign="top" align="center">38.6</td>
<td valign="top" align="center">40.0</td>
<td valign="top" align="center">39.8</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>37.6</bold></td>
<td valign="top" align="center">37.3</td>
<td valign="top" align="center">31.3</td>
<td valign="top" align="center">32.5</td>
<td valign="top" align="center">33.2</td>
<td valign="top" align="center">33.3</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>75.2</bold></td>
<td valign="top" align="center">75.0</td>
<td valign="top" align="center">61.6</td>
<td valign="top" align="center">62.0</td>
<td valign="top" align="center">62.2</td>
<td valign="top" align="center">62.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">74.3</td>
<td valign="top" align="center"><bold>75.2</bold></td>
<td valign="top" align="center">60.2</td>
<td valign="top" align="center">61.2</td>
<td valign="top" align="center">61.1</td>
<td valign="top" align="center">61.0</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">77.4</td>
<td valign="top" align="center"><bold>78.8</bold></td>
<td valign="top" align="center">60.8</td>
<td valign="top" align="center">61.5</td>
<td valign="top" align="center">61.8</td>
<td valign="top" align="center">62.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>71.0</bold></td>
<td valign="top" align="center">66.7</td>
<td valign="top" align="center">58.2</td>
<td valign="top" align="center">54.2</td>
<td valign="top" align="center">59.6</td>
<td valign="top" align="center">59.5</td>
</tr>
</tbody> 
</table>
<table-wrap-foot>
<p><italic>The highest power per setting is indicated in bold</italic>.</p>
</table-wrap-foot>
</table-wrap>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p>Empirical power from the simulation study with the indicators having an overall reliability of 80% and a non-linear relationship with the latent variable.</p></caption>
<table frame="hsides" rules="groups">
 <thead><tr>
<th/>
<th valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Non-linear</bold></th>
</tr>
<tr>
<th/>
</tr>
</thead>
<tbody>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><inline-formula><mml:math id="M61"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">N</mml:mi></mml:mrow><mml:mtext>&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">29.7</td>
<td valign="top" align="center">30.3</td>
<td valign="top" align="center"><bold>32.2</bold></td>
<td valign="top" align="center">31.4</td>
<td valign="top" align="center">31.8</td>
<td valign="top" align="center">31.0</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">28.1</td>
<td valign="top" align="center">29.0</td>
<td valign="top" align="center">29.3</td>
<td valign="top" align="center"><bold>30.3</bold></td>
<td valign="top" align="center">28.7</td>
<td valign="top" align="center">28.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">28.1</td>
<td valign="top" align="center">28.4</td>
<td valign="top" align="center">30.6</td>
<td valign="top" align="center"><bold>31.4</bold></td>
<td valign="top" align="center">30.8</td>
<td valign="top" align="center">29.8</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">23.7</td>
<td valign="top" align="center">25.5</td>
<td valign="top" align="center">25.7</td>
<td valign="top" align="center">26.2</td>
<td valign="top" align="center"><bold>29.1</bold></td>
<td valign="top" align="center">27.0</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">78.0</td>
<td valign="top" align="center">78.8</td>
<td valign="top" align="center">79.0</td>
<td valign="top" align="center">79.8</td>
<td valign="top" align="center"><bold>80.7</bold></td>
<td valign="top" align="center">80.6</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">77.7</td>
<td valign="top" align="center">77.7</td>
<td valign="top" align="center">79.4</td>
<td valign="top" align="center">79.4</td>
<td valign="top" align="center">80.6</td>
<td valign="top" align="center"><bold>81.2</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">78.9</td>
<td valign="top" align="center">78.5</td>
<td valign="top" align="center">79.8</td>
<td valign="top" align="center">80.6</td>
<td valign="top" align="center">81.2</td>
<td valign="top" align="center"><bold>81.3</bold></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">73.5</td>
<td valign="top" align="center">70.9</td>
<td valign="top" align="center">76.1</td>
<td valign="top" align="center">73.6</td>
<td valign="top" align="center">78.8</td>
<td valign="top" align="center"><bold>79.1</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">97.6</td>
<td valign="top" align="center">97.7</td>
<td valign="top" align="center">98.2</td>
<td valign="top" align="center">98.2</td>
<td valign="top" align="center"><bold>98.3</bold></td>
<td valign="top" align="center">98.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">97.6</td>
<td valign="top" align="center">97.5</td>
<td valign="top" align="center">97.9</td>
<td valign="top" align="center">97.9</td>
<td valign="top" align="center"><bold>98.0</bold></td>
<td valign="top" align="center"><bold>98.0</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">97.3</td>
<td valign="top" align="center">96.9</td>
<td valign="top" align="center"><bold>98.0</bold></td>
<td valign="top" align="center">97.8</td>
<td valign="top" align="center"><bold>98.0</bold></td>
<td valign="top" align="center"><bold>98.0</bold></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">96.3</td>
<td valign="top" align="center">94.8</td>
<td valign="top" align="center">97.1</td>
<td valign="top" align="center">95.4</td>
<td valign="top" align="center"><bold>97.6</bold></td>
<td valign="top" align="center"><bold>97.6</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>t</italic><sub>5</sub></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">23.3</td>
<td valign="top" align="center">23.1</td>
<td valign="top" align="center">23.6</td>
<td valign="top" align="center">24.0</td>
<td valign="top" align="center"><bold>24.1</bold></td>
<td valign="top" align="center">23.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">23.0</td>
<td valign="top" align="center">23.4</td>
<td valign="top" align="center">23.4</td>
<td valign="top" align="center"><bold>24.1</bold></td>
<td valign="top" align="center">22.8</td>
<td valign="top" align="center">23.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">23.5</td>
<td valign="top" align="center">23.9</td>
<td valign="top" align="center">23.8</td>
<td valign="top" align="center"><bold>24.2</bold></td>
<td valign="top" align="center">24.0</td>
<td valign="top" align="center">23.9</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">22.2</td>
<td valign="top" align="center">22.7</td>
<td valign="top" align="center">21.8</td>
<td valign="top" align="center"><bold>23.3</bold></td>
<td valign="top" align="center">22.4</td>
<td valign="top" align="center">22.7</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">65.4</td>
<td valign="top" align="center"><bold>65.8</bold></td>
<td valign="top" align="center">60.8</td>
<td valign="top" align="center">61.1</td>
<td valign="top" align="center">62.2</td>
<td valign="top" align="center">62.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">68.5</td>
<td valign="top" align="center"><bold>69.2</bold></td>
<td valign="top" align="center">63.6</td>
<td valign="top" align="center">63.3</td>
<td valign="top" align="center">64.8</td>
<td valign="top" align="center">64.9</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">66.5</td>
<td valign="top" align="center"><bold>67.1</bold></td>
<td valign="top" align="center">62.9</td>
<td valign="top" align="center">62.9</td>
<td valign="top" align="center">64.5</td>
<td valign="top" align="center">65.1</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>62.9</bold></td>
<td valign="top" align="center">57.5</td>
<td valign="top" align="center">58.5</td>
<td valign="top" align="center">55.2</td>
<td valign="top" align="center">61.4</td>
<td valign="top" align="center">61.1</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>92.2</bold></td>
<td valign="top" align="center">92.1</td>
<td valign="top" align="center">88.9</td>
<td valign="top" align="center">88.7</td>
<td valign="top" align="center">89.3</td>
<td valign="top" align="center">89.5</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>91.6</bold></td>
<td valign="top" align="center">91.3</td>
<td valign="top" align="center">88.8</td>
<td valign="top" align="center">88.7</td>
<td valign="top" align="center">89.3</td>
<td valign="top" align="center">89.5</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">93.0</td>
<td valign="top" align="center"><bold>93.2</bold></td>
<td valign="top" align="center">90.6</td>
<td valign="top" align="center">90.9</td>
<td valign="top" align="center">90.8</td>
<td valign="top" align="center">90.9</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>90.4</bold></td>
<td valign="top" align="center">86.7</td>
<td valign="top" align="center">87.5</td>
<td valign="top" align="center">84.6</td>
<td valign="top" align="center">88.9</td>
<td valign="top" align="center">88.8</td>
</tr>  <tr>
<td/>
<td valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Non-linear</bold></td>
</tr>
 <tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>Laplace</italic> (0, 1.25)</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">21.1</td>
<td valign="top" align="center"><bold>22.3</bold></td>
<td valign="top" align="center">19.5</td>
<td valign="top" align="center">20.2</td>
<td valign="top" align="center">20.2</td>
<td valign="top" align="center">19.7</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>22.9</bold></td>
<td valign="top" align="center">22.4</td>
<td valign="top" align="center">22.2</td>
<td valign="top" align="center">21.8</td>
<td valign="top" align="center">20.9</td>
<td valign="top" align="center">21.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">22.5</td>
<td valign="top" align="center"><bold>22.7</bold></td>
<td valign="top" align="center">21.2</td>
<td valign="top" align="center">20.5</td>
<td valign="top" align="center">20.4</td>
<td valign="top" align="center">21.4</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">19.3</td>
<td valign="top" align="center">20.2</td>
<td valign="top" align="center">18.0</td>
<td valign="top" align="center">19.9</td>
<td valign="top" align="center"><bold>21.0</bold></td>
<td valign="top" align="center">20.2</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">61.6</td>
<td valign="top" align="center"><bold>62.9</bold></td>
<td valign="top" align="center">53.9</td>
<td valign="top" align="center">55.1</td>
<td valign="top" align="center">55.9</td>
<td valign="top" align="center">55.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">62.1</td>
<td valign="top" align="center"><bold>63.5</bold></td>
<td valign="top" align="center">55.4</td>
<td valign="top" align="center">55.8</td>
<td valign="top" align="center">56.3</td>
<td valign="top" align="center">56.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">64.4</td>
<td valign="top" align="center"><bold>65.4</bold></td>
<td valign="top" align="center">55.6</td>
<td valign="top" align="center">55.4</td>
<td valign="top" align="center">57.4</td>
<td valign="top" align="center">57.9</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>59.8</bold></td>
<td valign="top" align="center">56.2</td>
<td valign="top" align="center">54.4</td>
<td valign="top" align="center">50.1</td>
<td valign="top" align="center">57.0</td>
<td valign="top" align="center">56.9</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">92.0</td>
<td valign="top" align="center"><bold>92.6</bold></td>
<td valign="top" align="center">85.6</td>
<td valign="top" align="center">85.5</td>
<td valign="top" align="center">86.1</td>
<td valign="top" align="center">85.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">88.3</td>
<td valign="top" align="center"><bold>89.0</bold></td>
<td valign="top" align="center">80.6</td>
<td valign="top" align="center">80.2</td>
<td valign="top" align="center">81.2</td>
<td valign="top" align="center">81.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">91.1</td>
<td valign="top" align="center"><bold>91.7</bold></td>
<td valign="top" align="center">85.2</td>
<td valign="top" align="center">85.7</td>
<td valign="top" align="center">86.1</td>
<td valign="top" align="center">86.1</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>86.1</bold></td>
<td valign="top" align="center">81.9</td>
<td valign="top" align="center">80.7</td>
<td valign="top" align="center">75.9</td>
<td valign="top" align="center">81.9</td>
<td valign="top" align="center">81.8</td>
</tr> <tr>
<td valign="top" align="left" colspan="7">Exp</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">25.0</td>
<td valign="top" align="center"><bold>25.9</bold></td>
<td valign="top" align="center">23.6</td>
<td valign="top" align="center">23.3</td>
<td valign="top" align="center">23.7</td>
<td valign="top" align="center">24.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>23.1</bold></td>
<td valign="top" align="center">22.5</td>
<td valign="top" align="center">21.2</td>
<td valign="top" align="center">21.5</td>
<td valign="top" align="center">22.1</td>
<td valign="top" align="center">23.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center"><bold>22.4</bold></td>
<td valign="top" align="center">22.2</td>
<td valign="top" align="center">20.9</td>
<td valign="top" align="center">21.2</td>
<td valign="top" align="center">21.4</td>
<td valign="top" align="center">21.6</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>19.4</bold></td>
<td valign="top" align="center">18.8</td>
<td valign="top" align="center">18.7</td>
<td valign="top" align="center">18.7</td>
<td valign="top" align="center">18.2</td>
<td valign="top" align="center">18.5</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>67.9</bold></td>
<td valign="top" align="center"><bold>67.9</bold></td>
<td valign="top" align="center">58.3</td>
<td valign="top" align="center">57.1</td>
<td valign="top" align="center">60.8</td>
<td valign="top" align="center">59.7</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">63.7</td>
<td valign="top" align="center"><bold>64.6</bold></td>
<td valign="top" align="center">54.5</td>
<td valign="top" align="center">54.0</td>
<td valign="top" align="center">56.8</td>
<td valign="top" align="center">56.9</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">69.6</td>
<td valign="top" align="center"><bold>70.4</bold></td>
<td valign="top" align="center">60.6</td>
<td valign="top" align="center">59.1</td>
<td valign="top" align="center">62.3</td>
<td valign="top" align="center">61.7</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>58.7</bold></td>
<td valign="top" align="center">55.9</td>
<td valign="top" align="center">49.7</td>
<td valign="top" align="center">49.0</td>
<td valign="top" align="center">54.8</td>
<td valign="top" align="center">54.6</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center"><bold>91.8</bold></td>
<td valign="top" align="center">91.7</td>
<td valign="top" align="center">84.7</td>
<td valign="top" align="center">84.1</td>
<td valign="top" align="center">85.4</td>
<td valign="top" align="center">85.7</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center"><bold>89.5</bold></td>
<td valign="top" align="center"><bold>89.5</bold></td>
<td valign="top" align="center">82.6</td>
<td valign="top" align="center">82.5</td>
<td valign="top" align="center">83.2</td>
<td valign="top" align="center">84.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center"><bold>92.8</bold></td>
<td valign="top" align="center">92.7</td>
<td valign="top" align="center">86.8</td>
<td valign="top" align="center">85.5</td>
<td valign="top" align="center">87.1</td>
<td valign="top" align="center">87.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>90.5</bold></td>
<td valign="top" align="center">85.3</td>
<td valign="top" align="center">81.6</td>
<td valign="top" align="center">79.4</td>
<td valign="top" align="center">84.6</td>
<td valign="top" align="center">84.5</td>
</tr>
</tbody> 
</table>
<table-wrap-foot>
<p><italic>The highest power per setting is indicated in bold</italic>.</p>
</table-wrap-foot>
</table-wrap>
<table-wrap position="float" id="T4">
<label>Table 4</label>
<caption><p>Empirical power from the simulation study with the indicators having an overall reliability of 60% and a non-linear relationship with the latent variable.</p></caption>
<table frame="hsides" rules="groups">
 <thead><tr>
<th/>
<th valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Non-linear</bold></th>
</tr>
<tr>
<th/>
</tr>
</thead>
<tbody>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><inline-formula><mml:math id="M62"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">N</mml:mi></mml:mrow><mml:mtext>&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">24.6</td>
<td valign="top" align="center">25.6</td>
<td valign="top" align="center">27.7</td>
<td valign="top" align="center"><bold>28.3</bold></td>
<td valign="top" align="center">27.5</td>
<td valign="top" align="center">28.0</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">22.6</td>
<td valign="top" align="center">24.6</td>
<td valign="top" align="center">25.1</td>
<td valign="top" align="center"><bold>26.7</bold></td>
<td valign="top" align="center">24.6</td>
<td valign="top" align="center">24.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">24.0</td>
<td valign="top" align="center">26.3</td>
<td valign="top" align="center">26.8</td>
<td valign="top" align="center"><bold>28.3</bold></td>
<td valign="top" align="center">26.6</td>
<td valign="top" align="center">25.9</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">21.2</td>
<td valign="top" align="center">21.6</td>
<td valign="top" align="center">23.1</td>
<td valign="top" align="center"><bold>23.2</bold></td>
<td valign="top" align="center"><bold>23.3</bold></td>
<td valign="top" align="center"><bold>23.2</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">71.6</td>
<td valign="top" align="center">72.4</td>
<td valign="top" align="center">73.2</td>
<td valign="top" align="center">74.0</td>
<td valign="top" align="center">75.2</td>
<td valign="top" align="center"><bold>75.3</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">71.6</td>
<td valign="top" align="center">71.8</td>
<td valign="top" align="center">74.3</td>
<td valign="top" align="center">74.0</td>
<td valign="top" align="center">75.4</td>
<td valign="top" align="center"><bold>75.5</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">73.4</td>
<td valign="top" align="center">73.3</td>
<td valign="top" align="center">74.7</td>
<td valign="top" align="center">75.4</td>
<td valign="top" align="center"><bold>76.2</bold></td>
<td valign="top" align="center">76.1</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">65.8</td>
<td valign="top" align="center">64.9</td>
<td valign="top" align="center">68.9</td>
<td valign="top" align="center">66.8</td>
<td valign="top" align="center">72.2</td>
<td valign="top" align="center"><bold>72.3</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">95.4</td>
<td valign="top" align="center">95.8</td>
<td valign="top" align="center">96.5</td>
<td valign="top" align="center"><bold>96.6</bold></td>
<td valign="top" align="center"><bold>96.9</bold></td>
<td valign="top" align="center">96.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">94.9</td>
<td valign="top" align="center">95.0</td>
<td valign="top" align="center">96.1</td>
<td valign="top" align="center">96.0</td>
<td valign="top" align="center">96.1</td>
<td valign="top" align="center"><bold>96.2</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">94.9</td>
<td valign="top" align="center">94.6</td>
<td valign="top" align="center">95.5</td>
<td valign="top" align="center">95.6</td>
<td valign="top" align="center">95.7</td>
<td valign="top" align="center"><bold>95.8</bold></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">92.6</td>
<td valign="top" align="center">91.5</td>
<td valign="top" align="center">94.0</td>
<td valign="top" align="center">92.2</td>
<td valign="top" align="center">94.6</td>
<td valign="top" align="center"><bold>94.7</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>t</italic><sub>5</sub></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">19.2</td>
<td valign="top" align="center">19.2</td>
<td valign="top" align="center">20.4</td>
<td valign="top" align="center"><bold>21.2</bold></td>
<td valign="top" align="center">19.0</td>
<td valign="top" align="center">18.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">18.6</td>
<td valign="top" align="center">19.1</td>
<td valign="top" align="center">18.2</td>
<td valign="top" align="center">19.0</td>
<td valign="top" align="center">18.6</td>
<td valign="top" align="center"><bold>19.3</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">20.3</td>
<td valign="top" align="center">20.5</td>
<td valign="top" align="center"><bold>21.1</bold></td>
<td valign="top" align="center"><bold>21.1</bold></td>
<td valign="top" align="center">20.3</td>
<td valign="top" align="center">20.8</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">19.5</td>
<td valign="top" align="center">20.2</td>
<td valign="top" align="center">19.7</td>
<td valign="top" align="center"><bold>20.6</bold></td>
<td valign="top" align="center">17.8</td>
<td valign="top" align="center">17.7</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">56.6</td>
<td valign="top" align="center">56.6</td>
<td valign="top" align="center">55.9</td>
<td valign="top" align="center">56.1</td>
<td valign="top" align="center">56.6</td>
<td valign="top" align="center"><bold>56.9</bold></td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">60.0</td>
<td valign="top" align="center"><bold>61.3</bold></td>
<td valign="top" align="center">57.1</td>
<td valign="top" align="center">57.3</td>
<td valign="top" align="center">57.7</td>
<td valign="top" align="center">57.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center"><bold>60.6</bold></td>
<td valign="top" align="center">60.2</td>
<td valign="top" align="center">57.8</td>
<td valign="top" align="center">58.3</td>
<td valign="top" align="center">58.9</td>
<td valign="top" align="center">59.2</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">51.9</td>
<td valign="top" align="center">50.8</td>
<td valign="top" align="center">51.0</td>
<td valign="top" align="center">49.2</td>
<td valign="top" align="center">52.0</td>
<td valign="top" align="center"><bold>52.9</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">87.4</td>
<td valign="top" align="center"><bold>87.7</bold></td>
<td valign="top" align="center">84.2</td>
<td valign="top" align="center">84.8</td>
<td valign="top" align="center">84.6</td>
<td valign="top" align="center">84.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">85.2</td>
<td valign="top" align="center"><bold>85.9</bold></td>
<td valign="top" align="center">83.6</td>
<td valign="top" align="center">84.4</td>
<td valign="top" align="center">84.0</td>
<td valign="top" align="center">84.4</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">88.4</td>
<td valign="top" align="center"><bold>88.9</bold></td>
<td valign="top" align="center">86.3</td>
<td valign="top" align="center">86.8</td>
<td valign="top" align="center">86.8</td>
<td valign="top" align="center">87.2</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>84.2</bold></td>
<td valign="top" align="center">81.9</td>
<td valign="top" align="center">81.4</td>
<td valign="top" align="center">80.0</td>
<td valign="top" align="center">81.9</td>
<td valign="top" align="center">82.1</td>
</tr>  <tr>
<td/>
<td valign="top" align="center" colspan="6" style="border-bottom: thin solid #000000;"><bold>Non-linear</bold></td>
</tr>
 <tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>WMW&#x02013;max rel</bold></td>
<td valign="top" align="center"><bold>WMW&#x02013;mean</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold></td>
<td valign="top" align="center"><italic><bold>t</bold></italic><bold>-test&#x02013;mean</bold></td>
<td valign="top" align="center"><bold>SEM</bold></td>
<td valign="top" align="center"><bold>SEM&#x02013;corrected</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="7"><italic>Laplace</italic> (0, 1.25)</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">16.8</td>
<td valign="top" align="center"><bold>19.0</bold></td>
<td valign="top" align="center">16.5</td>
<td valign="top" align="center">17.7</td>
<td valign="top" align="center">16.9</td>
<td valign="top" align="center">16.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">18.2</td>
<td valign="top" align="center"><bold>19.2</bold></td>
<td valign="top" align="center">18.5</td>
<td valign="top" align="center">19.0</td>
<td valign="top" align="center">17.6</td>
<td valign="top" align="center">17.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">17.6</td>
<td valign="top" align="center">18.5</td>
<td valign="top" align="center">18.6</td>
<td valign="top" align="center"><bold>19.3</bold></td>
<td valign="top" align="center">18.3</td>
<td valign="top" align="center">17.4</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">15.6</td>
<td valign="top" align="center"><bold>17.4</bold></td>
<td valign="top" align="center">15.8</td>
<td valign="top" align="center">17.3</td>
<td valign="top" align="center">17.2</td>
<td valign="top" align="center"><bold>17.4</bold></td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">50.9</td>
<td valign="top" align="center"><bold>52.2</bold></td>
<td valign="top" align="center">48.4</td>
<td valign="top" align="center">49.3</td>
<td valign="top" align="center">49.4</td>
<td valign="top" align="center">49.3</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">52.8</td>
<td valign="top" align="center"><bold>54.9</bold></td>
<td valign="top" align="center">49.8</td>
<td valign="top" align="center">51.1</td>
<td valign="top" align="center">51.2</td>
<td valign="top" align="center">51.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">56.0</td>
<td valign="top" align="center"><bold>56.5</bold></td>
<td valign="top" align="center">49.0</td>
<td valign="top" align="center">49.3</td>
<td valign="top" align="center">51.2</td>
<td valign="top" align="center">51.8</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">49.5</td>
<td valign="top" align="center">49.6</td>
<td valign="top" align="center">46.8</td>
<td valign="top" align="center">46.5</td>
<td valign="top" align="center"><bold>50.0</bold></td>
<td valign="top" align="center">49.5</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">85.1</td>
<td valign="top" align="center"><bold>85.8</bold></td>
<td valign="top" align="center">80.6</td>
<td valign="top" align="center">80.2</td>
<td valign="top" align="center">81.0</td>
<td valign="top" align="center">81.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">81.0</td>
<td valign="top" align="center"><bold>81.6</bold></td>
<td valign="top" align="center">74.1</td>
<td valign="top" align="center">74.9</td>
<td valign="top" align="center">74.8</td>
<td valign="top" align="center">75.0</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">85.8</td>
<td valign="top" align="center"><bold>86.1</bold></td>
<td valign="top" align="center">81.3</td>
<td valign="top" align="center">81.2</td>
<td valign="top" align="center">81.4</td>
<td valign="top" align="center">81.4</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>78.1</bold></td>
<td valign="top" align="center">74.7</td>
<td valign="top" align="center">72.7</td>
<td valign="top" align="center">70.5</td>
<td valign="top" align="center">74.2</td>
<td valign="top" align="center">74.2</td>
</tr> <tr>
<td valign="top" align="left" colspan="7">Exp</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">19.1</td>
<td valign="top" align="center"><bold>21.4</bold></td>
<td valign="top" align="center">19.0</td>
<td valign="top" align="center">20.1</td>
<td valign="top" align="center">18.7</td>
<td valign="top" align="center">19.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">18.7</td>
<td valign="top" align="center">19.1</td>
<td valign="top" align="center">18.1</td>
<td valign="top" align="center"><bold>19.4</bold></td>
<td valign="top" align="center">18.3</td>
<td valign="top" align="center">19.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">17.9</td>
<td valign="top" align="center"><bold>19.0</bold></td>
<td valign="top" align="center">17.5</td>
<td valign="top" align="center">18.6</td>
<td valign="top" align="center">17.2</td>
<td valign="top" align="center">17.9</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">14.8</td>
<td valign="top" align="center"><bold>16.8</bold></td>
<td valign="top" align="center">15.9</td>
<td valign="top" align="center">16.5</td>
<td valign="top" align="center">13.7</td>
<td valign="top" align="center">14.1</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 50</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">58.0</td>
<td valign="top" align="center"><bold>58.5</bold></td>
<td valign="top" align="center">51.2</td>
<td valign="top" align="center">52.3</td>
<td valign="top" align="center">53.3</td>
<td valign="top" align="center">53.1</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">52.9</td>
<td valign="top" align="center"><bold>54.5</bold></td>
<td valign="top" align="center">46.9</td>
<td valign="top" align="center">49.3</td>
<td valign="top" align="center">48.6</td>
<td valign="top" align="center">49.2</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">58.3</td>
<td valign="top" align="center"><bold>59.0</bold></td>
<td valign="top" align="center">51.7</td>
<td valign="top" align="center">53.2</td>
<td valign="top" align="center">54.1</td>
<td valign="top" align="center">53.1</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center">46.2</td>
<td valign="top" align="center"><bold>47.7</bold></td>
<td valign="top" align="center">43.0</td>
<td valign="top" align="center">44.0</td>
<td valign="top" align="center">44.5</td>
<td valign="top" align="center">44.4</td>
</tr> <tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left" colspan="7"><bold><italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 100</bold></td>
</tr> <tr>
<td valign="top" align="left">Setting 1</td>
<td valign="top" align="center">83.8</td>
<td valign="top" align="center"><bold>84.6</bold></td>
<td valign="top" align="center">78.0</td>
<td valign="top" align="center">78.4</td>
<td valign="top" align="center">78.7</td>
<td valign="top" align="center">78.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 2</td>
<td valign="top" align="center">83.0</td>
<td valign="top" align="center"><bold>84.1</bold></td>
<td valign="top" align="center">77.0</td>
<td valign="top" align="center">77.7</td>
<td valign="top" align="center">77.5</td>
<td valign="top" align="center">77.8</td>
</tr>
<tr>
<td valign="top" align="left">Setting 3</td>
<td valign="top" align="center">85.5</td>
<td valign="top" align="center"><bold>86.5</bold></td>
<td valign="top" align="center">79.6</td>
<td valign="top" align="center">80.3</td>
<td valign="top" align="center">80.3</td>
<td valign="top" align="center">80.0</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">Setting 4</td>
<td valign="top" align="center"><bold>80.0</bold></td>
<td valign="top" align="center">78.4</td>
<td valign="top" align="center">74.8</td>
<td valign="top" align="center">74.6</td>
<td valign="top" align="center">76.6</td>
<td valign="top" align="center">75.8</td>
</tr>
</tbody> 
</table>
<table-wrap-foot>
<p><italic>The highest power per setting is indicated in bold</italic>.</p>
</table-wrap-foot>
</table-wrap>
<p>Regardless of the setting, sample size, reliability or type of function <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;), the following trend can be observed with respect to the empirical power. When the latent variable is normally distributed, the six methods have similar power although the methods based on SEM and the <italic>t</italic>-test are slightly more powerful. As the distribution becomes more heavily tailed, the more superior the WMW methods become. This pattern is in accordance with the observations in the context of observed outcome variables, as mentioned by Van der Vaart (<xref ref-type="bibr" rid="B32">2000</xref>) and Hollander et al. (<xref ref-type="bibr" rid="B14">2013</xref>).</p>
<p>When one of the three indicators has an extremely low reliability, i.e. setting 4, a slightly different pattern can be seen. <bold>SEM</bold> has now a more pronounced superior power for the normal distribution. For the more heavily tailed distributions, the superiority of the WMW methods overall re-emerges. The methods <bold>WMW&#x02013;max rel</bold> and <bold>WMW&#x02013;mean</bold> differ in terms of the aggregated indicator that is used as input for the WMW method. The added value of the optimization process is demonstrated when looking at setting 4. The ability to fine-tune the weights in line with the reliability of each indicator separately and hence giving less influence to an indicator that has a bad quality, results in a higher empirical power for <bold>WMW&#x02013;max rel</bold> in comparison with <bold>WMW&#x02013;mean</bold>. However, it should be noted that when the reliability of indicators is equal to one another and/or the sample size is small (i.e. <italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15), a small reduction in empirical power for <bold>WMW&#x02013;max rel</bold> in comparison with <bold>WMW&#x02013;mean</bold> is observed. Hence, estimating the weights in order to optimize the estimated reliability of the aggregated indicator only has an added value when the data requires such adaptation.</p>
<p>Comparing the methods <bold>SEM</bold> and <bold>SEM&#x02013;correction</bold>, the results with respect to the empirical power show no remarkable differences. Hence, the results suggest that the added value of robust Satorra-Bentler standard errors is limited in our simulation setup.</p>
<p>The lower the reliability of all three indicators, the lower the power, and this is true for all settings and all methods. To guard the readability of the tables, the point estimates and standard errors are not listed, but based on these results, the reason of the decrease in power is perhaps different between the WMW methods and SEM methods. For <bold>SEM</bold> and <bold>SEM&#x02013;correction</bold>, it seems that there is barely an effect on the precision of the estimation but an increased empirical standard error is observed, hence influencing the power. Contrary, for <bold>WMW&#x02013;max rel</bold> and <bold>WMW&#x02013;mean</bold>, the empirical standard error remains relatively stable over the different levels of reliability, but the estimation is less accurate and hence influencing the power.</p>
<p>Overall, the simulation results suggest that the attractive properties of the original WMW method as discussed in the introduction are preserved in the context of latent variables. The extended WMW method is thus robust against outliers, relevant for skewed data and has a superior power in skewed distributions.</p>
</sec>
</sec>
<sec id="s4">
<title>4. Illustration</title>
<p>We now reconsider the example of the introduction, where we want to examine the association between employment and depression in 455 patients with multiple sclerosis (MS).</p>
<p><xref ref-type="fig" rid="F3">Figure 3</xref> shows the pairwise scatterplots of the three depression questionnaires. The relation between PROMIS-D-8 and the other two indicators seems non-linear. This is in accordance with Amtmann et al. (<xref ref-type="bibr" rid="B1">2014</xref>), who argue that this questionnaire is subject to floor effects. Taking into account the characteristics of the data, i.e. the skewness depicted in <xref ref-type="fig" rid="F1">Figure 1</xref> and the observed non-linearity in <xref ref-type="fig" rid="F3">Figure 3</xref>, the proposed extension of the WMW method is an appropriate analysis.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>Figures <bold>A&#x02013;C</bold> depict the pairwise scatterplots of the three questionnaires (i.e. PHQ-9, CESD-10 and PROMIS-D-8) that are the indicators for the latent variable depression. The scatterplots reveal deviations from linearity for the scatterplots depicting the PROMIS-D-8.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-12-754898-g0003.tif"/>
</fig>
<p>Visual inspection via Q&#x02013;Q plots (see Figure 1 in <xref ref-type="supplementary-material" rid="SM2">Supplementary Data Sheet 2</xref>) shows that the assumption of location shift for the indicators is reasonably met. For the sake of completeness, the data are analyzed with the six estimation methods mentioned in the simulation study (Section 3): <bold>WMW&#x02013;max rel</bold>, <bold>WMW&#x02013;mean</bold>, <italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold>, <italic><bold>t</bold></italic><bold>-test&#x02013; mean</bold>, <bold>SEM</bold> and <bold>SEM&#x02013;correction</bold>. In order to compare two groups by using our proposed WMW extension or via SEM, model comparison tests indicate that measurement invariance across the two groups is established. After performing a data-driven model evaluation, a residual covariance between CESD-10 and PROMIS-D-8 is added for both groups in order to further enhance the fit of the model when using SEM. <xref ref-type="table" rid="T5">Table 5</xref> lists the <italic>p</italic>-values of all six methods and almost all methods lead to the conclusion that there is a significant difference in levels of depression between employed and unemployed people living with MS. Inference based on <italic><bold>t</bold></italic><bold>-test&#x02013;max rel</bold> is only marginally significant at the 5% significance level.</p>
<table-wrap position="float" id="T5">
<label>Table 5</label>
<caption><p>Effect sizes and inference based on the extended WMW methods, default and corrected SEM and the adapted <italic>t</italic>-tests when comparing the latent outcome variable depression between employed and unemployed patients with MS.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th/>
<th valign="top" align="center"><bold>Probabilistic index</bold></th>
<th valign="top" align="left"><bold><italic>p</italic>-value</bold></th>
</tr>
<tr>
<th/>
<th valign="top" align="center"><bold>P</bold> <bold>(Depres<sub><italic>unemployed</italic></sub>&#x0003E;Depres<sub><italic>employed</italic></sub>)</bold></th>
<th/>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">WMW&#x02013;max rel</td>
<td valign="top" align="center">56.48%</td>
<td valign="top" align="left">0.021</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td valign="top" align="left">WMW&#x02013;mean</td>
<td valign="top" align="center">56.92%</td>
<td valign="top" align="left">0.014</td>
</tr> <tr>
<td/>
<td valign="top" align="center"><bold>Mean difference</bold></td>
<td valign="top" align="left"><italic><bold>p</bold></italic><bold>-value</bold></td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center"><bold>unemployed - employed</bold></td>
<td/>
</tr> <tr>
<td valign="top" align="left"><italic>t</italic>-test&#x02013;max rel</td>
<td valign="top" align="center">1.14</td>
<td valign="top" align="left">0.055</td>
</tr>
<tr>
<td valign="top" align="left"><italic>t</italic>-test&#x02013;mean</td>
<td valign="top" align="center">0.19</td>
<td valign="top" align="left">0.034</td>
</tr>
<tr>
<td valign="top" align="left">SEM</td>
<td valign="top" align="center">1.444</td>
<td valign="top" align="left">0.0031</td>
</tr>
<tr>
<td valign="top" align="left">SEM&#x02013;correction</td>
<td valign="top" align="center">1.444</td>
<td valign="top" align="left">0.0033</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The interpretation of this difference does vary according to the used method. All parametric methods, i.e. the methods based on SEM and the <italic>t</italic>-test, give an estimation of the mean difference between the two groups where the employed patients are the reference group. Hence, a positive difference indicates that the unemployed patients have on average a higher score. Both <bold>SEM</bold> and <bold>SEM&#x02013;correction</bold> give an estimated difference in latent means for depression of 1.444. Analyses based on the <italic>t</italic>-test with a weighted sum or mean of the standardized indicators result in a difference of 1.14 and 0.19, respectively. Given that these modified <italic>t</italic>-tests do not take a measurement model into account, in contrast with SEM, its effect size merely reflects a difference in an overall outcome that aims to measure depression. Nevertheless, both analyses indicate that the group of unemployed patients score on average higher for depression than the group of employed patients.</p>
<p>For the extended WMW methods, the interpretation of the effect sizes is slightly different. Here, the effect size represents the probability that an unemployed patient has a higher score for depression than an employed patient. These probabilities are 56.48 and 56.92% for respectively <bold>WMW&#x02013;max rel</bold> and <bold>WMW&#x02013;mean</bold>. Consequently, patients who are unemployed have a significantly higher probability (around 56%) to have higher depression scores than patients that do have a job. The results thus show that the conclusion in this specific context for all methods points in the same direction. The R code for all analysis and figures is made available in the <xref ref-type="supplementary-material" rid="SM1">Supplementary Material</xref>.</p>
</sec>
<sec sec-type="discussion" id="s5">
<title>5. Discussion</title>
<p>In this paper, we proposed an extension of the Wilcoxon&#x02013;Mann&#x02013;Whitney test in the context of latent variables with a main focus on hypothesis testing. We introduced two strategies: one where the mean of the standardized indicators is used and one where a maximally reliable composite is created to form the input of the WMW test. The statistical properties of these proposed testing procedures were examined and compared with the performance of SEMs and two adapted <italic>t</italic>-tests in a simulation study. By comparing the two proposed strategies to each other in this simulation study, the costs and benefits of this maximization procedure were explored.</p>
<p>Even though SEM is omnipresent in the context of analyzing latent variables, we believe that a valuable alternative is provided in this paper. Concerning measurement model (1), the difference with the classical theory in SEM is the amount of knowledge that is required concerning the function <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;). Typically, <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) is assumed to be linear and is estimated during the analysis. In our proposed methodology where the mean of the standardized indicators is used, we relax this requirement since we only assume monotonicity and we do not need to estimate <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) as this is not our main interest. Some flexibility has thus been introduced in comparison with the traditional FA and SEM. Additionally, the variances of the measurement error do not need to be estimated in this strategy. For the theoretical validation of our methodology, we do need to impose the assumption of equal distributions for the measurement error per indicator across two groups. However, the simulation results suggest that violations have no detrimental consequences with respect to both the empirical Type I error rate and the power.</p>
<p>When a maximally reliable composite is used, i.e. the second strategy of the extended WMW test, the variance of the measurement error needs to be estimated in order to obtain the weights for this aggregated indicator. In the proposed formulas for the estimation of the variance, linearity among the indicators was imposed. Also here, deviations from this linearity assumption did not lead to problems with respect to hypothesis testing in our simulation study as long as the function <italic>h</italic><sub><italic>p</italic></sub>(&#x000B7;) is monotone. It may be clear that the use of a maximally reliable composite requires additional assumptions in comparison with the use of a simple mean, but based on our simulation study, these assumptions seem to be flexible in practice.</p>
<p>Using an aggregated indicator based on the data-driven maximization procedure entails an improvement compared with the use of an unweighted mean when the reliability of the indicators differs. These results confirm the theoretical expectation that fine-tuning the weights in line with the reliability of each indicator separately can result in an increase in empirical power. However, one should not needlessly use this maximization procedure as it can result in a small loss of power when e.g. the reliability of indicators is equal to one another and/or the sample size is small (i.e. <italic>m</italic> &#x0003D; <italic>n</italic> &#x0003D; 15).</p>
<p>Most interestingly for this paper is that the results show that the attractive properties of the original WMW method are transferred to the context of latent variables. The results confirm that the procedures based on the WMW test have superior power when the distribution is heavily tailed. This is a pattern that is also observed when simulating data in the context of observed outcome variables, as mentioned by Van der Vaart (<xref ref-type="bibr" rid="B32">2000</xref>) and Hollander et al. (<xref ref-type="bibr" rid="B14">2013</xref>).</p>
<p>A possible limitation for the user might be the effect size of the extended WMW test. Where SEM provides an estimate of the difference between two groups on the scale of the latent variable which can be standardized, the effect size of our proposed method is a probabilistic index. On the other hand, from a theoretical point of view, this latter effect size has attractive properties. A probabilistic index is scale invariant and robust to outliers.</p>
<p>A second limitation of this study is the sole focus on hypothesis testing. It is known that using the standard Wilcoxon&#x02013;Mann&#x02013;Whitney test in the context of measurement error leads to an underestimation of the true effect size (Coffin and Sukhatme, <xref ref-type="bibr" rid="B6">1996</xref>, <xref ref-type="bibr" rid="B7">1997</xref>; Faraggi, <xref ref-type="bibr" rid="B11">2000</xref>; Schisterman et al., <xref ref-type="bibr" rid="B28">2001</xref>; Tosteson et al., <xref ref-type="bibr" rid="B31">2005</xref>; Fuller, <xref ref-type="bibr" rid="B12">2009</xref>). In future research, adaptations to the proposed methodology can be studied to enhance the point estimation. Other possible directions for research are extensions for paired groups, multiple group comparison or comparing groups over time.</p>
<p>A third limitation of the suggested methodology is that it is inherently impossible to model the monotonic relation between the indicator and the latent variable, in contrast to SEM. The methodology presented in this paper extends the standard Wilcoxon&#x02013;Mann&#x02013;Whitney method and hence only uses the rank of the data. Closely related is the remark that the extended WMW test can only be applied after measurement invariance is determined by using SEM. The additional use of the extended WMW test is especially justified when the distribution is heavily tailed in order to profit from the attractive properties of the extended WMW test as discussed earlier.</p>
<p>To conclude, this paper validated the use of the WMW method in the context of latent variables by implementing some small adaptations, i.e. the creation of an aggregated indicator. The use of existing concepts not only facilitates the practical implementation for researchers and practitioners, the advantages of the original WMW method are also carried over into the new context. We believe that the combination of flexibility in the measurement model, the ability to allocate weights reflecting the reliability of indicators and the superiority in heavily tailed distributions results in a valuable methodology.</p>
</sec>
<sec sec-type="data-availability" id="s6">
<title>Data Availability Statement</title>
<p>Requests to access data with respect to the simulation study should be directed to Heidelinde Dehaene, <email>heidelinde.dehaene&#x00040;ugent.be</email>. Some previously collected existing data was used in this article and is not publicly available. Therefore, requests to access these data should be directed to the corresponding authors [i.e. Amtmann et al. (<xref ref-type="bibr" rid="B1">2014</xref>)].</p>
</sec>
<sec id="s7">
<title>Author Contributions</title>
<p>HD, JDN, and YR: conceptualization and methodology of the presented idea. HD: implementation of the simulation studies, data analysis and writing of the manuscript. JDN and YR: review and editing of the manuscript and supervision. All authors approved the submitted version.</p>
</sec>
<sec sec-type="funding-information" id="s8">
<title>Funding</title>
<p>This work was financially supported by a Special Research Fund (BOF) Starting Grant 01N00717 from Ghent University. The computational resources (Stevin Supercomputer Infrastructure) and services used in this work were provided by the VSC (Flemish Supercomputer Center), funded by Ghent University, FWO and the Flemish Government&#x02013;department EWI.</p>
</sec>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of Interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s9">
<title>Publisher&#x00027;s Note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec> </body>
<back>
<ack><p>Some previously collected existing data used in this manuscript were supported by a grant from the National Institute of Arthritis and Musculoskeletal and Skin Diseases of the National Institutes of Health under award number U01AR052171. The content is solely the responsibility of the authors and does not necessarily represent the official views of the National Institutes of Health. Some previously collected existing data used in this manuscript were also supported by grants from the National Institute on Disability, Independent Living, and Rehabilitation Research (NIDILRR grant number 90RT5023 and 90AR5013). NIDILRR is a Center within the Administration for Community Living (ACL), Department of Health and Human Services (HHS). The contents of this manuscript do not necessarily represent the policy of NIDILRR, ACL, HHS, and you should not assume endorsement by the Federal Government.</p>
<p>We would also like to thank the editor, the associate editor and the three reviewers for their constructive comments on our manuscript.</p>
</ack>
<sec sec-type="supplementary-material" id="s10">
<title>Supplementary Material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fpsyg.2021.754898/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fpsyg.2021.754898/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.ZIP" id="SM1" mimetype="application/zip" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Data_Sheet_2.pdf" id="SM2" mimetype="application/pdf" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<p>The R code to recreate the simulation study and to reproduce all analysis and figures of the case study is made available in the <xref ref-type="supplementary-material" rid="SM1">Supplementary Material</xref>.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Amtmann</surname> <given-names>D.</given-names></name> <name><surname>Kim</surname> <given-names>J.</given-names></name> <name><surname>Chung</surname> <given-names>H.</given-names></name> <name><surname>Bamer</surname> <given-names>A. M.</given-names></name> <name><surname>Askew</surname> <given-names>R. L.</given-names></name> <name><surname>Wu</surname> <given-names>S.</given-names></name> <etal/></person-group>. (<year>2014</year>). <article-title>Comparing CESD-10, PROMIS-9, and PROMIS depression instruments in individuals with multiple sclerosis</article-title>. <source>Rehabil. Psychol</source>. <volume>59</volume>:<fpage>220</fpage>. <pub-id pub-id-type="doi">10.1037/a0035919</pub-id><pub-id pub-id-type="pmid">24661030</pub-id></citation></ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Andresen</surname> <given-names>E. M.</given-names></name> <name><surname>Malmgren</surname> <given-names>J. A.</given-names></name> <name><surname>Carter</surname> <given-names>W. B.</given-names></name> <name><surname>Patrick</surname> <given-names>D. L.</given-names></name></person-group> (<year>1994</year>). <article-title>Screening for depression in well older adults: evaluation of a short form of the CES-D</article-title>. <source>Am. J. Prev. Med</source>. <volume>10</volume>, <fpage>77</fpage>&#x02013;<lpage>84</lpage>. <pub-id pub-id-type="doi">10.1016/S0749-3797(18)30622-6</pub-id><pub-id pub-id-type="pmid">8037935</pub-id></citation></ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bentler</surname> <given-names>P.</given-names></name></person-group> (<year>1968</year>). <article-title>Alpha-maximized factor analysis (alphamax): Its relation to alpha and canonical factor analysis</article-title>. <source>Psychometrika</source> <volume>33</volume>, <fpage>335</fpage>&#x02013;<lpage>345</lpage>. <pub-id pub-id-type="doi">10.1007/BF02289328</pub-id><pub-id pub-id-type="pmid">5243965</pub-id></citation></ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Blair</surname> <given-names>R. C.</given-names></name> <name><surname>Higgins</surname> <given-names>J. J.</given-names></name></person-group> (<year>1980</year>). <article-title>A comparison of the power of wilcoxon&#x00027;s rank-sum statistic to that of student&#x00027;s t statistic under various nonnormal distributions</article-title>. <source>J. Educ. Stat</source>. <volume>5</volume>, <fpage>309</fpage>&#x02013;<lpage>335</lpage>. <pub-id pub-id-type="doi">10.2307/1164905</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chou</surname> <given-names>C.-P.</given-names></name> <name><surname>Bentler</surname> <given-names>P. M.</given-names></name> <name><surname>Satorra</surname> <given-names>A.</given-names></name></person-group> (<year>1991</year>). <article-title>Scaled test statistics and robust standard errors for non-normal data in covariance structure analysis: a monte carlo study</article-title>. <source>Br. J. Math. Stat. Psychol</source>. <volume>44</volume>, <fpage>347</fpage>&#x02013;<lpage>357</lpage>. <pub-id pub-id-type="doi">10.1111/j.2044-8317.1991.tb00966.x</pub-id><pub-id pub-id-type="pmid">1772802</pub-id></citation></ref>
<ref id="B6">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Coffin</surname> <given-names>M.</given-names></name> <name><surname>Sukhatme</surname> <given-names>S.</given-names></name></person-group> (<year>1996</year>). <article-title>A parametric approach to measurement errors in receiver operating characteristic studies,</article-title> in <source>Lifetime Data: Models in Reliability and Survival Analysis</source> (<publisher-loc>Boston, MA</publisher-loc>: <publisher-name>Springer</publisher-name>), <fpage>71</fpage>&#x02013;<lpage>75</lpage>.</citation>
</ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Coffin</surname> <given-names>M.</given-names></name> <name><surname>Sukhatme</surname> <given-names>S.</given-names></name></person-group> (<year>1997</year>). <article-title>Receiver operating characteristic studies and measurement errors</article-title>. <source>Biometrics</source> <volume>53</volume>, <fpage>823</fpage>&#x02013;<lpage>837</lpage>. <pub-id pub-id-type="doi">10.2307/2533545</pub-id><pub-id pub-id-type="pmid">9333348</pub-id></citation></ref>
<ref id="B8">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cohen</surname> <given-names>J.</given-names></name></person-group> (<year>1988</year>). <source>Statistical Power Analysis for the Behavioral Sciences, 2nd Edn</source>. <publisher-loc>Hillsdale, NJ</publisher-loc>: <publisher-name>Erlbaum</publisher-name>.</citation>
</ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>De Neve</surname> <given-names>J.</given-names></name> <name><surname>Dehaene</surname> <given-names>H.</given-names></name></person-group> (<year>2021</year>). <article-title>Semiparametric linear transformation models for indirectly observed outcomes</article-title>. <source>Stat. Med</source>. <volume>40</volume>, <fpage>2286</fpage>&#x02013;<lpage>2303</lpage>. <pub-id pub-id-type="doi">10.1002/sim.8903</pub-id><pub-id pub-id-type="pmid">33565108</pub-id></citation></ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>De Schryver</surname> <given-names>M.</given-names></name> <name><surname>De Neve</surname> <given-names>J.</given-names></name></person-group> (<year>2019</year>). <article-title>A tutorial on probabilistic index models: Regression models for the effect size P(Y1 &#x0003C; Y2)</article-title>. <source>Psychol. Methods</source> <volume>24</volume>:<fpage>403</fpage>. <pub-id pub-id-type="doi">10.1037/met0000194</pub-id><pub-id pub-id-type="pmid">30265047</pub-id></citation></ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Faraggi</surname> <given-names>D.</given-names></name></person-group> (<year>2000</year>). <article-title>The effect of random measurement error on receiver operating characteristic (roc) curves</article-title>. <source>Stat. Med</source>. <volume>19</volume>, <fpage>61</fpage>&#x02013;<lpage>70</lpage>. <pub-id pub-id-type="doi">10.1002/(SICI)1097-0258(20000115)19:1&#x0003C;61::AID-SIM297&#x0003E;3.0.CO;2-A</pub-id><pub-id pub-id-type="pmid">10623913</pub-id></citation></ref>
<ref id="B12">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Fuller</surname> <given-names>W. A.</given-names></name></person-group> (<year>2009</year>). <source>Measurement Error Models, Vol. 305</source>. <publisher-loc>New York, NY</publisher-loc>: <publisher-name>John Wiley &#x00026; Sons</publisher-name>.</citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Garcia-Marques</surname> <given-names>L.</given-names></name> <name><surname>Garcia-Marques</surname> <given-names>T.</given-names></name> <name><surname>Brauer</surname> <given-names>M.</given-names></name></person-group> (<year>2014</year>). <article-title>Buy three but get only two: The smallest effect in a 2 &#x000D7;2 anova is always uninterpretable</article-title>. <source>Psychon. Bull. Rev</source>. <volume>21</volume>, <fpage>1415</fpage>&#x02013;<lpage>1430</lpage>. <pub-id pub-id-type="doi">10.3758/s13423-014-0640-3</pub-id><pub-id pub-id-type="pmid">24841234</pub-id></citation></ref>
<ref id="B14">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Hollander</surname> <given-names>M.</given-names></name> <name><surname>Wolfe</surname> <given-names>D. A.</given-names></name> <name><surname>Chicken</surname> <given-names>E.</given-names></name></person-group> (<year>2013</year>). <source>Nonparametric Statistical Methods, Vol. 751</source>. John Wiley &#x00026; Sons.</citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jackson</surname> <given-names>D. L.</given-names></name></person-group> (<year>2001</year>). <article-title>Sample size and number of parameter estimates in maximum likelihood confirmatory factor analysis: a monte carlo investigation</article-title>. <source>Struct. Equ. Model</source>. <volume>8</volume>, <fpage>205</fpage>&#x02013;<lpage>223</lpage>. <pub-id pub-id-type="doi">10.1207/S15328007SEM0802_3</pub-id></citation>
</ref>
<ref id="B16">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Kline</surname> <given-names>R. B.</given-names></name></person-group> (<year>2015</year>). <source>Principles and Practice of Structural Equation Modeling</source>. <publisher-loc>New York, NY</publisher-loc>: <publisher-name>Guilford Publications</publisher-name>.</citation>
</ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kroenke</surname> <given-names>K.</given-names></name> <name><surname>Spitzer</surname> <given-names>R. L.</given-names></name></person-group> (<year>2002</year>). <article-title>The PHQ-9: a new depression diagnostic and severity measure</article-title>. <source>Psychiatr. Ann</source>. <volume>32</volume>, <fpage>509</fpage>&#x02013;<lpage>515</lpage>. <pub-id pub-id-type="doi">10.3928/0048-5713-20020901-06</pub-id></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kroenke</surname> <given-names>K.</given-names></name> <name><surname>Spitzer</surname> <given-names>R. L.</given-names></name> <name><surname>Williams</surname> <given-names>J. B.</given-names></name></person-group> (<year>2001</year>). <article-title>The PHQ-9: validity of a brief depression severity measure</article-title>. <source>J. Gen. Intern. Med</source>. <volume>16</volume>, <fpage>606</fpage>&#x02013;<lpage>613</lpage>. <pub-id pub-id-type="doi">10.1046/j.1525-1497.2001.016009606.x</pub-id><pub-id pub-id-type="pmid">11556941</pub-id></citation></ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lehmann</surname> <given-names>E. L.</given-names></name></person-group> (<year>1951</year>). <article-title>Consistency and unbiasedness of certain nonparametric tests</article-title>. <source>Ann. Math. Stat</source>. <volume>22</volume>, <fpage>165</fpage>&#x02013;<lpage>179</lpage>. <pub-id pub-id-type="doi">10.1214/aoms/1177729639</pub-id></citation>
</ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>H.</given-names></name></person-group> (<year>1997</year>). <article-title>A unifying expression for the maximal reliability of a linear composite</article-title>. <source>Psychometrika</source> <volume>62</volume>, <fpage>245</fpage>&#x02013;<lpage>249</lpage>. <pub-id pub-id-type="doi">10.1007/BF02295278</pub-id></citation>
</ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mann</surname> <given-names>H. B.</given-names></name> <name><surname>Whitney</surname> <given-names>D. R.</given-names></name></person-group> (<year>1947</year>). <article-title>On a test of whether one of two random variables is stochastically larger than the other</article-title>. <source>Ann. Math. Stat</source>. <volume>18</volume>, <fpage>50</fpage>&#x02013;<lpage>60</lpage>. <pub-id pub-id-type="doi">10.1214/aoms/1177730491</pub-id></citation>
</ref>
<ref id="B22">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Nunnally</surname> <given-names>J. C.</given-names></name> <name><surname>Bernstein</surname> <given-names>I. H.</given-names></name></person-group> (<year>1994</year>). <source>Psychometric Theory, 3rd Edn</source>. <publisher-loc>New York, NY</publisher-loc>: <publisher-name>McGraw-Hill</publisher-name>.</citation>
</ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Penev</surname> <given-names>S.</given-names></name> <name><surname>Raykov</surname> <given-names>T.</given-names></name></person-group> (<year>2006</year>). <article-title>Maximal reliability and power in covariance structure models</article-title>. <source>Br. J. Math. Stat. Psychol</source>. <volume>59</volume>, <fpage>75</fpage>&#x02013;<lpage>87</lpage>. <pub-id pub-id-type="doi">10.1348/000711005X68183</pub-id><pub-id pub-id-type="pmid">16709280</pub-id></citation></ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pilkonis</surname> <given-names>P. A.</given-names></name> <name><surname>Choi</surname> <given-names>S. W.</given-names></name> <name><surname>Reise</surname> <given-names>S. P.</given-names></name> <name><surname>Stover</surname> <given-names>A. M.</given-names></name> <name><surname>Riley</surname> <given-names>W. T.</given-names></name> <name><surname>Cella</surname> <given-names>D.</given-names></name></person-group> PROMIS Cooperative Group. (<year>2011</year>). <article-title>Item banks for measuring emotional distress from the Patient-Reported Outcomes Measurement Information System (PROMIS&#x000AE;): depression, anxiety, and anger</article-title>. <source>Assessment</source> <volume>18</volume>, <fpage>263</fpage>&#x02013;<lpage>283</lpage>.<pub-id pub-id-type="pmid">21697139</pub-id></citation></ref>
<ref id="B25">
<citation citation-type="book"><person-group person-group-type="author"><collab>R Core Team</collab></person-group> (<year>2020</year>). <source>R: A Language and Environment for Statistical Computing</source>. <publisher-loc>Vienna</publisher-loc>: <publisher-name>R Foundation for Statistical Computing</publisher-name>.</citation>
</ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rosseel</surname> <given-names>Y.</given-names></name></person-group> (<year>2012</year>). <article-title>lavaan: an R package for structural equation modeling</article-title>. <source>J. Stat. Softw</source>. <volume>48</volume>, <fpage>1</fpage>&#x02013;<lpage>36</lpage>. <pub-id pub-id-type="doi">10.18637/jss.v048.i02</pub-id><pub-id pub-id-type="pmid">25601849</pub-id></citation></ref>
<ref id="B27">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Satorra</surname> <given-names>A.</given-names></name> <name><surname>Bentler</surname> <given-names>P. M.</given-names></name></person-group> (<year>1994</year>). Corrections to test statistics and standard errors in co-variance structure analysis, in <source>Latent Variables Analysis: Applications for Developmental Research</source>, eds A. von Eye and C. C. Clogg (Thousands Oaks: Sage), <fpage>399</fpage>&#x02013;<lpage>419</lpage>.</citation>
</ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schisterman</surname> <given-names>E. F.</given-names></name> <name><surname>Faraggi</surname> <given-names>D.</given-names></name> <name><surname>Reiser</surname> <given-names>B.</given-names></name> <name><surname>Trevisan</surname> <given-names>M.</given-names></name></person-group> (<year>2001</year>). <article-title>Statistical inference for the area under the receiver operating characteristic curve in the presence of random measurement error</article-title>. <source>Am. J. Epidemiol</source>. <volume>154</volume>, <fpage>174</fpage>&#x02013;<lpage>179</lpage>. <pub-id pub-id-type="doi">10.1093/aje/154.2.174</pub-id><pub-id pub-id-type="pmid">11447052</pub-id></citation></ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Spitzer</surname> <given-names>R. L.</given-names></name> <name><surname>Kroenke</surname> <given-names>K.</given-names></name> <name><surname>Williams</surname> <given-names>J. B.</given-names></name> <name><surname>Group</surname> <given-names>P. H. Q. P. C. S.</given-names></name></person-group> (<year>1999</year>). <article-title>Validation and utility of a self-report version of PRIME-MD: the PHQ primary care study</article-title>. <source>JAMA</source> <volume>282</volume>, <fpage>1737</fpage>&#x02013;<lpage>1744</lpage>. <pub-id pub-id-type="doi">10.1001/jama.282.18.1737</pub-id><pub-id pub-id-type="pmid">10568646</pub-id></citation></ref>
<ref id="B30">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Thas</surname> <given-names>O.</given-names></name></person-group> (<year>2010</year>). <source>Comparing Distributions</source>. <publisher-loc>New York, NY</publisher-loc>: <publisher-name>Springer</publisher-name>.</citation>
</ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tosteson</surname> <given-names>T. D.</given-names></name> <name><surname>Buonaccorsi</surname> <given-names>J. P.</given-names></name> <name><surname>Demidenko</surname> <given-names>E.</given-names></name> <name><surname>Wells</surname> <given-names>W. A.</given-names></name></person-group> (<year>2005</year>). <article-title>Measurement error and confidence intervals for roc curves</article-title>. <source>Biometr. J</source>. <volume>47</volume>, <fpage>409</fpage>&#x02013;<lpage>416</lpage>. <pub-id pub-id-type="doi">10.1002/bimj.200310159</pub-id><pub-id pub-id-type="pmid">16161800</pub-id></citation></ref>
<ref id="B32">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Van der Vaart</surname> <given-names>A. W.</given-names></name></person-group> (<year>2000</year>). <source>Asymptotic Statistics, Vol. 3</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wilcoxon</surname> <given-names>F.</given-names></name></person-group> (<year>1945</year>). <article-title>Individual comparisons by ranking methods</article-title>. <source>Biometr. Bull</source>. <volume>1</volume>, <fpage>80</fpage>&#x02013;<lpage>83</lpage>. <pub-id pub-id-type="doi">10.2307/3001968</pub-id></citation>
</ref>
</ref-list> 
</back>
</article> 
