<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2023.1132128</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Comparison of false positive and false negative rates of two indices of individual reliable change: Jacobson-Truax and Hageman-Arrindell methods</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Ferrer-Urbina</surname>
<given-names>Rodrigo</given-names>
</name>
<xref rid="aff1" ref-type="aff"><sup>1</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/1255750/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Pardo</surname>
<given-names>Antonio</given-names>
</name>
<xref rid="aff2" ref-type="aff"><sup>2</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/655199/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Arrindell</surname>
<given-names>Willem A.</given-names>
</name>
<xref rid="aff3" ref-type="aff"><sup>3</sup></xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Puddu-Gallardo</surname>
<given-names>Giannina</given-names>
</name>
<xref rid="aff1" ref-type="aff"><sup>1</sup></xref>
<xref rid="c001" ref-type="corresp"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/2152537/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Universidad de Tarapac&#x00E1;, Av. General Vel&#x00E1;squez</institution>, <addr-line>Arica</addr-line>, <country>Chile</country></aff>
<aff id="aff2"><sup>2</sup><institution>Universidad Aut&#x00F3;noma de Madrid, Ciudad Universitaria de Cantoblanco</institution>, <addr-line>Madrid</addr-line>, <country>Spain</country></aff>
<aff id="aff3"><sup>3</sup><institution>University of Social Sciences and Humanities, Vietnam National University</institution>, <addr-line>Ho Chi Minh City</addr-line>, <country>Vietnam</country></aff>
<author-notes>
<fn fn-type="edited-by" id="fn0001">
<p>Edited by: Semira Tagliabue, Catholic University of the Sacred Heart, Italy</p>
</fn>
<fn fn-type="edited-by" id="fn0002">
<p>Reviewed by: Angela Sorgente, Catholic University of the Sacred Heart, Italy; Gary Baker, Champlain College, United States</p>
</fn>
<corresp id="c001">&#x002A;Correspondence: Giannina Puddu-Gallardo, <email>ninapuddu@gmail.com</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>13</day>
<month>07</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>14</volume>
<elocation-id>1132128</elocation-id>
<history>
<date date-type="received">
<day>26</day>
<month>12</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>23</day>
<month>06</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2023 Ferrer-Urbina, Pardo, Arrindell and Puddu-Gallardo.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Ferrer-Urbina, Pardo, Arrindell and Puddu-Gallardo</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<sec id="sec1">
<title>Background</title>
<p>Quantification of change is crucial for correctly estimating the effect of a treatment and for distinguishing random or non-systematic changes from substantive changes. The objective of the present study was to learn about the performance of two distribution-based methods [the Jacobson-Truax Reliable Change Index (RCI) and the Hageman-Arrindell (HA) approach] that were designed for evaluating individual reliable change.</p>
</sec>
<sec id="sec2">
<title>Methods</title>
<p>A pre-post design was simulated with the purpose to evaluate the false positive and false negative rates of RCI and HA methods. In this design, a first measurement is obtained before treatment and a second measurement is obtained after treatment, in the same group of subjects.</p>
</sec>
<sec id="sec3">
<title>Results</title>
<p>In relation to the rate of false positives, only the HA statistic provided acceptable results. Regarding the rate of false negatives, both statistics offered similar results, and both could claim to offer acceptable rates when Ferguson&#x2019;s stringent criteria were used to define effect sizes as opposed to when the conventional criteria advanced by Cohen were employed.</p>
</sec>
<sec id="sec4">
<title>Conclusion</title>
<p>Since the HA statistic appeared to be a better option than the RCI statistic, we have developed and presented an Excel macro so that the greater complexity of calculating HA would not represent an obstacle for the non-expert user.</p>
</sec>
</abstract>
<kwd-group>
<kwd>individual reliable change</kwd>
<kwd>assessment of change</kwd>
<kwd>Jacobson-Truax method</kwd>
<kwd>Hageman-Arrindell approach</kwd>
<kwd>false negatives</kwd>
<kwd>false positives</kwd>
</kwd-group>
<counts>
<fig-count count="0"/>
<table-count count="2"/>
<equation-count count="3"/>
<ref-count count="72"/>
<page-count count="10"/>
<word-count count="9197"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Psychology for Clinical Settings</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="sec5">
<title>Introduction</title>
<p>In the field of applied research, having objective and reliable measures to assess the change experienced after an intervention is crucial, for example in a clinical context, the interpretation of the results of a treatment will influence clinical decision-making, including the safety and efficacy of the patient (<xref ref-type="bibr" rid="ref50">Page, 2014</xref>). In recent decades, there has been an increase in pre-post study designs that include measures to assess the efficacy of an intervention or treatment, in an effort to redirect practice in a more oriented to &#x201C;Evidence-Based Practice&#x201D; (<xref ref-type="bibr" rid="ref50">Page, 2014</xref>). The pre-post design studies are especially useful in the clinical context since they allow to measure the variations in a variable of interest (e.g., depression and/or anxiety symptoms, consumption patterns of any substance, etc.) before and after an intervention and therefore evaluate the success of the strategy used, like to define the clinically meaningful change in the GAD-7 scale (<xref ref-type="bibr" rid="ref6">Bischoff et al., 2020</xref>); compare different treatment approaches as multi-family groups (MFG) (<xref ref-type="bibr" rid="ref67">Vardanian et al., 2020</xref>); or assess clinical change in mental health with psychiatric patients (<xref ref-type="bibr" rid="ref60">Shalaby et al., 2022</xref>). Although the quantification of this change is essential to correctly estimate the effect of a treatment, this itself is not enough since it must also be able to distinguish random or non-systematic changes from substantive changes.</p>
<p>In this context, a clinician or researcher could draw any of four conclusions: correctly conclude that change has taken place (true positive); correctly conclude that no change has taken place (true negative); erroneously conclude that significant change has taken place (the result is positive), when in reality such a change has not taken place (false positive); or erroneously conclude that no significant change has taken place (the result is negative), when in reality meaningful change has taken place (false negative).</p>
<p>Among the available strategies for assessing change, <italic>distribution-based methods</italic> are the most used (see 1). These are a set of techniques designed for identifying clinically meaningful change, based on the statistical properties of magnitude estimates of change and data variability, this mean that, can be estimated based on the distribution of observed scores in a relevant sample (<xref ref-type="bibr" rid="ref54">Revicki et al., 2008</xref>).</p>
<p>These methods have been designed in the context of assessing clinically meaningful change to identify reliable change, i.e., minimum variations that should occur in the patients&#x2019; answers to be able to conclude that significant change has been made (<xref ref-type="bibr" rid="ref46">McGlinchey et al., 2002</xref>; <xref ref-type="bibr" rid="ref17">Crosby et al., 2003</xref>; <xref ref-type="bibr" rid="ref28">Gatchel and Mayer, 2010</xref>; <xref ref-type="bibr" rid="ref66">Turner et al., 2010</xref>). To accomplish this purpose, distribution-based methods must be able to identify those substantive changes (true positive) other than randomness from randomly attributable changes (true negative).</p>
<p>For these reasons, some studies were conducted to compare the accuracy of the performance of the different methods by identifying misclassifications in simulated scenarios, specifically the quantity of changes detected when the variations were only random (false positive) (<xref ref-type="bibr" rid="ref51">Pardo and Ferrer, 2013</xref>) and the amount of undetected changes when the variations were systematic (false negative) (<xref ref-type="bibr" rid="ref26">Ferrer and Pardo, 2019</xref>).</p>
<p>More than three decades have elapsed since Jacobson, Follette and Revenstorf (<xref ref-type="bibr" rid="ref38">Jacobson et al., 1984</xref>) proposed the <italic>reliable change index</italic> (RCI) for assessing <italic>individual change</italic> as an alternative to the assessment of <italic>group change</italic> offered by the classical null hypothesis significance tests and measures of effect size. Along these years, the RCI has undergone some corrections by his own promoters (<xref ref-type="bibr" rid="ref40">Jacobson and Truax, 1991</xref>; <xref ref-type="bibr" rid="ref39">Jacobson et al., 1999</xref>) and many other researchers have proposed alternatives procedures for trying to improve accuracy and effectiveness in identifying significant or reliable changes (see, for example, <xref ref-type="bibr" rid="ref49">Nunnally and Kotsch, 1983</xref>; <xref ref-type="bibr" rid="ref10">Christensen and Mendoza, 1986</xref>; <xref ref-type="bibr" rid="ref35">Hsu, 1989</xref>, <xref ref-type="bibr" rid="ref36">1995</xref>, <xref ref-type="bibr" rid="ref37">1996</xref>; <xref ref-type="bibr" rid="ref64">Speer, 1992</xref>; <xref ref-type="bibr" rid="ref15">Crawford and Howell, 1998</xref>; <xref ref-type="bibr" rid="ref32">Hageman and Arrindell, 1999</xref>; <xref ref-type="bibr" rid="ref44">Maassen, 2004</xref>; <xref ref-type="bibr" rid="ref69">Wyrwich, 2004</xref>; <xref ref-type="bibr" rid="ref14">Crawford and Garthwaite, 2006</xref>; <xref ref-type="bibr" rid="ref8">Botella et al., 2018</xref>). It is important to emphasize that to estimate these measures of individual change, we require a distribution of observed data as a reference, which could be obtained from previous studies or field-related reference studies.</p>
<p>Despite the alternative proposals, the RCI statistic has become the most widely used index for assessing individual change in pre-post designs in clinical settings (according to &#x201C;Web of Science,&#x201D; the Jacobson and Truax paper (<xref ref-type="bibr" rid="ref40">Jacobson and Truax, 1991</xref>) has received 6,843 citations until December 2022). However, the fact that a method is widely used does not mean that it is free of problems. In a study designed to assess the performance of different indices of individual change, <xref ref-type="bibr" rid="ref26">Ferrer and Pardo (2019)</xref> have shown that false positive rates obtained with RCI are unacceptably high: depending on the context, these rates oscillate between 0 and 39.7% (between 5.0 and 34.3% when working with normal distributions), when in fact the expected values due to the cutoff points established should be around 5%.</p>
<sec id="sec6">
<title><italic>RCI</italic> versus <italic>HA</italic></title>
<p>Several researchers have proposed similar methods to RCI in an attempt to improve their performance (for a review, see <xref ref-type="bibr" rid="ref17">Crosby et al., 2003</xref>; <xref ref-type="bibr" rid="ref26">Ferrer and Pardo, 2019</xref>). Many of these methods have been compared with each other to determine whether or not they made equivalent classifications; and results of these studies have shown some consistency (<xref ref-type="bibr" rid="ref23">Estrada et al., 2019</xref>, <xref ref-type="bibr" rid="ref22">2020</xref>).</p>
<p><xref ref-type="bibr" rid="ref46">McGlinchey et al. (2002)</xref> compared five distribution-based methods: the reliable change index (RCI) (<xref ref-type="bibr" rid="ref38">Jacobson et al., 1984</xref>); the Edwards-Nunnally method (EN) (<xref ref-type="bibr" rid="ref64">Speer, 1992</xref>); the Gulliksen-Lord-Novick method (GLN) (<xref ref-type="bibr" rid="ref35">Hsu, 1989</xref>; <xref ref-type="bibr" rid="ref44">Maassen, 2004</xref>); a method based on the hierarchical linear modeling (HLM) (<xref ref-type="bibr" rid="ref64">Speer, 1992</xref>); and the Hageman-Arrindell method (HA) (<xref ref-type="bibr" rid="ref32">Hageman and Arrindell, 1999</xref>). <xref ref-type="bibr" rid="ref46">McGlinchey et al. (2002)</xref> concluded that all methods offer similar results, with the exception of the HA method, which tends to be more conservative, this means that it tends to identify fewer changes than the other methods: &#x201C;&#x2026;there will need to be relatively greater change with the HA method for an individual to be considered reliably improved&#x201D; (p. 543).</p>
<p>In a similar way, <xref ref-type="bibr" rid="ref56">Ronk et al. (2012)</xref> found that the HA method yielded different results from the rest of the methods studied (RCI, GLN, EN, and NK) (22), which offered similar performances to each other. <xref ref-type="bibr" rid="ref3">Bauer et al. (2004)</xref>, after comparing five distribution-based methods (RCI, GLN, EN, HLM and HA), conclude that the HA method is &#x201C;the most conservative.&#x201D; Despite that <xref ref-type="bibr" rid="ref57">Ronk et al. (2016)</xref>, in a comparative study between RCI and HA, conclude that there is no discernible advantage in the use of one method over the other, the results reported in this study (see Table 3, p. 5) shown that the percentage of patients classified as &#x201C;recovered&#x201D; were systematically lower with the HA method than with the RCI method.</p>
<p>The results obtained on these empirical studies are in close agreement with those obtained from simulation studies. <xref ref-type="bibr" rid="ref1">Atkins et al. (2005)</xref> have found that &#x201C;the HA method is the most conservative&#x201D; (p. 986) of the four compared (RCI, GLN, EN, HA), i.e., it is the method that classifies less cases as recovered. Indeed, <xref ref-type="bibr" rid="ref51">Pardo and Ferrer (2013)</xref> have shown that, although both RCI and HA offer unacceptably high false positive rates, the HA method offers a rate systematically lower than the one obtained with RCI.</p>
<p>In this context, one may wonder what makes HA work differently from RCI and other distribution-based methods. We believe that the answer to this could be that the HA statistic incorporates some details not taken into account by the RCI statistic (or by any other method based on distribution). While the RCI statistic is obtained by <xref ref-type="bibr" rid="ref38">Jacobson et al. (1984</xref>, p. 14).</p>
<disp-formula id="E1">
<mml:math id="M1">
<mml:mrow>
<mml:mi>R</mml:mi>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>Y</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
<mml:mrow>
<mml:msqrt>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:msup>
<mml:mrow>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:msub>
<mml:mi>S</mml:mi>
<mml:mi>X</mml:mi>
</mml:msub>
<mml:msqrt>
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>R</mml:mi>
<mml:mrow>
<mml:mi>X</mml:mi>
<mml:mi>X</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:msqrt>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
</mml:mrow>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:msqrt>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
</disp-formula>
<p>(<italic>X<sub>i</sub></italic>&#x2009;=&#x2009;individual pre-test score; <italic>Y<sub>i</sub></italic>&#x2009;=&#x2009;individual post-test score; <italic>S<sub>X</sub></italic>&#x2009;=&#x2009;standard deviation of pre-test; <italic>R<sub>XX</sub></italic>&#x2009;=&#x2009;reliability of test), the HA statistic (<xref ref-type="bibr" rid="ref39">Jacobson et al., 1999</xref>, p.1173) is obtained by</p>
<disp-formula id="E2">
<mml:math id="M2">
<mml:mrow>
<mml:mi>H</mml:mi>
<mml:mi>A</mml:mi>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:msub>
<mml:mi>Y</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
<mml:msub>
<mml:mi>R</mml:mi>
<mml:mrow>
<mml:mi>D</mml:mi>
<mml:mi>D</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>+</mml:mo>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:msub>
<mml:mi>M</mml:mi>
<mml:mi>Y</mml:mi>
</mml:msub>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>M</mml:mi>
<mml:mi>X</mml:mi>
</mml:msub>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>R</mml:mi>
<mml:mrow>
<mml:mi>D</mml:mi>
<mml:mi>D</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
</mml:mrow>
<mml:mrow>
<mml:msqrt>
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:msub>
<mml:mi>R</mml:mi>
<mml:mrow>
<mml:mi>D</mml:mi>
<mml:mi>D</mml:mi>
</mml:mrow>
</mml:msub>
<mml:msup>
<mml:mrow>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:msub>
<mml:mi>S</mml:mi>
<mml:mi>X</mml:mi>
</mml:msub>
<mml:msqrt>
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>R</mml:mi>
<mml:mrow>
<mml:mi>X</mml:mi>
<mml:mi>X</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:msqrt>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
</mml:mrow>
<mml:mn>2</mml:mn>
</mml:msup>
</mml:mrow>
</mml:msqrt>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
</disp-formula>
<p>(<italic>M<sub>X</sub></italic>&#x2009;=&#x2009;mean of the pre-test scores; <italic>M<sub>Y</sub></italic>&#x2009;=&#x2009;mean of the post-test scores; <italic>R<sub>DD</sub></italic>&#x2009;=&#x2009;reliability of the pre-post differences).</p>
<p>The approach of <xref ref-type="bibr" rid="ref32">Hageman and Arrindell (1999)</xref> tries to improve the accuracy of RCI by incorporating <italic>the reliability of the pre-post differences</italic>. Since working with pre-post differences has generated a lot of controversy among those who theorize about the psychometric properties of tests from classical test theory (due to the possible lack of reliability of this type of scores; see <xref ref-type="bibr" rid="ref42">Lord, 1956</xref>, <xref ref-type="bibr" rid="ref43">1963</xref>; <xref ref-type="bibr" rid="ref55">Rogosa and Willett, 1983</xref>), ignoring pre-post differences reliability does not seem the best way to proceed.</p>
<p>Therefore, the most remarkable difference between RCI and HA is that HA includes the reliability of differences (<italic>R<sub>DD</sub></italic>). If <italic>R<sub>DD</sub></italic> is perfect (<italic>R<sub>DD</sub></italic>&#x2009;=&#x2009;1), RCI and HA take identical values. If <italic>R<sub>DD</sub></italic> is not perfect (<italic>R<sub>DD</sub></italic>&#x2009;&#x003C;&#x2009;1), the HA formula does not clearly show what happens (because <italic>R<sub>DD</sub></italic> plays a different part in the numerator and denominator), but both empirical and simulation studies indicate that as the value of <italic>R<sub>DD</sub></italic> decreases, so does the value of HA, and that is why HA tends to make classifications more conservative than other distribution-based methods.</p>
</sec>
<sec id="sec7">
<title>How to estimate reliability</title>
<p>The confirmation that HA produces more conservative classifications than RCI (and more conservative than other distribution-based methods) is important considering that these methods tend to offer too high false positives rates.</p>
<p>But why do all distribution-based methods (including HA) offer excessively high false positives rates? The RCI and HA equations shown above (including the equations of other distribution-based methods) show that both statistics are based on the <italic>standard error of measurement</italic> (<italic>SEM</italic>), which is obtained by</p>
<disp-formula id="E3">
<mml:math id="M3">
<mml:mrow>
<mml:mi>S</mml:mi>
<mml:mi>E</mml:mi>
<mml:mi>M</mml:mi>
<mml:mo>=</mml:mo>
<mml:msub>
<mml:mi>S</mml:mi>
<mml:mi>X</mml:mi>
</mml:msub>
<mml:msqrt>
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>R</mml:mi>
<mml:mrow>
<mml:mi>X</mml:mi>
<mml:mi>X</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:msqrt>
</mml:mrow>
</mml:math>
</disp-formula>
<p>As we can see in the above equation, SEM depends on (a) the standard deviation of the pretest scores <italic>S<sub>X</sub></italic> and (b) test reliability <italic>R<sub>XX</sub></italic>. Now while there is only one way to calculate <italic>S<sub>X</sub></italic>, there are many ways to calculate <italic>R<sub>XX</sub></italic>. Each of these different approaches has advantages and disadvantages, but in the field of health sciences, the strategies most used are based on internal consistency (usually estimated by Cronbach&#x2019;s coefficient alpha) (<xref ref-type="bibr" rid="ref16">Cronbach, 1951</xref>) or on temporal stability (usually estimated by the test&#x2013;retest correlation). <xref ref-type="bibr" rid="ref45">Martinovich et al. (1996)</xref>, after reflecting on the pros and cons of both strategies in the field of individual change assessment, recommended estimating reliability using internal consistency, especially for clinical populations, because test&#x2013;retest reliability is reduced by the presence of true individual test&#x2013;retest change, even without patients being on therapy during that period. and <xref ref-type="bibr" rid="ref71">Wyrwich et al. (1999)</xref> also recommended estimating reliability by the alpha coefficient.</p>
<p>However, the psychometric literature contains numerous studies that advise against using alpha to estimate reliability (<xref ref-type="bibr" rid="ref59">Schmitt, 1996</xref>; <xref ref-type="bibr" rid="ref5">Bentler, 2009</xref>; <xref ref-type="bibr" rid="ref31">Green and Yang, 2009</xref>; <xref ref-type="bibr" rid="ref53">Revelle and Zinbarg, 2009</xref>; <xref ref-type="bibr" rid="ref62">Sijtsma, 2009</xref>; <xref ref-type="bibr" rid="ref21">Dunn et al., 2014</xref>; <xref ref-type="bibr" rid="ref18">Crutzen and Peters, 2017</xref>). On the one hand, there is evidence that Cronbach&#x2019;s alpha is not really an indicator of the internal consistency of a test (see, for example, <xref ref-type="bibr" rid="ref62">Sijtsma, 2009</xref>). On the other hand, if a test is unidimensional, it is known that: (a) Only when the tau-equivalent assumption is assumed does the alpha coefficient produce results that are comparable to those of other measures of internal consistency (<xref ref-type="bibr" rid="ref29">Graham, 2006</xref>), and (b) the reliability estimated through the alpha coefficient is higher than the one estimated using the test&#x2013;retest correlation (<xref ref-type="bibr" rid="ref4">Becker, 2000</xref>; <xref ref-type="bibr" rid="ref33">Hogan et al., 2000</xref>; <xref ref-type="bibr" rid="ref30">Green, 2003</xref>; <xref ref-type="bibr" rid="ref58">Schmidt et al., 2003</xref>).</p>
<p>When this is considered, it seems that the recommendations given by <xref ref-type="bibr" rid="ref45">Martinovich et al. (1996)</xref>, and <xref ref-type="bibr" rid="ref71">Wyrwich et al. (1999)</xref> would lead to evaluating statistically reliable change through the use of an underestimated value of SEM; and this is precisely what could justify, at least partially, the high false positives rate found in simulation studies. As a matter of fact, <xref ref-type="bibr" rid="ref51">Pardo and Ferrer (2013)</xref> have proved that, when reliability is estimated through the test&#x2013;retest correlation, both RCI and HA offer acceptable rates of false positives (which does not happen when reliability is estimated through Cronbach&#x2019;s alpha).</p>
<p>Therefore, estimating reliability through the test&#x2013;retest correlation implies not only working with a more realistic SEM, but also working with a value of SEM that has the direct consequence of reducing the false positive rate. But using the test&#x2013;retest correlation to estimate the reliability of a test has a serious drawback: its value depends on the time-interval between first testing and the retest. If that interval is too short, there is a risk of overestimating the true reliability due to the recall of the subjects and their desire to be congruent; if the elapsed time is too long, there is a risk of underestimating true reliability because what is being measured may have changed. Since there is no way of knowing what the ideal time-interval should be between the two measurements, the estimates based on the test&#x2013;retest correlation include an arbitrary component that is difficult to quantify and justify.</p>
<p>Accordingly, in this context, it is felt that the most reasonable measure for bypassing the interval issue would be to resort to alternative ways of estimating reliability. And among the available alternatives, McDonald&#x2019;s omega (&#x03C9;<italic>
<sub>h</sub>
</italic>) coefficient has been postulated as the most widely accepted and optimal measure of internal consistency (<xref ref-type="bibr" rid="ref61">Shevlin et al., 2000</xref>; <xref ref-type="bibr" rid="ref72">Zinbarg et al., 2005</xref>; <xref ref-type="bibr" rid="ref53">Revelle and Zinbarg, 2009</xref>; <xref ref-type="bibr" rid="ref21">Dunn et al., 2014</xref>). And what is more interesting, results obtained by <xref ref-type="bibr" rid="ref53">Revelle and Zinbarg, (2009)</xref> in several groups of data show that &#x03C9;<italic>
<sub>h</sub>
</italic> coefficient takes values systematically smaller than Cronbach&#x2019;s alpha. Of course, this would indicate that &#x03C9;<italic>
<sub>h</sub>
</italic> could be a good option for trying to reduce the rate of false positives associated with RCI and HA when reliability is estimated by Cronbach&#x2019;s alpha.</p>
</sec>
<sec id="sec8">
<title>Objectives</title>
<p>This study has two main aims. First, we intend to make a detailed comparison of the RCI and HA statistics in various scenarios incorporating the use of a new way of estimating reliability (&#x03C9;<italic>
<sub>h</sub>
</italic>). This will allow us to assess the false positive and false negative rates associated with each method in many new scenarios.</p>
<p>Second, since neither RCI nor HA can be calculated with the most widely used computer programs, we put forward to offer to non-expert users can Excel macro to easily calculate these statistics given the conceptual advantage of the HA method, it does not seem reasonable to suggest that the choice for RCI above HA should be based solely on the fact that it is easier to calculate RCI, as suggested by <xref ref-type="bibr" rid="ref57">Ronk et al. (2016)</xref>.</p>
</sec>
</sec>
<sec sec-type="methods" id="sec9">
<title>Methods</title>
<p>To evaluate the false positive and false negative rates of RCI and HA methods, a pre-post design were simulated. In this design, a first measurement is obtained before treatment (<italic>X,</italic> or pre-treatment score) and a second measurement is obtained after treatment (<italic>Y,</italic> or post-treatment score), in the same group of subjects.</p>
<p>The simulated scores were generated assuming no change (null effect size) and different changes (different effect sizes) between pre- and post-measures. The general simulated scenario was a 10 items pre-test measurement (pre-test score was computed by the arithmetic mean of these 10 items), with equal factorial loadings (a <italic>tau</italic>-equivalent scenario in classic test theory), to estimate the reliability (by internal consistency). A post-test score fixed to Pearson&#x2019;s correlation coefficient of 0.8 (<italic>R<sub>XY</sub></italic>&#x2009;=&#x2009;0.80) with the pre-test score to represent common levels of test&#x2013;retest reliability (<xref ref-type="bibr" rid="ref11">Cicchetti, 1994</xref>) (for a detailed comparison of the effects of different test&#x2013;retest correlation sizes, see <xref ref-type="bibr" rid="ref51">Pardo and Ferrer, 2013</xref>; <xref ref-type="bibr" rid="ref26">Ferrer and Pardo, 2019</xref>). To generate the different simulated situations, we used four criteria:</p>
<list list-type="alpha-lower">
<list-item><p><italic>The shape of the pre- and post-treatment score distribution</italic>. Given that moderate and severe deviations from normality are often found in applied contexts (<xref ref-type="bibr" rid="ref48">Micceri, 1989</xref>; <xref ref-type="bibr" rid="ref7">Blanca et al., 2013</xref>), we simulated different values for skewness, ranging from extremely negative to extremely positive, and kurtosis. Using the Pearson distribution system as a reference, we generated five different distributions, four of which represent different degrees of deviation from normality. The degree of deviation from normality was controlled manipulating the value of the skewness (<italic>g</italic><sub>1</sub>) and kurtosis (<italic>g</italic><sub>2</sub>) indexes in the following manner: (<italic>a</italic>) <italic>normal distribution: g</italic><sub>1</sub>&#x2009;=&#x2009;0, <italic>g</italic><sub>2</sub>&#x2009;=&#x2009;0; (<italic>b</italic>) <italic>negative very asymmetric distribution: g</italic><sub>1</sub>&#x2009;=&#x2009;&#x2212;4, <italic>g</italic><sub>2</sub>&#x2009;=&#x2009;18; (<italic>c</italic>) <italic>negative moderately asymmetric distribution: g</italic><sub>1</sub>&#x2009;=&#x2009;&#x2212;2, <italic>g</italic><sub>2</sub>&#x2009;=&#x2009;4; (<italic>d</italic>) <italic>positive moderately asymmetric distribution: g</italic><sub>1</sub>&#x2009;=&#x2009;2, <italic>g</italic><sub>2</sub>&#x2009;=&#x2009;4; (<italic>e</italic>) <italic>positive very asymmetric distribution: g</italic><sub>1</sub>&#x2009;=&#x2009;4, <italic>g</italic><sub>2</sub>&#x2009;=&#x2009;18.</p></list-item>
<list-item><p><italic>The sample size</italic> (<italic>n</italic>): 25, 50, 100. We selected different sample sizes with the intention of representing what is known in the clinical field as small, medium, and large sizes (see, for example, <xref ref-type="bibr" rid="ref15">Crawford and Howell, 1998</xref>).</p></list-item>
<list-item><p><italic>The effect size</italic> (&#x03B4;): 0, 0.2, 0.5, 0.8, 1.1, 1.4, 1.7 and 2 standard deviations of the differences. These values correspond to the systematic increase in the post-scores <italic>Y</italic> expressed in standard deviation of pre-post differences, in the different simulated conditions. Because individual changes have greater variability than average changes, we have chosen effect sizes that range from small (0.2 standard deviations) to very large (2 standard deviations), with increases of 0.3 points. First effect size (0.0) represent a non-change scenario to estimate the false positive rates; the remaining values correspond to the systematic increase on the post scores <italic>Y</italic> in the different simulated conditions to estimate the false negative rates.</p>
<list list-type="simple">
<list-item><p>For a pre-post design, the effect size is usually computed as the standardized pre-post difference (<xref ref-type="bibr" rid="ref12">Cohen, 1988</xref>). However, standardization can be carried out in two different ways: by dividing the mean of the pre-post differences between the standard deviation of pre-test scores (<italic>S<sub>X</sub></italic>), or between the standard deviation of pre-post differences (<italic>S<sub>D</sub></italic>). Following recommendations of some authors (<xref ref-type="bibr" rid="ref12">Cohen, 1988</xref>; <xref ref-type="bibr" rid="ref19">Cumming and Finch, 2001</xref>), we use the standard deviation of the pre-test (<italic>S<sub>X</sub></italic>) as a standardizer since the natural reference for thinking about original scores is the variability in the pre-test scores (<italic>S<sub>X</sub></italic>).</p></list-item>
</list>
</list-item>
<list-item><p><italic>Factorial loadings in the pre-test</italic> (<italic>&#x03BB;</italic>): 0.40, 0.50 and 0.60. These values were selected to represent common values observed in psychometrics factorial analyses (<xref ref-type="bibr" rid="ref52">Peterson, 2000</xref>) and were used to estimate reliability (by internal consistency) using Cronbach&#x2019;s alpha and McDonald&#x2019;s omega coefficients.</p></list-item>
</list>
<p>A total of 5(distributions)&#x2009;&#x00D7;&#x2009;3(sample sizes)&#x2009;&#x00D7;&#x2009;8(effect sizes)&#x2009;&#x00D7;&#x2009;3(factorial loadings)&#x2009;=&#x2009;360 conditions were defined combining these four criteria, and a thousand samples were generated for each of these 120 conditions. Details of the simulation are included in the additional documentation (see <xref ref-type="sec" rid="sec202">supplementary files</xref>).</p>
<p>For data analysis, we made the necessary computations to obtain RCI and HA in each simulated sample. Finally, the performance of each statistic was assessed by applying the corresponding criterion, that is, recording the observed false positive and false negative rates. We considered that a false positive occurred when, with effect size&#x2009;=&#x2009;0, a pre-post difference exceeded the corresponding cut-off point established as the change criterion: &#x2265; 1.65, in absolute value, for both RCI and HA, so make false positives and negatives rates were also comparable. We considered that a false negative occurred when, with effect size &#x003E;0, a pre-post difference did not exceed the corresponding cut-off point. The 1.65 criterion corresponds to the reference point in a normal distribution that should be below the distribution in 95% of the cases, that is, the cut-off point at which one would expect to observe a false-positive rate of approximately 5%.For simulation, and for many of the calculations, we used the MATLAB 20009b program. To compute the mean results from the samples of each condition, we used the IBM SPSS Statistics v. 22 program.</p>
</sec>
<sec sec-type="results" id="sec10">
<title>Results</title>
<p>Since publishing limitations prevent us from including all the results generated by the collection of simulated conditions, the present report only includes percentages of false negatives and false positives.</p>
<p><xref rid="tab1" ref-type="table">Table 1</xref> offers the percentage of false positives (when effect size&#x2009;=&#x2009;0) and false negatives (when effect size &#x003E;0) associated with the RCI statistic. <xref rid="tab2" ref-type="table">Table 2</xref> offers the same percentages for the HA statistic. These percentages were obtained by calculating the number of false positives and false negatives in the 1,000 samples for each condition. Following the liberal criterion of <xref ref-type="bibr" rid="ref9">Bradley (1978)</xref>, percentages of false positives between 2.5 and 7.5% were considered acceptable (and shaded). Following a similar logic, the percentages of false negatives under 25% were considered correct (and shaded).</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>RCI: mean (standard deviation) percentage of false positives and false negatives.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th colspan="3"></th>
<th align="center" valign="top" colspan="8">Effect size (<italic>&#x03B4;</italic>)</th>
</tr>
</thead>
<tbody>
<tr>
<td/>
<td align="center" valign="middle">
<italic>&#x03BB;</italic>
</td>
<td align="center" valign="top"><italic>g</italic><sub>1</sub>, <italic>g</italic><sub>2</sub></td>
<td align="center" valign="bottom">0</td>
<td align="center" valign="bottom">0.2</td>
<td align="center" valign="bottom">0.5</td>
<td align="center" valign="bottom">0.8</td>
<td align="center" valign="bottom">1.1</td>
<td align="center" valign="bottom">1.4</td>
<td align="center" valign="bottom">1.7</td>
<td align="center" valign="bottom">2.0</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="15"><italic>n</italic>&#x2009;=&#x2009;25</td>
<td align="center" valign="middle" rowspan="5">0.4</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">13.9 (0.08)</td>
<td align="center" valign="middle">84.7 (0.08)</td>
<td align="center" valign="middle">78.5 (0.10)</td>
<td align="center" valign="middle">68.1 (0.13)</td>
<td align="center" valign="middle">55.3 (0.15)</td>
<td align="center" valign="middle">41.5 (0.15)</td>
<td align="center" valign="middle">29.5 (0.15)</td>
<td align="center" valign="middle">18.8 (0.13)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">13.1 (0.07)</td>
<td align="center" valign="middle">86.4 (0.07)</td>
<td align="center" valign="middle">82.9 (0.08)</td>
<td align="center" valign="middle">74.9 (0.11)</td>
<td align="center" valign="middle">63.0 (0.18)</td>
<td align="center" valign="middle">46.3 (0.23)</td>
<td align="center" valign="middle">31.1 (0.23)</td>
<td align="center" valign="middle">18.9 (0.20)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">13.5 (0.07)</td>
<td align="center" valign="middle">85.1 (0.08)</td>
<td align="center" valign="middle">77.7 (0.14)</td>
<td align="center" valign="middle">61.1 (0.21)</td>
<td align="center" valign="middle">44.4 (0.24)</td>
<td align="center" valign="middle">31.7 (0.22)</td>
<td align="center" valign="middle">22.8 (0.18)</td>
<td align="center" valign="middle">17.0 (0.15)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">11.4 (0.06)</td>
<td align="center" valign="middle">88.8 (0.06)</td>
<td align="center" valign="middle">85.8 (0.11)</td>
<td align="center" valign="middle">71.4 (0.26)</td>
<td align="center" valign="middle">54.9 (0.32)</td>
<td align="center" valign="middle">40.3 (0.33)</td>
<td align="center" valign="middle">29.0 (0.31)</td>
<td align="center" valign="middle">20.0 (0.27)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">11.0 (0.06)</td>
<td align="center" valign="middle">87.3 (0.08)</td>
<td align="center" valign="middle">77.2 (0.19)</td>
<td align="center" valign="middle">58.7 (0.32)</td>
<td align="center" valign="middle">42.8 (0.35)</td>
<td align="center" valign="middle">31.2 (0.33)</td>
<td align="center" valign="middle">22.5 (0.28)</td>
<td align="center" valign="middle">16.5 (0.24)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.5</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">20.6 (0.09)</td>
<td align="center" valign="middle">77.6 (0.10)</td>
<td align="center" valign="middle">68.9 (0.11)</td>
<td align="center" valign="middle">55.5 (0.14)</td>
<td align="center" valign="middle">40.0 (0.14)</td>
<td align="center" valign="middle">26.3 (0.14)</td>
<td align="center" valign="middle">16.1 (0.12)</td>
<td align="center" valign="middle">8.7 (0.09)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">17.7 (0.08)</td>
<td align="center" valign="middle">82.0 (0.08)</td>
<td align="center" valign="middle">76.8 (0.10)</td>
<td align="center" valign="middle">64.2 (0.17)</td>
<td align="center" valign="middle">45.2 (0.22)</td>
<td align="center" valign="middle">28.0 (0.21)</td>
<td align="center" valign="middle">15.6 (0.16)</td>
<td align="center" valign="middle">7.8 (0.12)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">18.8 (0.09)</td>
<td align="center" valign="middle">78.8 (0.10)</td>
<td align="center" valign="middle">65.7 (0.18)</td>
<td align="center" valign="middle">45.6 (0.22)</td>
<td align="center" valign="middle">30.3 (0.19)</td>
<td align="center" valign="middle">20.7 (0.15)</td>
<td align="center" valign="middle">14.4 (0.12)</td>
<td align="center" valign="middle">10.4 (0.10)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">13.2 (0.07)</td>
<td align="center" valign="middle">86.9 (0.07)</td>
<td align="center" valign="middle">77.9 (0.20)</td>
<td align="center" valign="middle">58.1 (0.32)</td>
<td align="center" valign="middle">41.5 (0.33)</td>
<td align="center" valign="middle">29.2 (0.31)</td>
<td align="center" valign="middle">21.1 (0.28)</td>
<td align="center" valign="middle">14.7 (0.24)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">13.2 (0.07)</td>
<td align="center" valign="middle">84.1 (0.09)</td>
<td align="center" valign="middle">65.9 (0.26)</td>
<td align="center" valign="middle">44.0 (0.34)</td>
<td align="center" valign="middle">30.0 (0.33)</td>
<td align="center" valign="middle">21.5 (0.29)</td>
<td align="center" valign="middle">15.4 (0.24)</td>
<td align="center" valign="middle">11.1 (0.19)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.6</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">30.3 (0.11)</td>
<td align="center" valign="middle">67.4 (0.11)</td>
<td align="center" valign="middle">57.5 (0.12)</td>
<td align="center" valign="middle">42.3 (0.13)</td>
<td align="center" valign="middle">27.2 (0.12)</td>
<td align="center" valign="middle">15.4 (0.10)</td>
<td align="center" valign="middle">7.7 (0.07)</td>
<td align="center" valign="middle">3.5 (0.04)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">23.5 (0.09)</td>
<td align="center" valign="middle">76.1 (0.09)</td>
<td align="center" valign="middle">67.9 (0.14)</td>
<td align="center" valign="middle">47.7 (0.20)</td>
<td align="center" valign="middle">28.6 (0.19)</td>
<td align="center" valign="middle">15.7 (0.15)</td>
<td align="center" valign="middle">7.9 (0.11)</td>
<td align="center" valign="middle">3.3 (0.07)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">24.1 (0.09)</td>
<td align="center" valign="middle">72.3 (0.11)</td>
<td align="center" valign="middle">53.7 (0.20)</td>
<td align="center" valign="middle">33.0 (0.20)</td>
<td align="center" valign="middle">21.6 (0.16)</td>
<td align="center" valign="middle">14.5 (0.12)</td>
<td align="center" valign="middle">10.0 (0.10)</td>
<td align="center" valign="middle">7.0 (0.08)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">16.2 (0.08)</td>
<td align="center" valign="middle">83.6 (0.08)</td>
<td align="center" valign="middle">64.7 (0.28)</td>
<td align="center" valign="middle">42.7 (0.33)</td>
<td align="center" valign="middle">29.4 (0.31)</td>
<td align="center" valign="middle">20.4 (0.27)</td>
<td align="center" valign="middle">13.9 (0.23)</td>
<td align="center" valign="middle">9.4 (0.20)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">15.7 (0.07)</td>
<td align="center" valign="middle">80.4 (0.11)</td>
<td align="center" valign="middle">54.1 (0.31)</td>
<td align="center" valign="middle">34.4 (0.33)</td>
<td align="center" valign="middle">23.3 (0.30)</td>
<td align="center" valign="middle">16.5 (0.26)</td>
<td align="center" valign="middle">12.0 (0.21)</td>
<td align="center" valign="middle">8.8 (0.18)</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="15"><italic>n</italic>&#x2009;=&#x2009;50</td>
<td align="center" valign="middle" rowspan="5">0.4</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">14.2 (0.05)</td>
<td align="center" valign="middle">84.5 (0.06)</td>
<td align="center" valign="middle">78.5 (0.07)</td>
<td align="center" valign="middle">67.8 (0.09)</td>
<td align="center" valign="middle">54.4 (0.10)</td>
<td align="center" valign="middle">40.4 (0.11)</td>
<td align="center" valign="middle">27.7 (0.10)</td>
<td align="center" valign="middle">17.3 (0.09)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">12.3 (0.05)</td>
<td align="center" valign="middle">86.9 (0.05)</td>
<td align="center" valign="middle">83.5 (0.05)</td>
<td align="center" valign="middle">77.4 (0.07)</td>
<td align="center" valign="middle">66.0 (0.13)</td>
<td align="center" valign="middle">48.0 (0.18)</td>
<td align="center" valign="middle">29.0 (0.18)</td>
<td align="center" valign="middle">15.6 (0.14)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">12.3 (0.05)</td>
<td align="center" valign="middle">86.8 (0.05)</td>
<td align="center" valign="middle">80.4 (0.09)</td>
<td align="center" valign="middle">64.5 (0.16)</td>
<td align="center" valign="middle">44.3 (0.19)</td>
<td align="center" valign="middle">29.7 (0.15)</td>
<td align="center" valign="middle">20.6 (0.11)</td>
<td align="center" valign="middle">14.9 (0.08)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">10.1 (0.04)</td>
<td align="center" valign="middle">89.9 (0.04)</td>
<td align="center" valign="middle">88.4 (0.05)</td>
<td align="center" valign="middle">79.6 (0.17)</td>
<td align="center" valign="middle">60.5 (0.28)</td>
<td align="center" valign="middle">40.9 (0.30)</td>
<td align="center" valign="middle">26.4 (0.27)</td>
<td align="center" valign="middle">15.9 (0.21)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">10.3 (0.05)</td>
<td align="center" valign="middle">88.9 (0.05)</td>
<td align="center" valign="middle">82.7 (0.12)</td>
<td align="center" valign="middle">64.6 (0.27)</td>
<td align="center" valign="middle">44.0 (0.32)</td>
<td align="center" valign="middle">29.1 (0.29)</td>
<td align="center" valign="middle">19.3 (0.23)</td>
<td align="center" valign="middle">12.5 (0.17)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.5</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">21.1 (0.06)</td>
<td align="center" valign="middle">77.1 (0.07)</td>
<td align="center" valign="middle">68.7 (0.08)</td>
<td align="center" valign="middle">55.2 (0.09)</td>
<td align="center" valign="middle">40.1 (0.10)</td>
<td align="center" valign="middle">26.2 (0.09)</td>
<td align="center" valign="middle">15.3 (0.08)</td>
<td align="center" valign="middle">7.8 (0.06)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">17.5 (0.06)</td>
<td align="center" valign="middle">82.1 (0.06)</td>
<td align="center" valign="middle">77.7 (0.07)</td>
<td align="center" valign="middle">66.2 (0.12)</td>
<td align="center" valign="middle">45.2 (0.17)</td>
<td align="center" valign="middle">25.4 (0.16)</td>
<td align="center" valign="middle">12.5 (0.11)</td>
<td align="center" valign="middle">8.8 (0.07)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">17.5 (0.06)</td>
<td align="center" valign="middle">80.4 (0.07)</td>
<td align="center" valign="middle">68.1 (0.13)</td>
<td align="center" valign="middle">48.7 (0.17)</td>
<td align="center" valign="middle">28.5 (0.14)</td>
<td align="center" valign="middle">18.8 (0.10)</td>
<td align="center" valign="middle">13.4 (0.07)</td>
<td align="center" valign="middle">9.4 (0.06)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">12.6 (0.05)</td>
<td align="center" valign="middle">87.4 (0.05)</td>
<td align="center" valign="middle">82.7 (0.11)</td>
<td align="center" valign="middle">61.5 (0.28)</td>
<td align="center" valign="middle">39.6 (0.31)</td>
<td align="center" valign="middle">24.9 (8.27)</td>
<td align="center" valign="middle">15.3 (0.22)</td>
<td align="center" valign="middle">9.2 (0.17)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">12.4 (0.05)</td>
<td align="center" valign="middle">85.9 (0.06)</td>
<td align="center" valign="middle">72.3 (0.19)</td>
<td align="center" valign="middle">46.1 (0.31)</td>
<td align="center" valign="middle">27.8 (0.28)</td>
<td align="center" valign="middle">17.4 (0.22)</td>
<td align="center" valign="middle">11.4 (0.17)</td>
<td align="center" valign="middle">7.8 (0.13)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.6</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">29.9 (0.07)</td>
<td align="center" valign="middle">67.9 (0.07)</td>
<td align="center" valign="middle">57.6 (0.09)</td>
<td align="center" valign="middle">42.3 (0.09)</td>
<td align="center" valign="middle">27.0 (0.09)</td>
<td align="center" valign="middle">15.0 (0.7)</td>
<td align="center" valign="middle">7.1 (0.05)</td>
<td align="center" valign="middle">3.0 (0.03)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="top">23.3 (0.07)</td>
<td align="center" valign="top">76.7 (0.07)</td>
<td align="center" valign="top">69.1 (0.09)</td>
<td align="center" valign="top">47.0 (0.16)</td>
<td align="center" valign="top">25.3 (0.14)</td>
<td align="center" valign="top">12.3 (0.09)</td>
<td align="center" valign="top">8.5 (0.04)</td>
<td align="center" valign="top">2.1 (0.02)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">23.3 (0.07)</td>
<td align="center" valign="top">73.0 (0.08)</td>
<td align="center" valign="top">53.0 (0.15)</td>
<td align="center" valign="top">30.0 (0.14)</td>
<td align="center" valign="top">18.8 (0.10)</td>
<td align="center" valign="top">12.4 (0.07)</td>
<td align="center" valign="top">8.4 (0.06)</td>
<td align="center" valign="top">5.7 (0.05)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">14.8 (0.05)</td>
<td align="center" valign="top">84.9 (0.05)</td>
<td align="center" valign="top">71.0 (0.21)</td>
<td align="center" valign="top">41.7 (0.30)</td>
<td align="center" valign="top">23.5 (0.26)</td>
<td align="center" valign="top">13.8 (0.19)</td>
<td align="center" valign="top">08.1 (0.14)</td>
<td align="center" valign="top">4.8 (0.10)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;4, 18</td>
<td align="center" valign="top">14.8 (0.05)</td>
<td align="center" valign="top">82.0 (0.07)</td>
<td align="center" valign="top">57.5 (0.26)</td>
<td align="center" valign="top">30.6 (0.29)</td>
<td align="center" valign="top">17.5 (0.22)</td>
<td align="center" valign="top">10.7 (0.17)</td>
<td align="center" valign="top">7.2 (0.13)</td>
<td align="center" valign="top">5.2 (0.10)</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="15"><italic>n</italic>&#x2009;=&#x2009;100</td>
<td align="center" valign="top" rowspan="5">0.4</td>
<td align="center" valign="top">0, 0</td>
<td align="center" valign="top">14.0 (0.4)</td>
<td align="center" valign="top">84.9 (0.04)</td>
<td align="center" valign="top">78.8 (0.05)</td>
<td align="center" valign="top">68.2 (0.06)</td>
<td align="center" valign="top">54.8 (0.07)</td>
<td align="center" valign="top">40.7 (0.08)</td>
<td align="center" valign="top">27.7 (0.07)</td>
<td align="center" valign="top">17.1 (0.06)</td>
</tr>
<tr>
<td align="center" valign="top">2, 4</td>
<td align="center" valign="top">12.1 (0.03)</td>
<td align="center" valign="top">87.1 (0.03)</td>
<td align="center" valign="top">83.9 (0.04)</td>
<td align="center" valign="top">78.0 (0.05)</td>
<td align="center" valign="top">66.8 (0.09)</td>
<td align="center" valign="top">46.9 (0.14)</td>
<td align="center" valign="top">25.7 (0.13)</td>
<td align="center" valign="top">11.9 (0.08)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">12.1 (0.03)</td>
<td align="center" valign="top">87.2 (0.04)</td>
<td align="center" valign="top">81.5 (0.06)</td>
<td align="center" valign="top">65.9 (0.12)</td>
<td align="center" valign="top">43.8 (0.14)</td>
<td align="center" valign="top">28.2 (0.10)</td>
<td align="center" valign="top">19.6 (0.07)</td>
<td align="center" valign="top">14.5 (0.05)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">9.2 (0.03)</td>
<td align="center" valign="top">90.8 (0.03)</td>
<td align="center" valign="top">89.6 (0.03)</td>
<td align="center" valign="top">84.8 (0.10)</td>
<td align="center" valign="top">66.7 (0.23)</td>
<td align="center" valign="top">42.2 (0.27)</td>
<td align="center" valign="top">23.6 (0.23)</td>
<td align="center" valign="top">12.3 (0.15)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;4, 18</td>
<td align="center" valign="top">9.3 (0.03)</td>
<td align="center" valign="top">90.1 (0.03)</td>
<td align="center" valign="top">85.5 (0.07)</td>
<td align="center" valign="top">69.0 (0.20)</td>
<td align="center" valign="top">42.4 (0.27)</td>
<td align="center" valign="top">22.7 (0.22)</td>
<td align="center" valign="top">12.4 (0.14)</td>
<td align="center" valign="top">8.3 (0.14)</td>
</tr>
<tr>
<td align="center" valign="top" rowspan="5">0.5</td>
<td align="center" valign="top">0, 0</td>
<td align="center" valign="top">21.1 (0.05)</td>
<td align="center" valign="top">77.1 (0.05)</td>
<td align="center" valign="top">68.4 (0.06)</td>
<td align="center" valign="top">54.8 (0.07)</td>
<td align="center" valign="top">39.3 (0.07)</td>
<td align="center" valign="top">25.0 (0.06)</td>
<td align="center" valign="top">14.0 (0.05)</td>
<td align="center" valign="top">6.9 (0.03)</td>
</tr>
<tr>
<td align="center" valign="top">2, 4</td>
<td align="center" valign="top">17.4 (0.04)</td>
<td align="center" valign="top">82.2 (0.04)</td>
<td align="center" valign="top">77.8 (0.05)</td>
<td align="center" valign="top">66.4 (0.09)</td>
<td align="center" valign="top">43.0 (0.14)</td>
<td align="center" valign="top">21.6 (0.11)</td>
<td align="center" valign="top">9.9 (0.06)</td>
<td align="center" valign="top">4.2 (0.03)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">17.2 (0.04)</td>
<td align="center" valign="top">80.8 (0.05)</td>
<td align="center" valign="top">69.3 (0.09)</td>
<td align="center" valign="top">45.5 (0.13)</td>
<td align="center" valign="top">27.2 (0.10)</td>
<td align="center" valign="top">18.2 (0.06)</td>
<td align="center" valign="top">12.9 (0.05)</td>
<td align="center" valign="top">9.1 (0.04)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">12.0 (0.03)</td>
<td align="center" valign="top">88.0 (0.03)</td>
<td align="center" valign="top">85.5 (0.06)</td>
<td align="center" valign="top">65.4 (0.23)</td>
<td align="center" valign="top">37.3 (0.26)</td>
<td align="center" valign="top">19.0 (0.20)</td>
<td align="center" valign="top">8.5 (0.12)</td>
<td align="center" valign="top">4.8 (0.06)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;4, 18</td>
<td align="center" valign="top">11.8 (0.03)</td>
<td align="center" valign="top">86.8 (0.04)</td>
<td align="center" valign="top">76.3 (0.12)</td>
<td align="center" valign="top">46.7 (0.26)</td>
<td align="center" valign="top">22.5 (0.22)</td>
<td align="center" valign="top">12.0 (0.14)</td>
<td align="center" valign="top">7.6 (0.08)</td>
<td align="center" valign="top">5.6 (0.05)</td>
</tr>
<tr>
<td align="center" valign="top" rowspan="5">0.6</td>
<td align="center" valign="top">0, 0</td>
<td align="center" valign="top">29.9 (0.05)</td>
<td align="center" valign="top">67.9 (0.5)</td>
<td align="center" valign="top">57.6 (0.06)</td>
<td align="center" valign="top">42.2 (0.06)</td>
<td align="center" valign="top">26.6 (0.06)</td>
<td align="center" valign="top">14.3 (0.05)</td>
<td align="center" valign="top">6.6 (0.03)</td>
<td align="center" valign="top">2.7 (0.02)</td>
</tr>
<tr>
<td align="center" valign="top">2, 4</td>
<td align="center" valign="top">23.1 (0.04)</td>
<td align="center" valign="top">76.6 (0.04)</td>
<td align="center" valign="top">69.5 (0.06)</td>
<td align="center" valign="top">46.4 (0.12)</td>
<td align="center" valign="top">22.2 (0.09)</td>
<td align="center" valign="top">10.2 (0.04)</td>
<td align="center" valign="top">4.5 (0.02)</td>
<td align="center" valign="top">1.9 (0.01)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">23.0 (0.04)</td>
<td align="center" valign="top">73.6 (0.05)</td>
<td align="center" valign="top">54.4 (0.11)</td>
<td align="center" valign="top">29.0 (0.09)</td>
<td align="center" valign="top">17.7 (0.06)</td>
<td align="center" valign="top">11.9 (0.04)</td>
<td align="center" valign="top">8.0 (0.03)</td>
<td align="center" valign="top">5.3 (0.03)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">14.1 (0.04)</td>
<td align="center" valign="top">85.6 (0.04)</td>
<td align="center" valign="top">76.3 (0.15)</td>
<td align="center" valign="top">41.8 (0.27)</td>
<td align="center" valign="top">19.6 (0.20)</td>
<td align="center" valign="top">9.9 (0.13)</td>
<td align="center" valign="top">5.4 (0.08)</td>
<td align="center" valign="top">3.1 (0.04)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;4, 18</td>
<td align="center" valign="top">14.3 (0.04)</td>
<td align="center" valign="top">83.3 (0.05)</td>
<td align="center" valign="top">62.0 (0.19)</td>
<td align="center" valign="top">26.5 (0.23)</td>
<td align="center" valign="top">12.2 (0.14)</td>
<td align="center" valign="top">7.3 (0.09)</td>
<td align="center" valign="top">5.1 (0.05)</td>
<td align="center" valign="top">3.9 (0.03)</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap position="float" id="tab2">
<label>Table 2</label>
<caption>
<p>HA: mean (standard deviation) percentage of false positives and false negatives.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th colspan="3"></th>
<th align="center" valign="top" colspan="8">Effect size (<italic>&#x03B4;</italic>)</th>
</tr>
</thead>
<tbody>
<tr>
<td/>
<td align="center" valign="middle">
<italic>&#x03BB;</italic>
</td>
<td align="center" valign="top"><italic>g</italic><sub>1</sub>, <italic>g</italic><sub>2</sub></td>
<td align="center" valign="bottom">0</td>
<td align="center" valign="bottom">0.2</td>
<td align="center" valign="bottom">0.5</td>
<td align="center" valign="bottom">0.8</td>
<td align="center" valign="bottom">1.1</td>
<td align="center" valign="bottom">1.4</td>
<td align="center" valign="bottom">1.7</td>
<td align="center" valign="bottom">2.0</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="15"><italic>n</italic>&#x2009;=&#x2009;25</td>
<td align="center" valign="middle" rowspan="5">0.4</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">5.3 (0.15)</td>
<td align="center" valign="middle">89.5 (0.21)</td>
<td align="center" valign="middle">69.9 (0.34)</td>
<td align="center" valign="middle">50.6 (0.39)</td>
<td align="center" valign="middle">39.1 (0.43)</td>
<td align="center" valign="middle">34.2 (0.45)</td>
<td align="center" valign="middle">32.3 (0.45)</td>
<td align="center" valign="middle">31.6 (0.46)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">6.7 (0.14)</td>
<td align="center" valign="middle">89.3 (0.18)</td>
<td align="center" valign="middle">75.5 (0.29)</td>
<td align="center" valign="middle">55.0 (0.40)</td>
<td align="center" valign="middle">40.4 (0.44)</td>
<td align="center" valign="middle">36.0 (0.46)</td>
<td align="center" valign="middle">34.8 (0.47)</td>
<td align="center" valign="middle">34.6 (0.47)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">5.9 (0.10)</td>
<td align="center" valign="middle">91.0 (0.18)</td>
<td align="center" valign="middle">71.5 (0.32)</td>
<td align="center" valign="middle">48.5 (0.39)</td>
<td align="center" valign="middle">39.8 (0.42)</td>
<td align="center" valign="middle">36.3 (0.44)</td>
<td align="center" valign="middle">34.4 (0.45)</td>
<td align="center" valign="middle">33.3 (0.46)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">7.5 (0.10)</td>
<td align="center" valign="middle">90.9 (0.15)</td>
<td align="center" valign="middle">74.7 (0.27)</td>
<td align="center" valign="middle">51.9 (0.40)</td>
<td align="center" valign="middle">40.6 (0.45)</td>
<td align="center" valign="middle">38.1 (0.47)</td>
<td align="center" valign="middle">37.1 (0.47)</td>
<td align="center" valign="middle">36.8 (0.48)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">7.5 (0.11)</td>
<td align="center" valign="middle">90.2 (0.16)</td>
<td align="center" valign="middle">72.8 (0.32)</td>
<td align="center" valign="middle">47.8 (0.42)</td>
<td align="center" valign="middle">41.2 (0.45)</td>
<td align="center" valign="middle">40.0 (0.45)</td>
<td align="center" valign="middle">40.0 (0.46)</td>
<td align="center" valign="middle">38.5 (0.46)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.5</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">8.9 (0.13)</td>
<td align="center" valign="middle">85.1 (0.18)</td>
<td align="center" valign="middle">60.8 (0.289)</td>
<td align="center" valign="middle">36.6 (0.30)</td>
<td align="center" valign="middle">21.6 (0.30)</td>
<td align="center" valign="middle">14.7 (0.30)</td>
<td align="center" valign="middle">11.8 (0.30)</td>
<td align="center" valign="middle">10.8 (0.30)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="top">11.1 (0.14)</td>
<td align="center" valign="top">85.6 (0.17)</td>
<td align="center" valign="top">68.9 (0.28)</td>
<td align="center" valign="top">39.8 (0.33)</td>
<td align="center" valign="top">21.8 (0.33)</td>
<td align="center" valign="top">16.2 (0.34)</td>
<td align="center" valign="top">14.6 (0.34)</td>
<td align="center" valign="top">14.3 (0.34)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">11.2 (0.13)</td>
<td align="center" valign="middle">84.2 (0.18)</td>
<td align="center" valign="middle">59.5 (0.30)</td>
<td align="center" valign="middle">33.9 (0.31)</td>
<td align="center" valign="middle">24.0 (0.31)</td>
<td align="center" valign="middle">19.3 (0.32)</td>
<td align="center" valign="middle">16.9 (0.33)</td>
<td align="center" valign="middle">15.4 (0.33)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">10.1 (0.13)</td>
<td align="center" valign="middle">88.5 (0.16)</td>
<td align="center" valign="middle">73.6 (0.30)</td>
<td align="center" valign="middle">40.3 (0.38)</td>
<td align="center" valign="middle">30.3 (0.41)</td>
<td align="center" valign="middle">27.7 (0.42)</td>
<td align="center" valign="middle">26.4 (0.43)</td>
<td align="center" valign="middle">26.0 (0.43)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">9.9 (0.11)</td>
<td align="center" valign="middle">86.6 (0.17)</td>
<td align="center" valign="middle">61.6 (0.34)</td>
<td align="center" valign="middle">36.0 (0.39)</td>
<td align="center" valign="middle">29.5 (0.41)</td>
<td align="center" valign="middle">28.1 (0.41)</td>
<td align="center" valign="middle">27.3 (0.42)</td>
<td align="center" valign="middle">26.8 (0.42)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.6</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">17.1 (0.11)</td>
<td align="center" valign="middle">77.8 (0.14)</td>
<td align="center" valign="middle">54.5 (0.19)</td>
<td align="center" valign="middle">29.5 (0.18)</td>
<td align="center" valign="middle">13.4 (0.14)</td>
<td align="center" valign="middle">5.4 (0.11)</td>
<td align="center" valign="middle">2.4 (0.09)</td>
<td align="center" valign="middle">1.3 (0.09)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">16.2 (0.12)</td>
<td align="center" valign="middle">81.6 (0.14)</td>
<td align="center" valign="middle">64.9 (0.22)</td>
<td align="center" valign="middle">30.5 (0.24)</td>
<td align="center" valign="middle">13.8 (0.22)</td>
<td align="center" valign="middle">8.0 (0.22)</td>
<td align="center" valign="middle">5.9 (0.22)</td>
<td align="center" valign="middle">5.5 (0.22)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">16.3 (0.12)</td>
<td align="center" valign="middle">78.0 (0.17)</td>
<td align="center" valign="middle">50.1 (0.27)</td>
<td align="center" valign="middle">25.1 (0.23)</td>
<td align="center" valign="middle">16.0 (0.21)</td>
<td align="center" valign="middle">11.2 (0.21)</td>
<td align="center" valign="middle">8.4 (0.21)</td>
<td align="center" valign="middle">6.7 (0.21)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">12.9 (0.10)</td>
<td align="center" valign="middle">86.2 (0.12)</td>
<td align="center" valign="middle">62.5 (0.31)</td>
<td align="center" valign="middle">28.3 (0.32)</td>
<td align="center" valign="middle">19.9 (0.33)</td>
<td align="center" valign="middle">17.1 (0.34)</td>
<td align="center" valign="middle">15.5 (0.35)</td>
<td align="center" valign="middle">14.7 (0.35)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">12.5 (0.11)</td>
<td align="center" valign="middle">82.7 (0.16)</td>
<td align="center" valign="middle">48.0 (0.36)</td>
<td align="center" valign="middle">25.2 (0.34)</td>
<td align="center" valign="middle">18.8 (0.33)</td>
<td align="center" valign="middle">17.4 (0.33)</td>
<td align="center" valign="middle">16.6 (0.34)</td>
<td align="center" valign="middle">16.0 (0.34)</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="15"><italic>n</italic>&#x2009;=&#x2009;50</td>
<td align="center" valign="middle" rowspan="5">0.4</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">3.8 (0.14)</td>
<td align="center" valign="middle">90.6 (0.21)</td>
<td align="center" valign="middle">63.4 (0.34)</td>
<td align="center" valign="middle">37.2 (0.36)</td>
<td align="center" valign="middle">24.9 (0.37)</td>
<td align="center" valign="middle">20.7 (0.38)</td>
<td align="center" valign="middle">19.5 (0.39)</td>
<td align="center" valign="middle">19.2 (0.39)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">4.5 (0.09)</td>
<td align="center" valign="middle">90.8 (0.15)</td>
<td align="center" valign="middle">72.1 (0.30)</td>
<td align="center" valign="middle">44.7 (0.39)</td>
<td align="center" valign="middle">30.7 (0.42)</td>
<td align="center" valign="middle">27.8 (0.44)</td>
<td align="center" valign="middle">27.2 (0.44)</td>
<td align="center" valign="middle">27.1 (0.44)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">4.3 (0.10)</td>
<td align="center" valign="middle">93.1 (0.17)</td>
<td align="center" valign="middle">67.3 (0.34)</td>
<td align="center" valign="middle">40.5 (0.38)</td>
<td align="center" valign="middle">32.2 (0.41)</td>
<td align="center" valign="middle">29.1 (0.42)</td>
<td align="center" valign="middle">27.5 (0.43)</td>
<td align="center" valign="middle">26.7 (0.43)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">5.0 (0.07)</td>
<td align="center" valign="middle">92.5 (0.14)</td>
<td align="center" valign="middle">81.5 (0.28)</td>
<td align="center" valign="middle">50.8 (0.40)</td>
<td align="center" valign="middle">38.7 (0.45)</td>
<td align="center" valign="middle">36.4 (0.47)</td>
<td align="center" valign="middle">35.6 (0.47)</td>
<td align="center" valign="middle">35.4 (0.47)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">6.0 (0.10)</td>
<td align="center" valign="middle">92.6 (0.13)</td>
<td align="center" valign="middle">72.8 (0.31)</td>
<td align="center" valign="middle">58.0 (0.41)</td>
<td align="center" valign="middle">35.2 (0.43)</td>
<td align="center" valign="middle">33.8 (0.44)</td>
<td align="center" valign="middle">33.0. (44)</td>
<td align="center" valign="middle">32.4 (0.45)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.5</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">6.5 (0.07)</td>
<td align="center" valign="middle">88.2 (0.12)</td>
<td align="center" valign="middle">62.0 (0.21)</td>
<td align="center" valign="middle">31.5 (0.20)</td>
<td align="center" valign="middle">12.5 (0.16)</td>
<td align="center" valign="middle">4.9 (0.13)</td>
<td align="center" valign="middle">2.5 (0.13)</td>
<td align="center" valign="middle">2.0 (0.13)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="middle">8.3 (0.09)</td>
<td align="center" valign="middle">88.3 (0.12)</td>
<td align="center" valign="middle">71.5 (0.22)</td>
<td align="center" valign="middle">35.6 (0.27)</td>
<td align="center" valign="middle">14.2 (0.25)</td>
<td align="center" valign="middle">9.0 (0.25)</td>
<td align="center" valign="middle">7.6 (0.25)</td>
<td align="center" valign="middle">7.4 (0.26)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;2, 4</td>
<td align="center" valign="middle">8.8 (0.09)</td>
<td align="center" valign="middle">88.3 (0.13)</td>
<td align="center" valign="middle">61.5 (0.26)</td>
<td align="center" valign="middle">29.3 (0.25)</td>
<td align="center" valign="middle">18.2 (0.25)</td>
<td align="center" valign="middle">13.2 (0.26)</td>
<td align="center" valign="middle">10.6 (0.26)</td>
<td align="center" valign="middle">9.2 (0.26)</td>
</tr>
<tr>
<td align="center" valign="middle">4, 18</td>
<td align="center" valign="middle">8.1 (0.07)</td>
<td align="center" valign="middle">89.8 (0.12)</td>
<td align="center" valign="middle">76.3 (0.26)</td>
<td align="center" valign="middle">32.6 (0.33)</td>
<td align="center" valign="middle">20.5 (0.35)</td>
<td align="center" valign="middle">18.0 (0.36)</td>
<td align="center" valign="middle">17.0 (0.36)</td>
<td align="center" valign="middle">16.5 (0.36)</td>
</tr>
<tr>
<td align="center" valign="middle">&#x2212;4, 18</td>
<td align="center" valign="middle">8.4 (0.10)</td>
<td align="center" valign="middle">89.1 (0.14)</td>
<td align="center" valign="middle">63.0 (0.31)</td>
<td align="center" valign="middle">27.7 (0.34)</td>
<td align="center" valign="middle">21.1 (0.34)</td>
<td align="center" valign="middle">19.5 (0.35)</td>
<td align="center" valign="middle">18.5 (0.35)</td>
<td align="center" valign="middle">17.7 (0.36)</td>
</tr>
<tr>
<td align="center" valign="middle" rowspan="5">0.6</td>
<td align="center" valign="middle">0, 0</td>
<td align="center" valign="middle">16.0 (0.08)</td>
<td align="center" valign="middle">79.7 (0.09)</td>
<td align="center" valign="middle">57.9 (0.13)</td>
<td align="center" valign="middle">31.1 (0.13)</td>
<td align="center" valign="middle">12.7 (0.09)</td>
<td align="center" valign="middle">4.0 (0.05)</td>
<td align="center" valign="middle">1.2 (0.03)</td>
<td align="center" valign="middle">0.4 (0.03)</td>
</tr>
<tr>
<td align="center" valign="middle">2, 4</td>
<td align="center" valign="top">14.4 (0.08)</td>
<td align="center" valign="top">83.7 (0.08)</td>
<td align="center" valign="top">68.2 (0.14)</td>
<td align="center" valign="top">28.6 (0.17)</td>
<td align="center" valign="top">9.7 (0.12)</td>
<td align="center" valign="top">4.1 (0.12)</td>
<td align="center" valign="top">2.1 (0.12)</td>
<td align="center" valign="top">1.7 (0.12)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">14.5 (0.07)</td>
<td align="center" valign="top">80.9 (0.11)</td>
<td align="center" valign="top">50.1 (0.20)</td>
<td align="center" valign="top">21.9 (0.14)</td>
<td align="center" valign="top">12.0 (0.12)</td>
<td align="center" valign="top">7.3 (0.11)</td>
<td align="center" valign="top">4.4 (0.11)</td>
<td align="center" valign="top">2.9 (0.11)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">11.0 (0.07)</td>
<td align="center" valign="top">87.3 (0.11)</td>
<td align="center" valign="top">67.8 (0.26)</td>
<td align="center" valign="top">20.7 (0.23)</td>
<td align="center" valign="top">10.8 (0.22)</td>
<td align="center" valign="top">8.5 (0.23)</td>
<td align="center" valign="top">7.2 (0.23)</td>
<td align="center" valign="top">6.5 (0.23)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;4, 18</td>
<td align="center" valign="top">11.2 (0.08)</td>
<td align="center" valign="top">85.8 (0.12)</td>
<td align="center" valign="top">52.3 (0.31)</td>
<td align="center" valign="top">20.0 (0.27)</td>
<td align="center" valign="top">13.7 (0.26)</td>
<td align="center" valign="top">12.0 (0.26)</td>
<td align="center" valign="top">11.0 (0.27)</td>
<td align="center" valign="top">10.3 (0.27)</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="15"><italic>n</italic>&#x2009;=&#x2009;100</td>
<td align="center" valign="top" rowspan="5">0.4</td>
<td align="center" valign="top">0, 0</td>
<td align="center" valign="top">2.0 (0.10)</td>
<td align="center" valign="top">91.7 (0.20)</td>
<td align="center" valign="top">56.3 (0.32)</td>
<td align="center" valign="top">24.8 (0.30)</td>
<td align="center" valign="top">13.3 (0.29)</td>
<td align="center" valign="top">10.7 (0.30)</td>
<td align="center" valign="top">10.2 (0.30)</td>
<td align="center" valign="top">10.2 (0.30)</td>
</tr>
<tr>
<td align="center" valign="top">2, 4</td>
<td align="center" valign="top">3.4 (0.09)</td>
<td align="center" valign="top">91.7 (0.15)</td>
<td align="center" valign="top">67.5 (0.31)</td>
<td align="center" valign="top">32.2 (0.36)</td>
<td align="center" valign="top">18.8 (0.36)</td>
<td align="center" valign="top">16.9 (0.37)</td>
<td align="center" valign="top">16.6 (0.37)</td>
<td align="center" valign="top">16.5 (0.37)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">3.0 (0.08)</td>
<td align="center" valign="top">92.8 (0.21)</td>
<td align="center" valign="top">64.8 (0.33)</td>
<td align="center" valign="top">31.9 (0.35)</td>
<td align="center" valign="top">23.6 (0.37)</td>
<td align="center" valign="top">20.7 (0.38)</td>
<td align="center" valign="top">19.6 (0.38)</td>
<td align="center" valign="top">19.2 (0.38)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">3.8 (0.08)</td>
<td align="center" valign="top">92.9 (0.15)</td>
<td align="center" valign="top">77.2 (0.32)</td>
<td align="center" valign="top">40.3 (0.40)</td>
<td align="center" valign="top">30.1 (0.43)</td>
<td align="center" valign="top">28.7 (0.44)</td>
<td align="center" valign="top">28.3 (0.44)</td>
<td align="center" valign="top">28.2 (0.44)</td>
</tr>
<tr>
<td align="center" valign="top">-4, 18</td>
<td align="center" valign="top">3.8 (0.06)</td>
<td align="center" valign="top">92.7 (0.16)</td>
<td align="center" valign="top">71.3 (0.33)</td>
<td align="center" valign="top">35.9 (0.40)</td>
<td align="center" valign="top">30.6 (0.42)</td>
<td align="center" valign="top">29.1 (0.43)</td>
<td align="center" valign="top">28.4 (0.43)</td>
<td align="center" valign="top">27.9 (0.43)</td>
</tr>
<tr>
<td align="center" valign="top" rowspan="5">0.5</td>
<td align="center" valign="top">0, 0</td>
<td align="center" valign="top">5.3 (0.04)</td>
<td align="center" valign="top">90.3 (0.06)</td>
<td align="center" valign="top">63.7 (0.15)</td>
<td align="center" valign="top">29.9 (0.15)</td>
<td align="center" valign="top">9.8 (0.08)</td>
<td align="center" valign="top">2.5 (0.04)</td>
<td align="center" valign="top">0.6 (0.03)</td>
<td align="center" valign="top">0.2 (0.03)</td>
</tr>
<tr>
<td align="center" valign="top">2, 4</td>
<td align="center" valign="top">7.5 (0.07)</td>
<td align="center" valign="top">89.5 (0.07)</td>
<td align="center" valign="top">72.5 (0.17)</td>
<td align="center" valign="top">30.5 (0.20)</td>
<td align="center" valign="top">7.9 (0.13)</td>
<td align="center" valign="top">3.1 (0.13)</td>
<td align="center" valign="top">2.0 (0.13)</td>
<td align="center" valign="top">1.9 (0.13)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">7.2 (0.06)</td>
<td align="center" valign="top">90.5 (0.11)</td>
<td align="center" valign="top">61.6 (0.21)</td>
<td align="center" valign="top">22.6 (0.14)</td>
<td align="center" valign="top">11.0 (0.12)</td>
<td align="center" valign="top">6.0 (0.11)</td>
<td align="center" valign="top">3.4 (0.11)</td>
<td align="center" valign="top">2.1 (0.11)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">7.2 (0.06)</td>
<td align="center" valign="top">91.0 (0.11)</td>
<td align="center" valign="top">80.5 (0.21)</td>
<td align="center" valign="top">27.0 (0.25)</td>
<td align="center" valign="top">12.6 (0.26)</td>
<td align="center" valign="top">10.2 (0.27)</td>
<td align="center" valign="top">9.1 (0.27)</td>
<td align="center" valign="top">8.7 (0.27)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;4, 18</td>
<td align="center" valign="top">7.0 (0.6)</td>
<td align="center" valign="top">90.2 (0.13)</td>
<td align="center" valign="top">64.7 (0.28)</td>
<td align="center" valign="top">20.8 (0.27)</td>
<td align="center" valign="top">14.7 (0.28)</td>
<td align="center" valign="top">12.9 (0.28)</td>
<td align="center" valign="top">11.9 (0.28)</td>
<td align="center" valign="top">11.1 (0.29)</td>
</tr>
<tr>
<td align="center" valign="top" rowspan="5">0.6</td>
<td align="center" valign="top">0, 0</td>
<td align="center" valign="top">15.1 (0.05)</td>
<td align="center" valign="top">80.4 (0.06)</td>
<td align="center" valign="top">59.2 (0.08)</td>
<td align="center" valign="top">31.9 (0.09)</td>
<td align="center" valign="top">12.5 (0.06)</td>
<td align="center" valign="top">3.7 (0.03)</td>
<td align="center" valign="top">0.9 (0.01)</td>
<td align="center" valign="top">0.2 (0.00)</td>
</tr>
<tr>
<td align="center" valign="top">2, 4</td>
<td align="center" valign="top">13.9 (0.05)</td>
<td align="center" valign="top">84.1 (0.05)</td>
<td align="center" valign="top">70.4 (0.08)</td>
<td align="center" valign="top">27.8 (0.11)</td>
<td align="center" valign="top">8.1 (0.04)</td>
<td align="center" valign="top">2.4 (0.01)</td>
<td align="center" valign="top">0.5 (0.00)</td>
<td align="center" valign="top">0.1 (0.00)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;2, 4</td>
<td align="center" valign="top">13.9 (0.05)</td>
<td align="center" valign="top">83.0 (0.06)</td>
<td align="center" valign="top">54.2 (0.14)</td>
<td align="center" valign="top">21.4 (0.07)</td>
<td align="center" valign="top">11.1 (0.05)</td>
<td align="center" valign="top">5.8 (0.03)</td>
<td align="center" valign="top">2.9 (0.02)</td>
<td align="center" valign="top">1.4 (0.01)</td>
</tr>
<tr>
<td align="center" valign="top">4, 18</td>
<td align="center" valign="top">10.1 (0.04)</td>
<td align="center" valign="top">89.0 (0.06)</td>
<td align="center" valign="top">74.9 (0.19)</td>
<td align="center" valign="top">18.4 (0.17)</td>
<td align="center" valign="top">7.9 (0.16)</td>
<td align="center" valign="top">5.5 (0.16)</td>
<td align="center" valign="top">4.2 (0.17)</td>
<td align="center" valign="top">3.5 (0.17)</td>
</tr>
<tr>
<td align="center" valign="top">&#x2212;4, 18</td>
<td align="center" valign="top">10.3 (0.04)</td>
<td align="center" valign="top">87.7 (0.07)</td>
<td align="center" valign="top">56.3 (0.24)</td>
<td align="center" valign="top">12.5 (0.14)</td>
<td align="center" valign="top">7.1 (0.12)</td>
<td align="center" valign="top">5.4 (0.12)</td>
<td align="center" valign="top">4.3 (0.12)</td>
<td align="center" valign="top">3.5 (0.12)</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Information regarding the accuracy of the performed simulation, provided evidence that the simulated data reproduced the imposed conditions reasonably well (see Simulation Tables in the <xref ref-type="sec" rid="sec202">supplementary files</xref>). However, as in other studies (<xref ref-type="bibr" rid="ref51">Pardo and Ferrer, 2013</xref>; <xref ref-type="bibr" rid="ref26">Ferrer and Pardo, 2019</xref>), only skewness and kurtosis deviated from what was expected (the smaller the sample size, the greater the deviation). This occurred because the standard errors of the statistics used to evaluate skewness and kurtosis increased as the sample size decreased (<xref ref-type="bibr" rid="ref68">Wright and Herrington, 2011</xref>).</p>
<sec id="sec11">
<title>False positives</title>
<p>Percentages of false positives obtained with the RCI statistic were systematically higher than the standard nominal level: where one would have expected to find values around 5%, we found values that ranged from 9.2 to 30.3%. These percentages were not significantly altered, neither by the shape of the simulated distributions nor by the different sample sizes used in the present study.</p>
<p>The percentages of false positives obtained with the HA statistic were more acceptable; in fact, these percentages took correct values when <italic>&#x03BB;</italic>&#x2009;=&#x2009;0.4 (regardless of the shape of the distribution and of the sample size) and when <italic>&#x03BB;</italic>&#x2009;=&#x2009;0.5 if <italic>n</italic>&#x2009;=&#x2009;100 (regardless of the shape of the distribution). In the rest of the simulated conditions, percentages higher than the nominal level were obtained, although in no case were values observed as high as those obtained with the RCI statistic.</p>
</sec>
<sec id="sec12">
<title>False negatives</title>
<p>RCI and HA were better comparable in terms of the percentage of false negatives they generated. With the RCI statistic, these percentages tended to improve as the value of &#x03BB; increased; but correct percentages were only obtained if &#x03B4; was greater than 1. With the HA statistic, the percentages of false negatives were also better when <italic>&#x03BB;</italic> equaled 0.5 or 0.6 than when it equaled 0.4, but some correct percentages were also obtained when <italic>&#x03B4;</italic>&#x2009;=&#x2009;0.8. It also occurred that the percentages of false negatives improved slightly as sample size increased (this occurred in relation to both the RCI and the HA statistic).</p>
</sec>
</sec>
<sec sec-type="discussions" id="sec13">
<title>Discussion</title>
<p>
<list list-type="simple">
<list-item><p>The aim of the present study was to estimate the rate of false positives and false negatives associated with RCI and HA, incorporating the use of a new way of estimating reliability. Since false positives and false negatives represent classification errors, it would be reasonable to expect a good diagnostic method to be able to make proper classifications while maintaining low rates of false positives and false negatives.</p></list-item>
<list-item><p>It is commonly assumed that the false positive rate should be around 0.05. How low the false negative rate should be is also a subjective issue, but in applied research and clinical practice, it is common to consider that this rate should not exceed 20% (<xref ref-type="bibr" rid="ref12">Cohen, 1988</xref>, <xref ref-type="bibr" rid="ref13">1992</xref>). Taking these two conventional values as a reference (5 and 20%, respectively), the results of the present study indicate that:</p></list-item>
</list>
<list list-type="alpha-lower">
<list-item><p>RCI offers unacceptable false positives rates in all simulated conditions. As this occurs when reliability is estimated by Cronbach&#x2019;s alpha coefficient (<xref ref-type="bibr" rid="ref26">Ferrer and Pardo, 2019</xref>), when reliability is estimated by the <italic>omega<sub>h</sub></italic> coefficient, false positive rates associated with RCI take values well above the nominal value. These unacceptable values increase slightly when &#x03BB; increases. When the samples come from normal distributions, they also tend to be higher than when they come from asymmetric distributions.</p></list-item>
<list-item><p>HA offers acceptable false positive rates in some simulated conditions. When <italic>&#x03BB;</italic>&#x2009;=&#x2009;0.4, all false positive rates take correct values (regardless of the sample size and the shape of the simulated distributions). When the value of <italic>&#x03BB;</italic> increases, the false positives rate also increases. The presence of acceptable rates of false positives in several of the simulated conditions indicates that significantly better results are obtained when using the <italic>omega<sub>h</sub></italic> coefficient to estimate reliability than when estimating reliability with the alpha coefficient. It is true that estimating reliability with the test&#x2013;retest correlation provides better results than estimations through alpha (<xref ref-type="bibr" rid="ref26">Ferrer and Pardo, 2019</xref>); however, estimates based on the <italic>omega<sub>h</sub></italic> coefficient do not have the aforementioned drawbacks that estimates based on the test&#x2013;retest correlation have.</p></list-item>
<list-item><p>Both RCI and HA offer unacceptable rates of false negative. All false negative rates decrease as the effect size increases: this is to be expected if we take into account that the greater the mean of the pre-post differences, the greater a randomly selected individual difference is to be expected. But, even though false negative rates should not exceed 20% (25% applying a criterion similar to the criterion of Bradley for false positives), with RCI statistic rates were found that ranged from 67.4 to 90.8% when the effect size was 0.2 (a small effect size according to Cohen&#x2019;s criteria); and rates that ranged from 12.2 to 66.8% when the effect size was 0.8 (a large effect size according to Cohen&#x2019;s criteria). With the HA statistic, rates were found that ranged from 77.8 to 93.1% when the effect size was 0.2; and rates that ranged from 12.5 to 58.0% when the effect size was 0.8. Therefore, neither RCI nor HA perform well regarding false negative rates.</p></list-item>
</list>
</p>
<p>Nevertheless, to be able to correctly interpret these results, it is necessary to take into account some considerations related to Cohen&#x2019;s standardized difference (&#x03B4;) and the reference values specifically proposed by <xref ref-type="bibr" rid="ref12">Cohen (1988)</xref> to interpret <italic>&#x03B4;</italic>. The cut-off points proposed by Cohen to identify small, medium, and large effect sizes (0.2, 0.5, and 0.8, respectively) do not seem to have been sufficiently justified in order to be accepted as reference values. Indeed, both Cohen and other experts recommended using these cut-off points as mere guides and not as fixed, rigid criteria (<xref ref-type="bibr" rid="ref13">Cohen, 1992</xref>; <xref ref-type="bibr" rid="ref63">Snyder and Lawson, 1993</xref>; <xref ref-type="bibr" rid="ref65">Thompson, 2002</xref>). <xref ref-type="bibr" rid="ref25">Ferguson (2009)</xref>, for example, based on previous reviews (<xref ref-type="bibr" rid="ref27">Franzblau, 1958</xref>; <xref ref-type="bibr" rid="ref41">Lipsey and Hurley, 2009</xref>), proposed reference values that depart markedly from those proposed by Cohen. Ferguson&#x2019;s specific proposal is as follows: 0.41 for a <italic>minimum</italic> effect, 1.15 for a <italic>moderate</italic> effect and 2.70 for a <italic>strong</italic> effect. It is clear that the criteria initially proposed by Cohen (criteria considered valid by most researchers) differ meaningfully from those proposed by Ferguson.</p>
<p>One illustration will suffice. At the evaluation level, for example, the observation of a large therapeutic effect (<italic>&#x03B4;</italic>&#x2009;=&#x2009;0.80 according to Cohen) in the positive direction (i.e., less complaints/negative affect or greater well-being/positive affect) suggests that 19.9% of the clients obtain pre-post differences that represent a reliable change (i.e., differences that surpass the cutoff point 1.645, the 95th percentile of a normal distribution). When a large effect size is achieved following the directives of Ferguson (<italic>&#x03B4;</italic>&#x2009;=&#x2009;2.70), 85.4% of clients obtain pre-post differences that represent a reliable change (in calculating these percentages we assume that pre-post differences are normally distributed).</p>
<p>These considerations about the cut-off points used to define small, medium and large effect sizes lead to the following conclusion: taking 1.15 (instead of 0.5) as a reference value for an effect of medium size, the false positive rates associated with the HA statistic seem quite correct when <italic>&#x03BB;</italic>&#x2009;&#x003E;&#x2009;0.4. Therefore, the false negative rate obtained does not seem as unacceptable as it initially appeared.</p>
<p>Finally, classifications resulting from the application of these cutoffs could be improved if the results obtained by applying distribution-based methods such as RCI and HA were supplemented by information provided by anchor-based methods (<xref ref-type="bibr" rid="ref2">Barrett et al., 2008</xref>; <xref ref-type="bibr" rid="ref20">de Vet and Terwee, 2010</xref>; <xref ref-type="bibr" rid="ref34">Houweling, 2010</xref>; <xref ref-type="bibr" rid="ref66">Turner et al., 2010</xref>) or the cumulative proportion of responders (<xref ref-type="bibr" rid="ref24">Farrar et al., 2006</xref>; <xref ref-type="bibr" rid="ref47">McLeod et al., 2011</xref>; <xref ref-type="bibr" rid="ref70">Wyrwich et al., 2013</xref>). This, however, is an area in need of further research.</p>
</sec>
<sec sec-type="conclusions" id="sec14">
<title>Conclusion</title>
<p>The objective of the present study was to learn about the false negative and false negative rates associated with two distribution-based methods (RCI and HA) designed to evaluate individual change (reliable change) in pre-post designs. The novelty of this study is that reliability has been estimated by the <italic>omega<sub>h</sub></italic> coefficient rather than with the alpha coefficient or the test&#x2013;retest correlation.</p>
<p>Regarding the rate of false positives, only the HA statistic provides acceptable results. Regarding the rate of false negatives, both statistics offer similar results, and both can claim to offer acceptable rates when Ferguson&#x2019;s stringent criteria are used to define effect sizes rather than when the conventional criteria advanced by Cohen is employed.</p>
<p>Since the HA statistic seems to be a better option than the RCI statistic, we have developed an Excel macro (see <xref ref-type="sec" rid="sec202">Supplementary files</xref>) so that the greater complexity of calculating HA does not represent an obstacle for the non-expert user.</p>
<p>The methods used to establish the minimally reliable change analyzed in the present study offer an opportunity to assess the change experienced by a person or a group of people as a consequence of an intervention. So far, we have used the clinical context as an example, but this approach could be used in a wide range of contexts, e.g., in educational, community, and/or social intervention areas, to assess the effectiveness of skills training program, to test interventions in the organizational area, to evaluate cognitive stimulation and/or learning programs, etc.</p>
<p>However, some considerations must be taken into account before applying these reliable change measures. This approach is used in pre-post research designs; the trait or symptoms of interest should be susceptible to change as a result of the intervention; the scales used must have evidence of validity and sufficient reliability (because reliability is an important parameter within the equation for its estimation); and certain minimum reference information must be available or there must be a sufficient sample to estimate this information. For example, it could be applied with scales commonly used in psychotherapeutic contexts, e.g., the Beck Depression Inventory (BDI), the Hamilton Anxiety Rating Scale (HAM-A), and the Global Assessment of Functioning (GAF).</p>
</sec>
<sec sec-type="data-availability" id="sec15">
<title>Data availability statement</title>
<p>The datasets presented in this study can be found in online repositories. The names of the repository/repositories and accession number(s) can be found in the article/<xref ref-type="sec" rid="sec202">supplementary material</xref>.</p>
</sec>
<sec id="sec16">
<title>Author contributions</title>
<p>RF-U: conceptualization, data curation, formal analysis, methodology, writing &#x2013; original draft, writing &#x2013; review, and editing. AP: conceptualization, formal analysis, methodology, supervision, validation, writing &#x2013; original draft, writing &#x2013; review, and editing. WA: conceptualization, supervision, validation, writing &#x2013; review, and editing. GP-G: formal analysis, writing &#x2013; review, editing, and validation. All authors contributed to the article and approved the submitted version.</p>
</sec>
<sec sec-type="COI-statement" id="sec17">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="sec100" sec-type="disclaimer">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<sec id="sec202" sec-type="supplementary-material">
<title>Supplementary material</title>
<p>The Supplementary material for this article can be found online at: <ext-link xlink:href="https://www.frontiersin.org/articles/10.3389/fpsyg.2023.1132128/full#supplementary-material" ext-link-type="uri">https://www.frontiersin.org/articles/10.3389/fpsyg.2023.1132128/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.zip" id="SM1" mimetype="application/zip" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="ref1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Atkins</surname> <given-names>D. C.</given-names></name> <name><surname>Bedics</surname> <given-names>J. D.</given-names></name> <name><surname>Mcglinchey</surname> <given-names>J. B.</given-names></name> <name><surname>Beauchaine</surname> <given-names>T. P.</given-names></name></person-group> (<year>2005</year>). <article-title>Assessing clinical significance: does it matter which method we use?</article-title> <source>J. Consult. Clin. Psychol.</source> <volume>73</volume>, <fpage>982</fpage>&#x2013;<lpage>989</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0022-006X.73.5.982</pub-id>, PMID: <pub-id pub-id-type="pmid">16287398</pub-id></citation>
</ref>
<ref id="ref2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Barrett</surname> <given-names>B.</given-names></name> <name><surname>Brown</surname> <given-names>R.</given-names></name> <name><surname>Mundt</surname> <given-names>M.</given-names></name></person-group> (<year>2008</year>). <article-title>Comparison of anchor-based and distributional approaches in estimating important difference in common cold</article-title>. <source>Qual. Life Res.</source> <volume>17</volume>, <fpage>75</fpage>&#x2013;<lpage>85</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11136-007-9277-2</pub-id>, PMID: <pub-id pub-id-type="pmid">18027107</pub-id></citation>
</ref>
<ref id="ref3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bauer</surname> <given-names>S.</given-names></name> <name><surname>Lambert</surname> <given-names>M. J.</given-names></name> <name><surname>Nielsen</surname> <given-names>S. L.</given-names></name></person-group> (<year>2004</year>). <article-title>Clinical significance methods: a comparison of statistical techniques</article-title>. <source>J. Pers. Assess.</source> <volume>82</volume>, <fpage>60</fpage>&#x2013;<lpage>70</lpage>. doi: <pub-id pub-id-type="doi">10.1207/s15327752jpa8201_11</pub-id></citation>
</ref>
<ref id="ref4">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Becker</surname> <given-names>G.</given-names></name>
</person-group> (<year>2000</year>). <article-title>How important is transient error in estimating reliability? Going beyond simulation studies</article-title>. <source>Psychol. Methods</source> <volume>5</volume>, <fpage>370</fpage>&#x2013;<lpage>379</lpage>. doi: <pub-id pub-id-type="doi">10.1037/1082-989x.5.3.370</pub-id>, PMID: <pub-id pub-id-type="pmid">11004874</pub-id></citation>
</ref>
<ref id="ref5">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Bentler</surname> <given-names>P. M.</given-names></name>
</person-group> (<year>2009</year>). <article-title>Alpha, dimension-free, and model-based internal consistency reliability</article-title>. <source>Psychometrika</source> <volume>74</volume>, <fpage>137</fpage>&#x2013;<lpage>143</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11336-008-9100-1</pub-id>, PMID: <pub-id pub-id-type="pmid">20161430</pub-id></citation>
</ref>
<ref id="ref6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bischoff</surname> <given-names>T.</given-names></name> <name><surname>Anderson</surname> <given-names>S. R.</given-names></name> <name><surname>Heafner</surname> <given-names>J.</given-names></name> <name><surname>Tambling</surname> <given-names>R.</given-names></name></person-group> (<year>2020</year>). <article-title>Establishment of a reliable change index for the GAD-7</article-title>. <source>Psychol. Community Health</source> <volume>8</volume>, <fpage>176</fpage>&#x2013;<lpage>187</lpage>. doi: <pub-id pub-id-type="doi">10.5964/pch.v8i1.309</pub-id>, PMID: <pub-id pub-id-type="pmid">37309101</pub-id></citation>
</ref>
<ref id="ref7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Blanca</surname> <given-names>M. J.</given-names></name> <name><surname>Arnau</surname> <given-names>J.</given-names></name> <name><surname>L&#x00F3;pez-Montiel</surname> <given-names>D.</given-names></name> <name><surname>Bono</surname> <given-names>R.</given-names></name> <name><surname>Bendayan</surname> <given-names>R.</given-names></name></person-group> (<year>2013</year>). <article-title>Skewness and kurtosis in real data samples</article-title>. <source>Methodology</source> <volume>9</volume>, <fpage>78</fpage>&#x2013;<lpage>84</lpage>. doi: <pub-id pub-id-type="doi">10.1027/1614-2241/a000057</pub-id></citation>
</ref>
<ref id="ref8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Botella</surname> <given-names>J.</given-names></name> <name><surname>Bl&#x00E1;zquez</surname> <given-names>D.</given-names></name> <name><surname>Suero</surname> <given-names>M.</given-names></name> <name><surname>Juola</surname> <given-names>J. F.</given-names></name></person-group> (<year>2018</year>). <article-title>Assessing individual change without knowing the test properties: item bootstrapping</article-title>. <source>Front. Psychol.</source> <volume>9</volume>:<fpage>223</fpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2018.00223</pub-id>, PMID: <pub-id pub-id-type="pmid">29593591</pub-id></citation>
</ref>
<ref id="ref9">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Bradley</surname> <given-names>J. V.</given-names></name>
</person-group> (<year>1978</year>). <article-title>Robustness?</article-title> <source>Br. J. Math. Stat. Psychol.</source> <volume>31</volume>, <fpage>144</fpage>&#x2013;<lpage>152</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.2044-8317.1978.tb00581.x</pub-id>, PMID: <pub-id pub-id-type="pmid">37356068</pub-id></citation>
</ref>
<ref id="ref10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Christensen</surname> <given-names>L.</given-names></name> <name><surname>Mendoza</surname> <given-names>J. L.</given-names></name></person-group> (<year>1986</year>). <article-title>A method of assessing change in a single subject: an alteration of the RC index</article-title>. <source>Behav. Ther.</source> <volume>17</volume>, <fpage>305</fpage>&#x2013;<lpage>308</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0005-7894(86)80060-0</pub-id></citation>
</ref>
<ref id="ref11">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Cicchetti</surname> <given-names>D. V.</given-names></name>
</person-group> (<year>1994</year>). <article-title>Guidelines, criteria, and rules of thumb for evaluating normed and standardized assessment instruments in psychology</article-title>. <source>Psychol. Assess.</source> <volume>6</volume>, <fpage>284</fpage>&#x2013;<lpage>290</lpage>. doi: <pub-id pub-id-type="doi">10.1037/1040-3590.6.4.284</pub-id></citation>
</ref>
<ref id="ref12">
<citation citation-type="book"><person-group person-group-type="author">
<name><surname>Cohen</surname> <given-names>J.</given-names></name>
</person-group> (<year>1988</year>). <source>Statistical power analysis for the behavioral sciences</source>. (<edition>2nd Edn.</edition>) <publisher-loc>New York</publisher-loc>: <publisher-name>Routledge</publisher-name>.</citation>
</ref>
<ref id="ref13">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Cohen</surname> <given-names>J.</given-names></name>
</person-group> (<year>1992</year>). <article-title>A power primer</article-title>. <source>Psychol. Bull.</source> <volume>112</volume>, <fpage>155</fpage>&#x2013;<lpage>159</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0033-2909.112.1.155</pub-id>, PMID: <pub-id pub-id-type="pmid">19565683</pub-id></citation>
</ref>
<ref id="ref14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crawford</surname> <given-names>J. R.</given-names></name> <name><surname>Garthwaite</surname> <given-names>P. H.</given-names></name></person-group> (<year>2006</year>). <article-title>Comparing patients&#x2019; predicted test scores from a regression equation with their obtained scores: a significance test and point estimate of abnormality with accompanying confidence limits</article-title>. <source>Neuropsychology</source> <volume>20</volume>, <fpage>259</fpage>&#x2013;<lpage>271</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0894-4105.20.3.259</pub-id>, PMID: <pub-id pub-id-type="pmid">16719619</pub-id></citation>
</ref>
<ref id="ref15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crawford</surname> <given-names>J. R.</given-names></name> <name><surname>Howell</surname> <given-names>D. C.</given-names></name></person-group> (<year>1998</year>). <article-title>Regression equations in clinical neuropsychology: an evaluation of statistical methods for comparing predicted and obtained scores</article-title>. <source>J. Clin. Exp. Neuropsychol.</source> <volume>20</volume>, <fpage>755</fpage>&#x2013;<lpage>762</lpage>. doi: <pub-id pub-id-type="doi">10.1076/jcen.20.5.755.1132</pub-id>, PMID: <pub-id pub-id-type="pmid">10079050</pub-id></citation>
</ref>
<ref id="ref16">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Cronbach</surname> <given-names>L. J.</given-names></name>
</person-group> (<year>1951</year>). <article-title>Coefficient alpha and the internal structure of tests</article-title>. <source>Psychometrika</source> <volume>16</volume>, <fpage>297</fpage>&#x2013;<lpage>334</lpage>. doi: <pub-id pub-id-type="doi">10.1007/BF02310555</pub-id>, PMID: <pub-id pub-id-type="pmid">37355567</pub-id></citation>
</ref>
<ref id="ref17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crosby</surname> <given-names>R. D.</given-names></name> <name><surname>Kolotkin</surname> <given-names>R. L.</given-names></name> <name><surname>Williams</surname> <given-names>G. R.</given-names></name></person-group> (<year>2003</year>). <article-title>Defining clinically meaningful change in health-related quality of life</article-title>. <source>J. Clin. Epidemiol.</source> <volume>56</volume>, <fpage>395</fpage>&#x2013;<lpage>407</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0895-4356(03)00044-1</pub-id>, PMID: <pub-id pub-id-type="pmid">12812812</pub-id></citation>
</ref>
<ref id="ref18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crutzen</surname> <given-names>R.</given-names></name> <name><surname>Peters</surname> <given-names>G.-J. Y.</given-names></name></person-group> (<year>2017</year>). <article-title>Scale quality: alpha is an inadequate estimate and factor-analytic evidence is needed first of all</article-title>. <source>Health Psychol. Rev.</source> <volume>11</volume>, <fpage>242</fpage>&#x2013;<lpage>247</lpage>. doi: <pub-id pub-id-type="doi">10.1080/17437199.2015.1124240</pub-id>, PMID: <pub-id pub-id-type="pmid">26602990</pub-id></citation>
</ref>
<ref id="ref19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cumming</surname> <given-names>G.</given-names></name> <name><surname>Finch</surname> <given-names>S.</given-names></name></person-group> (<year>2001</year>). <article-title>A primer on the understanding, use, and calculation of confidence intervals that are based on central and noncentral distributions</article-title>. <source>Educ. Psychol. Meas.</source> <volume>61</volume>, <fpage>532</fpage>&#x2013;<lpage>574</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0013164401614002</pub-id></citation>
</ref>
<ref id="ref20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>de Vet</surname> <given-names>H. C. W.</given-names></name> <name><surname>Terwee</surname> <given-names>C. B.</given-names></name></person-group> (<year>2010</year>). <article-title>The minimal detectable change should not replace the minimal important difference</article-title>. <source>J. Clin. Epidemiol.</source> <volume>63</volume>, <fpage>804</fpage>&#x2013;<lpage>805</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.jclinepi.2009.12.015</pub-id>, PMID: <pub-id pub-id-type="pmid">20399609</pub-id></citation>
</ref>
<ref id="ref21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dunn</surname> <given-names>T. J.</given-names></name> <name><surname>Baguley</surname> <given-names>T.</given-names></name> <name><surname>Brunsden</surname> <given-names>V.</given-names></name></person-group> (<year>2014</year>). <article-title>From alpha to omega: a practical solution to the pervasive problem of internal consistency estimation</article-title>. <source>Br. J. Psychol. Lond. Engl.</source> <volume>105</volume>, <fpage>399</fpage>&#x2013;<lpage>412</lpage>. doi: <pub-id pub-id-type="doi">10.1111/bjop.12046</pub-id></citation>
</ref>
<ref id="ref22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Estrada</surname> <given-names>E.</given-names></name> <name><surname>Caperos</surname> <given-names>J. M.</given-names></name> <name><surname>Pardo</surname> <given-names>A.</given-names></name></person-group> (<year>2020</year>). <article-title>Change in the center of the distribution and in the individual scores: relation with heteroskedastic pre- and post-test distributions</article-title>. <source>Psicothema</source> <volume>32</volume>, <fpage>410</fpage>&#x2013;<lpage>419</lpage>. doi: <pub-id pub-id-type="doi">10.7334/psicothema2019.396</pub-id>, PMID: <pub-id pub-id-type="pmid">32711677</pub-id></citation>
</ref>
<ref id="ref23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Estrada</surname> <given-names>E.</given-names></name> <name><surname>Ferrer</surname> <given-names>E.</given-names></name> <name><surname>Pardo</surname> <given-names>A.</given-names></name></person-group> (<year>2019</year>). <article-title>Statistics for evaluating pre-post change: relation between change in the distribution center and change in the individual scores</article-title>. <source>Front. Psychol.</source> <volume>9</volume>:<fpage>2696</fpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2018.02696</pub-id>, PMID: <pub-id pub-id-type="pmid">30671008</pub-id></citation>
</ref>
<ref id="ref24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Farrar</surname> <given-names>J. T.</given-names></name> <name><surname>Dworkin</surname> <given-names>R. H.</given-names></name> <name><surname>Max</surname> <given-names>M. B.</given-names></name></person-group> (<year>2006</year>). <article-title>Use of the cumulative proportion of responders analysis graph to present pain data over a range of cut-off points: making clinical trial data more understandable</article-title>. <source>J. Pain Symptom Manag.</source> <volume>31</volume>, <fpage>369</fpage>&#x2013;<lpage>377</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.jpainsymman.2005.08.018</pub-id>, PMID: <pub-id pub-id-type="pmid">16632085</pub-id></citation>
</ref>
<ref id="ref25">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Ferguson</surname> <given-names>C. J.</given-names></name>
</person-group> (<year>2009</year>). <article-title>An effect size primer: a guide for clinicians and researchers</article-title>. <source>Prof. Psychol. Res. Pract.</source> <volume>40</volume>, <fpage>532</fpage>&#x2013;<lpage>538</lpage>. doi: <pub-id pub-id-type="doi">10.1037/a0015808</pub-id></citation>
</ref>
<ref id="ref26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ferrer</surname> <given-names>R.</given-names></name> <name><surname>Pardo</surname> <given-names>A.</given-names></name></person-group> (<year>2019</year>). <article-title>Clinically meaningful change: false negatives in the estimation of individual change</article-title>. <source>Methodology</source> <volume>15</volume>, <fpage>97</fpage>&#x2013;<lpage>105</lpage>. doi: <pub-id pub-id-type="doi">10.1027/1614-2241/a000168</pub-id></citation>
</ref>
<ref id="ref27">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Franzblau</surname> <given-names>A. N.</given-names></name> <name><surname>Abraham</surname> <given-names>N.</given-names></name></person-group> (<year>1958</year>). <source>A primer of statistics for non-statisticians</source>. <publisher-loc>New York</publisher-loc>, <publisher-name>Harcourt, Brace</publisher-name>.</citation>
</ref>
<ref id="ref28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gatchel</surname> <given-names>R. J.</given-names></name> <name><surname>Mayer</surname> <given-names>T. G.</given-names></name></person-group> (<year>2010</year>). <article-title>Testing minimal clinically important difference: consensus or conundrum?</article-title> <source>Spine J</source> <volume>10</volume>, <fpage>321</fpage>&#x2013;<lpage>327</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.spinee.2009.10.015</pub-id></citation>
</ref>
<ref id="ref29">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Graham</surname> <given-names>J. M.</given-names></name>
</person-group> (<year>2006</year>). <article-title>Congeneric and (essentially) Tau-equivalent estimates of score reliability: what they are and how to use them</article-title>. <source>Educ. Psychol. Meas.</source> <volume>66</volume>, <fpage>930</fpage>&#x2013;<lpage>944</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0013164406288165</pub-id></citation>
</ref>
<ref id="ref30">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Green</surname> <given-names>S. B.</given-names></name>
</person-group> (<year>2003</year>). <article-title>A coefficient alpha for test-retest data</article-title>. <source>Psychol. Methods</source> <volume>8</volume>, <fpage>88</fpage>&#x2013;<lpage>101</lpage>. doi: <pub-id pub-id-type="doi">10.1037/1082-989X.8.1.88</pub-id>, PMID: <pub-id pub-id-type="pmid">12741675</pub-id></citation>
</ref>
<ref id="ref31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Green</surname> <given-names>S. B.</given-names></name> <name><surname>Yang</surname> <given-names>Y.</given-names></name></person-group> (<year>2009</year>). <article-title>Commentary on coefficient alpha: a cautionary tale</article-title>. <source>Psychometrika</source> <volume>74</volume>, <fpage>121</fpage>&#x2013;<lpage>135</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11336-008-9098-4</pub-id></citation>
</ref>
<ref id="ref32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hageman</surname> <given-names>W. J.</given-names></name> <name><surname>Arrindell</surname> <given-names>W. A.</given-names></name></person-group> (<year>1999</year>). <article-title>Establishing clinically significant change: increment of precision and the distinction between individual and group level of analysis</article-title>. <source>Behav. Res. Ther.</source> <volume>37</volume>, <fpage>1169</fpage>&#x2013;<lpage>1193</lpage>. doi: <pub-id pub-id-type="doi">10.1016/s0005-7967(99)00032-7</pub-id>, PMID: <pub-id pub-id-type="pmid">10596464</pub-id></citation>
</ref>
<ref id="ref33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hogan</surname> <given-names>T. P.</given-names></name> <name><surname>Benjamin</surname> <given-names>A.</given-names></name> <name><surname>Brezinski</surname> <given-names>K. L.</given-names></name></person-group> (<year>2000</year>). <article-title>Reliability methods: a note on the frequency of use of various types</article-title>. <source>Educ. Psychol. Meas.</source> <volume>60</volume>, <fpage>523</fpage>&#x2013;<lpage>531</lpage>. doi: <pub-id pub-id-type="doi">10.1177/00131640021970691</pub-id>, PMID: <pub-id pub-id-type="pmid">36003241</pub-id></citation>
</ref>
<ref id="ref34">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Houweling</surname> <given-names>T. A. W.</given-names></name>
</person-group> (<year>2010</year>). <article-title>Reporting improvement from patient-reported outcome measures: a review</article-title>. <source>Clin. Chiropr.</source> <volume>13</volume>, <fpage>15</fpage>&#x2013;<lpage>22</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.clch.2009.12.003</pub-id>, PMID: <pub-id pub-id-type="pmid">37355186</pub-id></citation>
</ref>
<ref id="ref35">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Hsu</surname> <given-names>L. M.</given-names></name>
</person-group> (<year>1989</year>). <article-title>Reliable changes in psychotherapy: taking into account regression toward the mean</article-title>. <source>Behav. Assess.</source> <volume>11</volume>, <fpage>459</fpage>&#x2013;<lpage>467</lpage>.</citation>
</ref>
<ref id="ref36">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Hsu</surname> <given-names>L. M.</given-names></name>
</person-group> (<year>1995</year>). <article-title>Regression toward the mean associated with measurement error and the identification of improvement and deterioration in psychotherapy</article-title>. <source>J. Consult. Clin. Psychol.</source> <volume>63</volume>, <fpage>141</fpage>&#x2013;<lpage>144</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0022-006X.63.1.141</pub-id>, PMID: <pub-id pub-id-type="pmid">7896979</pub-id></citation>
</ref>
<ref id="ref37">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Hsu</surname> <given-names>L. M.</given-names></name>
</person-group> (<year>1996</year>). <article-title>On the identification of clinically significant client changes: Reinterpretation of Jacobson&#x2019;s cut scores</article-title>. <source>J. Psychopathol. Behav. Assess.</source> <volume>18</volume>, <fpage>371</fpage>&#x2013;<lpage>385</lpage>. doi: <pub-id pub-id-type="doi">10.1007/BF02229141</pub-id></citation>
</ref>
<ref id="ref38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobson</surname> <given-names>N. S.</given-names></name> <name><surname>Follette</surname> <given-names>W. C.</given-names></name> <name><surname>Revenstorf</surname> <given-names>D.</given-names></name></person-group> (<year>1984</year>). <article-title>Psychotherapy outcome research: methods for reporting variability and evaluating clinical significance</article-title>. <source>Behav. Ther.</source> <volume>15</volume>, <fpage>336</fpage>&#x2013;<lpage>352</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0005-7894(84)80002-7</pub-id>, PMID: <pub-id pub-id-type="pmid">36879204</pub-id></citation>
</ref>
<ref id="ref39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobson</surname> <given-names>N. S.</given-names></name> <name><surname>Roberts</surname> <given-names>L. J.</given-names></name> <name><surname>Berns</surname> <given-names>S. B.</given-names></name> <name><surname>McGlinchey</surname> <given-names>J. B.</given-names></name></person-group> (<year>1999</year>). <article-title>Methods for defining and determining the clinical significance of treatment effects: description, application, and alternatives</article-title>. <source>J. Consult. Clin. Psychol.</source> <volume>67</volume>, <fpage>300</fpage>&#x2013;<lpage>307</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0022-006X.67.3.300</pub-id>, PMID: <pub-id pub-id-type="pmid">10369050</pub-id></citation>
</ref>
<ref id="ref40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobson</surname> <given-names>N. S.</given-names></name> <name><surname>Truax</surname> <given-names>P.</given-names></name></person-group> (<year>1991</year>). <article-title>Clinical significance: a statistical approach to defining meaningful change in psychotherapy research</article-title>. <source>J. Consult. Clin. Psychol.</source> <volume>59</volume>, <fpage>12</fpage>&#x2013;<lpage>19</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0022-006X.59.1.12</pub-id>, PMID: <pub-id pub-id-type="pmid">2002127</pub-id></citation>
</ref>
<ref id="ref41">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lipsey</surname> <given-names>M.</given-names></name> <name><surname>Hurley</surname> <given-names>S.</given-names></name></person-group> (<year>2009</year>). <source>The SAGE handbook of applied social research methods</source> (<publisher-loc>Thousand Oaks</publisher-loc>: <publisher-name>SAGE Publications, Inc.</publisher-name>), <fpage>44</fpage>&#x2013;<lpage>76</lpage>.</citation>
</ref>
<ref id="ref42">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Lord</surname> <given-names>F. M.</given-names></name>
</person-group> (<year>1956</year>). <article-title>The measurement of growth</article-title>. <source>ETS Res. Bull. Ser.</source> <volume>1956</volume>, <fpage>i</fpage>&#x2013;<lpage>22</lpage>. doi: <pub-id pub-id-type="doi">10.1002/j.2333-8504.1956.tb00058.x</pub-id></citation>
</ref>
<ref id="ref43">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Lord</surname> <given-names>F. M.</given-names></name>
</person-group> (<year>1963</year>). <article-title>Elementary models for measuring change</article-title>. <source>Probl. Meas. Change</source>, <fpage>21</fpage>&#x2013;<lpage>38</lpage>.</citation>
</ref>
<ref id="ref44">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Maassen</surname> <given-names>G. H.</given-names></name>
</person-group> (<year>2004</year>). <article-title>The standard error in the Jacobson and Truax Reliable Change Index: the classical approach to the assessment of reliable change</article-title>. <source>J. Int. Neuropsychol. Soc.</source> <volume>10</volume>, <fpage>888</fpage>&#x2013;<lpage>893</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S1355617704106097</pub-id>, PMID: <pub-id pub-id-type="pmid">15637779</pub-id></citation>
</ref>
<ref id="ref45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Martinovich</surname> <given-names>Z.</given-names></name> <name><surname>Saunders</surname> <given-names>S.</given-names></name> <name><surname>Howard</surname> <given-names>K.</given-names></name></person-group> (<year>1996</year>). <article-title>Some comments on assessing clinical significance</article-title>. <source>Psychother. Res.</source> <volume>6</volume>, <fpage>124</fpage>&#x2013;<lpage>132</lpage>. doi: <pub-id pub-id-type="doi">10.1080/10503309612331331648</pub-id></citation>
</ref>
<ref id="ref46">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>McGlinchey</surname> <given-names>J. B.</given-names></name> <name><surname>Atkins</surname> <given-names>D. C.</given-names></name> <name><surname>Jacobson</surname> <given-names>N. S.</given-names></name></person-group> (<year>2002</year>). <article-title>Clinical significance methods: which one to use and how useful are they?</article-title> <source>Behav. Ther.</source> <volume>33</volume>, <fpage>529</fpage>&#x2013;<lpage>550</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0005-7894(02)80015-6</pub-id>, PMID: <pub-id pub-id-type="pmid">37336849</pub-id></citation>
</ref>
<ref id="ref47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>McLeod</surname> <given-names>L. D.</given-names></name> <name><surname>Coon</surname> <given-names>C. D.</given-names></name> <name><surname>Martin</surname> <given-names>S. A.</given-names></name> <name><surname>Fehnel</surname> <given-names>S. E.</given-names></name> <name><surname>Hays</surname> <given-names>R. D.</given-names></name></person-group> (<year>2011</year>). <article-title>Interpreting patient-reported outcome results: US FDA guidance and emerging methods</article-title>. <source>Expert Rev. Pharmacoecon. Outcomes Res.</source> <volume>11</volume>, <fpage>163</fpage>&#x2013;<lpage>169</lpage>. doi: <pub-id pub-id-type="doi">10.1586/erp.11.12</pub-id>, PMID: <pub-id pub-id-type="pmid">21476818</pub-id></citation>
</ref>
<ref id="ref48">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Micceri</surname> <given-names>T.</given-names></name>
</person-group> (<year>1989</year>). <article-title>The unicorn, the normal curve, and other improbable creatures</article-title>. <source>Psychol. Bull.</source> <volume>105</volume>, <fpage>156</fpage>&#x2013;<lpage>166</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0033-2909.105.1.156</pub-id></citation>
</ref>
<ref id="ref49">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nunnally</surname> <given-names>J. C.</given-names></name> <name><surname>Kotsch</surname> <given-names>W. E.</given-names></name></person-group> (<year>1983</year>). <article-title>Studies of individual subjects: logic and methods of analysis</article-title>. <source>Br. J. Clin. Psychol.</source> <volume>22</volume>, <fpage>83</fpage>&#x2013;<lpage>93</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.2044-8260.1983.tb00582.x</pub-id>, PMID: <pub-id pub-id-type="pmid">37120641</pub-id></citation>
</ref>
<ref id="ref50">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Page</surname> <given-names>P.</given-names></name>
</person-group> (<year>2014</year>). <article-title>Beyond statistical significance: clinical interpretation of rehabilitation research literature</article-title>. <source>Int. J. Sports Phys. Ther.</source> <volume>9</volume>, <fpage>726</fpage>&#x2013;<lpage>736</lpage>. PMID: <pub-id pub-id-type="pmid">25328834</pub-id></citation>
</ref>
<ref id="ref51">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pardo</surname> <given-names>A.</given-names></name> <name><surname>Ferrer</surname> <given-names>R.</given-names></name></person-group> (<year>2013</year>). <article-title>Significaci&#x00F3;n cl&#x00ED;nica: falsos positivos en la estimaci&#x00F3;n del cambio individual</article-title>. <source>An. Psicol.</source> <volume>29</volume>, <fpage>301</fpage>&#x2013;<lpage>310</lpage>. doi: <pub-id pub-id-type="doi">10.6018/analesps.29.2.139031</pub-id></citation>
</ref>
<ref id="ref52">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Peterson</surname> <given-names>R. A.</given-names></name>
</person-group> (<year>2000</year>). <article-title>A meta-analysis of variance accounted for and factor loadings in exploratory factor analysis</article-title>. <source>Mark. Lett.</source> <volume>11</volume>, <fpage>261</fpage>&#x2013;<lpage>275</lpage>. doi: <pub-id pub-id-type="doi">10.1023/A:1008191211004</pub-id></citation>
</ref>
<ref id="ref53">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Revelle</surname> <given-names>W.</given-names></name> <name><surname>Zinbarg</surname> <given-names>R. E.</given-names></name></person-group> (<year>2009</year>). <article-title>Coefficients alpha, beta, omega, and the glb: comments on Sijtsma</article-title>. <source>Psychometrika</source> <volume>74</volume>, <fpage>145</fpage>&#x2013;<lpage>154</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11336-008-9102-z</pub-id></citation>
</ref>
<ref id="ref54">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Revicki</surname> <given-names>D.</given-names></name> <name><surname>Hays</surname> <given-names>R. D.</given-names></name> <name><surname>Cella</surname> <given-names>D.</given-names></name> <name><surname>Sloan</surname> <given-names>J.</given-names></name></person-group> (<year>2008</year>). <article-title>Recommended methods for determining responsiveness and minimally important differences for patient-reported outcomes</article-title>. <source>J. Clin. Epidemiol.</source> <volume>61</volume>, <fpage>102</fpage>&#x2013;<lpage>109</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.jclinepi.2007.03.012</pub-id>, PMID: <pub-id pub-id-type="pmid">18177782</pub-id></citation>
</ref>
<ref id="ref55">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rogosa</surname> <given-names>D. R.</given-names></name> <name><surname>Willett</surname> <given-names>J. B.</given-names></name></person-group> (<year>1983</year>). <article-title>Demonstrating the reliability of the difference score in the measurement of change</article-title>. <source>J. Educ. Meas.</source> <volume>20</volume>, <fpage>335</fpage>&#x2013;<lpage>343</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.1745-3984.1983.tb00211.x</pub-id>, PMID: <pub-id pub-id-type="pmid">37354340</pub-id></citation>
</ref>
<ref id="ref56">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ronk</surname> <given-names>F. R.</given-names></name> <name><surname>Hooke</surname> <given-names>G. R.</given-names></name> <name><surname>Page</surname> <given-names>A. C.</given-names></name></person-group> (<year>2012</year>). <article-title>How consistent are clinical significance classifications when calculation methods and outcome measures differ?</article-title> <source>Clin. Psychol. Sci. Pract.</source> <volume>19</volume>, <fpage>167</fpage>&#x2013;<lpage>179</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.1468-2850.2012.01281.x</pub-id></citation>
</ref>
<ref id="ref57">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ronk</surname> <given-names>F. R.</given-names></name> <name><surname>Hooke</surname> <given-names>G. R.</given-names></name> <name><surname>Page</surname> <given-names>A. C.</given-names></name></person-group> (<year>2016</year>). <article-title>Validity of clinically significant change classifications yielded by Jacobson-Truax and Hageman-Arrindell methods</article-title>. <source>BMC Psychiatry</source> <volume>16</volume>:<fpage>187</fpage>. doi: <pub-id pub-id-type="doi">10.1186/s12888-016-0895-5</pub-id>, PMID: <pub-id pub-id-type="pmid">27267986</pub-id></citation>
</ref>
<ref id="ref58">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schmidt</surname> <given-names>F. L.</given-names></name> <name><surname>Le</surname> <given-names>H.</given-names></name> <name><surname>Ilies</surname> <given-names>R.</given-names></name></person-group> (<year>2003</year>). <article-title>Beyond alpha: an empirical examination of the effects of different sources of measurement error on reliability estimates for measures of individual differences constructs</article-title>. <source>Psychol. Methods</source> <volume>8</volume>, <fpage>206</fpage>&#x2013;<lpage>224</lpage>. doi: <pub-id pub-id-type="doi">10.1037/1082-989X.8.2.206</pub-id>, PMID: <pub-id pub-id-type="pmid">12924815</pub-id></citation>
</ref>
<ref id="ref59">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Schmitt</surname> <given-names>N.</given-names></name>
</person-group> (<year>1996</year>). <article-title>Uses and abuses of coefficient alpha</article-title>. <source>Psychol. Assess.</source> <volume>8</volume>, <fpage>350</fpage>&#x2013;<lpage>353</lpage>. doi: <pub-id pub-id-type="doi">10.1037/1040-3590.8.4.350</pub-id>, PMID: <pub-id pub-id-type="pmid">12720614</pub-id></citation>
</ref>
<ref id="ref60">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shalaby</surname> <given-names>R.</given-names></name> <name><surname>Spurvey</surname> <given-names>P.</given-names></name> <name><surname>Knox</surname> <given-names>M.</given-names></name> <name><surname>Rathwell</surname> <given-names>R.</given-names></name> <name><surname>Vuong</surname> <given-names>W.</given-names></name> <name><surname>Surood</surname> <given-names>S.</given-names></name> <etal/></person-group>. (<year>2022</year>). <article-title>Clinical outcomes in routine evaluation measures for patients discharged from acute psychiatric care: four-arm peer and text messaging support controlled observational study</article-title>. <source>Int. J. Environ. Res. Public Health</source> <volume>19</volume>:<fpage>3798</fpage>. doi: <pub-id pub-id-type="doi">10.3390/ijerph19073798</pub-id>, PMID: <pub-id pub-id-type="pmid">35409483</pub-id></citation>
</ref>
<ref id="ref61">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shevlin</surname> <given-names>M.</given-names></name> <name><surname>Miles</surname> <given-names>J. N. V.</given-names></name> <name><surname>Davies</surname> <given-names>M. N. O.</given-names></name> <name><surname>Walker</surname> <given-names>S.</given-names></name></person-group> (<year>2000</year>). <article-title>Coefficient alpha: a useful indicator of reliability?</article-title> <source>Personal. Individ. Differ.</source> <volume>28</volume>, <fpage>229</fpage>&#x2013;<lpage>237</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0191-8869(99)00093-8</pub-id></citation>
</ref>
<ref id="ref62">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Sijtsma</surname> <given-names>K.</given-names></name>
</person-group> (<year>2009</year>). <article-title>On the use, the misuse, and the very limited usefulness of Cronbach&#x2019;s alpha</article-title>. <source>Psychometrika</source> <volume>74</volume>, <fpage>107</fpage>&#x2013;<lpage>120</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11336-008-9101-0</pub-id>, PMID: <pub-id pub-id-type="pmid">20037639</pub-id></citation>
</ref>
<ref id="ref63">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Snyder</surname> <given-names>P.</given-names></name> <name><surname>Lawson</surname> <given-names>S.</given-names></name></person-group> (<year>1993</year>). <article-title>Evaluating results using corrected and uncorrected effect size estimates</article-title>. <source>J. Exp. Educ.</source> <volume>61</volume>, <fpage>334</fpage>&#x2013;<lpage>349</lpage>. doi: <pub-id pub-id-type="doi">10.1080/00220973.1993.10806594</pub-id></citation>
</ref>
<ref id="ref64">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Speer</surname> <given-names>D. C.</given-names></name>
</person-group> (<year>1992</year>). <article-title>Clinically significant change: Jacobson and Truax (1991) revisited</article-title>. <source>J. Consult. Clin. Psychol.</source> <volume>60</volume>, <fpage>402</fpage>&#x2013;<lpage>408</lpage>. doi: <pub-id pub-id-type="doi">10.1037/0022-006X.60.3.402</pub-id>, PMID: <pub-id pub-id-type="pmid">1619094</pub-id></citation>
</ref>
<ref id="ref65">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Thompson</surname> <given-names>B.</given-names></name>
</person-group> (<year>2002</year>). <article-title>&#x201C;Statistical,&#x201D; &#x201C;practical,&#x201D; and &#x201C;clinical&#x201D;: how many kinds of significance do counselors need to consider?</article-title> <source>J. Couns. Dev.</source> <volume>80</volume>, <fpage>64</fpage>&#x2013;<lpage>71</lpage>. doi: <pub-id pub-id-type="doi">10.1002/j.1556-6678.2002.tb00167.x</pub-id></citation>
</ref>
<ref id="ref66">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Turner</surname> <given-names>D.</given-names></name> <name><surname>Sch&#x00FC;nemann</surname> <given-names>H. J.</given-names></name> <name><surname>Griffith</surname> <given-names>L. E.</given-names></name> <name><surname>Beaton</surname> <given-names>D. E.</given-names></name> <name><surname>Griffiths</surname> <given-names>A. M.</given-names></name> <name><surname>Critch</surname> <given-names>J. N.</given-names></name> <etal/></person-group>. (<year>2010</year>). <article-title>The minimal detectable change cannot reliably replace the minimal important difference</article-title>. <source>J. Clin. Epidemiol.</source> <volume>63</volume>, <fpage>28</fpage>&#x2013;<lpage>36</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.jclinepi.2009.01.024</pub-id>, PMID: <pub-id pub-id-type="pmid">19800198</pub-id></citation>
</ref>
<ref id="ref67">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Vardanian</surname> <given-names>M. M.</given-names></name> <name><surname>Ramakrishnan</surname> <given-names>A.</given-names></name> <name><surname>Peralta</surname> <given-names>S.</given-names></name> <name><surname>Siddiqui</surname> <given-names>Y.</given-names></name> <name><surname>Shah</surname> <given-names>S. P.</given-names></name> <name><surname>Clark-Whitney</surname> <given-names>E.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>Clinically significant and reliable change: comparing an evidence-based intervention to usual care</article-title>. <source>J. Child Fam. Stud.</source> <volume>29</volume>, <fpage>921</fpage>&#x2013;<lpage>933</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10826-019-01621-3</pub-id>, PMID: <pub-id pub-id-type="pmid">36658663</pub-id></citation>
</ref>
<ref id="ref68">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wright</surname> <given-names>D. B.</given-names></name> <name><surname>Herrington</surname> <given-names>J. A.</given-names></name></person-group> (<year>2011</year>). <article-title>Problematic standard errors and confidence intervals for skewness and kurtosis</article-title>. <source>Behav. Res. Methods</source> <volume>43</volume>, <fpage>8</fpage>&#x2013;<lpage>17</lpage>. doi: <pub-id pub-id-type="doi">10.3758/s13428-010-0044-x</pub-id>, PMID: <pub-id pub-id-type="pmid">21298573</pub-id></citation>
</ref>
<ref id="ref69">
<citation citation-type="journal"><person-group person-group-type="author">
<name><surname>Wyrwich</surname> <given-names>K. W.</given-names></name>
</person-group> (<year>2004</year>). <article-title>Minimal important difference thresholds and the standard error of measurement: is there a connection?</article-title> <source>J. Biopharm. Stat.</source> <volume>14</volume>, <fpage>97</fpage>&#x2013;<lpage>110</lpage>. doi: <pub-id pub-id-type="doi">10.1081/BIP-120028508</pub-id>, PMID: <pub-id pub-id-type="pmid">15027502</pub-id></citation>
</ref>
<ref id="ref70">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wyrwich</surname> <given-names>K. W.</given-names></name> <name><surname>Norquist</surname> <given-names>J. M.</given-names></name> <name><surname>Lenderking</surname> <given-names>W. R.</given-names></name> <name><surname>Acaster</surname> <given-names>S.</given-names></name><collab id="coll1">Industry advisory committee of international society for quality of life research (ISOQOL)</collab></person-group> (<year>2013</year>). <article-title>Methods for interpreting change over time in patient-reported outcome measures</article-title>. <source>Qual. Life Res</source> <volume>22</volume>, <fpage>475</fpage>&#x2013;<lpage>483</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11136-012-0175-x</pub-id>, PMID: <pub-id pub-id-type="pmid">22528240</pub-id></citation>
</ref>
<ref id="ref71">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wyrwich</surname> <given-names>K. W.</given-names></name> <name><surname>Tierney</surname> <given-names>W. M.</given-names></name> <name><surname>Wolinsky</surname> <given-names>F. D.</given-names></name></person-group> (<year>1999</year>). <article-title>Further evidence supporting an SEM-based criterion for identifying meaningful intra-individual changes in health-related quality of life</article-title>. <source>J. Clin. Epidemiol.</source> <volume>52</volume>, <fpage>861</fpage>&#x2013;<lpage>873</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0895-4356(99)00071-2</pub-id>, PMID: <pub-id pub-id-type="pmid">10529027</pub-id></citation>
</ref>
<ref id="ref72">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zinbarg</surname> <given-names>R. E.</given-names></name> <name><surname>Revelle</surname> <given-names>W.</given-names></name> <name><surname>Yovel</surname> <given-names>I.</given-names></name> <name><surname>Li</surname> <given-names>W.</given-names></name></person-group> (<year>2005</year>). <article-title>Cronbach&#x2019;s &#x03B1;, Revelle&#x2019;s &#x03B2;, and Mcdonald&#x2019;s &#x03C9;H: their relations with each other and two alternative conceptualizations of reliability</article-title>. <source>Psychometrika</source> <volume>70</volume>, <fpage>123</fpage>&#x2013;<lpage>133</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11336-003-0974-7</pub-id></citation>
</ref>
</ref-list>
<glossary>
<def-list>
<title>Abbreviations</title>
<def-item>
<term>RCI</term>
<def>
<p>The reliable change index</p>
</def>
</def-item>
<def-item>
<term>EN</term>
<def>
<p>The Edwards-Nunnally method</p>
</def>
</def-item>
<def-item>
<term>GLN</term>
<def>
<p>The Gulliksen-Lord-Novick method</p>
</def>
</def-item>
<def-item>
<term>HLM</term>
<def>
<p>The hierarchical linear modeling</p>
</def>
</def-item>
<def-item>
<term>HA</term>
<def>
<p>The Hageman-Arrindell method</p>
</def>
</def-item>
<def-item>
<term>SEM</term>
<def>
<p>Standard error of measurement</p>
</def>
</def-item>
</def-list>
</glossary>
</back>
</article>